Qwen3.8 27B API

Multimodal reasoner with 256K context, image and video input, function tools, structured JSON, and thinking on by default.

Alibaba CloudGénération de texte256K contextePublié 14 août 2026Inférence nativeNouveau

À propos de Qwen3.8 27B

Multimodal reasoner with 256K context, image and video input, function tools, structured JSON, and thinking on by default.

Supports text, image, and video input, streaming, function tools, structured JSON output, seed control, and thinking mode on by default. Use reasoning_effort (xhigh, medium, low) or enable_thinking=false for direct answers. Automatic cache reads are billed at the cached-input rate when reported by the model service. Explicit cache controls are not supported. Open-weight context is 256K; hosted 1M context and official built-in tools are not included.

Aussi connu sous le nom Alibaba Cloud Qwen3.8 27B

reasoningvisionvideofunction callingcachemultimodaljson modelogprobs

Caractéristiques de Qwen3.8 27B

ID du modèle
qwen3-8-27b
Fournisseur
Alibaba Cloud
Catégorie
Génération de texte
Publié
14 août 2026
Fenêtre de contexte
256K tokens
Sortie max
32 768 tokens
Entrée
TexteImageVidéo
Sortie
Texte
Sortie structurée
JSON Schema
Endpoints
POST/v1/chat/completionsPOST/v1/responsesPOST/v1/messagesPOST/v1/completionsPOST/v1beta/models/qwen3-8-27b:generateContent
Identifiants de modèle alternatifs
qwen3.8-27bqwen/qwen3-8-27bqwen/qwen3.8-27b

Tarifs de l'API Qwen3.8 27BÉconomisez jusqu'à 84%

Tarifs à l'usage en direct du catalogue EmpirioLabs. Vous ne payez que ce que vous utilisez, sans minimum mensuel.

Type
Spéc.
Tarif
Entrée
per 1M prompt tokens
$0.45$0.17
Sortie
per 1M generated tokens
$3.20$0.50
Implicit cache read
per 1M cached input tokens
$0.08
Web Search (Linkup)
per call when invoked
$0.013
Comparer sur la page complète des tarifs

Comment appeler l'API Qwen3.8 27B

Qwen3.8 27B sert l'API Chat Completions compatible OpenAI. Pointez n'importe quel SDK OpenAI vers https://api.empiriolabs.ai/v1 avec votre clé API EmpirioLabs et utilisez l'id de modèle qwen3-8-27b. Obtenez une clé API depuis le tableau de bord EmpirioLabs.

cURL
curl https://api.empiriolabs.ai/v1/chat/completions \
  -H "Authorization: Bearer $EMPIRIOLABS_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen3-8-27b",
    "messages": [
      {"role": "user", "content": "Write a haiku about the ocean."}
    ]
  }'
Python (OpenAI SDK)
from openai import OpenAI

client = OpenAI(
    base_url="https://api.empiriolabs.ai/v1",
    api_key="YOUR_EMPIRIOLABS_API_KEY",
)

response = client.chat.completions.create(
    model="qwen3-8-27b",
    messages=[{"role": "user", "content": "Write a haiku about the ocean."}],
)
print(response.choices[0].message.content)
Référence complète de l'API Qwen3.8 27B

Paramètres de l'API Qwen3.8 27B

Paramètres de requête pris en charge par l'API Qwen3.8 27B sur EmpirioLabs. Les valeurs par défaut s'appliquent quand un champ est omis.

ParamètreTypePar défautPlage / valeursDescription
temperaturenumber10 à 2Sampling temperature. 0 is deterministic and 2 is maximum randomness.
top_pnumber0.950 à 1Nucleus sampling probability mass. Lower values make outputs more focused.
max_tokensinteger40961 à 32768Maximum output tokens.
stopstring--Up to 4 strings where the model will stop generating further tokens.
enable_thinkingbooleantrue-Enable reasoning before answering.
reasoning_effortenumxhighnone, low, medium, xhighReasoning effort level. none disables thinking. low, medium, and xhigh select discrete thinking depths. There is no token budget control.
top_kinteger201 à 200Limit sampling to the top K candidate tokens when supported.
min_pnumber00 à 1Minimum probability threshold for token sampling.
frequency_penaltynumber0-2 à 2Penalty based on how often a token has already appeared.
presence_penaltynumber0-2 à 2Penalty for tokens that already appeared in the generated text.
repetition_penaltynumber10.1 à 2Penalty used by SGLang to reduce repeated text.
seedinteger-0 à 2147483647Optional random seed for reproducible sampling.
logprobsbooleanfalse-Return token log probabilities when supported.
top_logprobsinteger-0 à 20Return up to this many top token log probabilities.
7 paramètres de plus dans la doc

Bon à savoir

Supports text, image, and video input, streaming, function tools, structured JSON output, seed control, and thinking mode on by default. Use reasoning_effort (xhigh, medium, low) or enable_thinking=false for direct answers. Automatic cache reads are billed at the cached-input rate when reported by the model service. Explicit cache controls are not supported. Open-weight context is 256K; hosted 1M context and official built-in tools are not included.

API Qwen3.8 27B : questions fréquentes

Combien coûte l'API Qwen3.8 27B ?

Sur EmpirioLabs, Qwen3.8 27B est facturé à l'usage. La grille tarifaire en direct de cette page correspond toujours à ce que l'API facture.

Quelle est la fenêtre de contexte de Qwen3.8 27B ?

Qwen3.8 27B prend en charge une fenêtre de contexte de 256K tokens avec jusqu'à 32 768 tokens de sortie par réponse.

L'API Qwen3.8 27B est-elle compatible OpenAI ?

Oui. Qwen3.8 27B sert l'API Chat Completions compatible OpenAI : les SDKs OpenAI existants fonctionnent en pointant base_url vers https://api.empiriolabs.ai/v1 et en utilisant l'id de modèle qwen3-8-27b.

Puis-je essayer Qwen3.8 27B dans le navigateur avant d'intégrer ?

Oui. Le playground EmpirioLabs exécute Qwen3.8 27B dans le navigateur avec les mêmes paramètres que l'API expose, pour tester vos prompts avant d'écrire du code.

Comment obtenir une clé API Qwen3.8 27B ?

Créez un compte EmpirioLabs, puis générez une clé sous API Keys dans le tableau de bord. La facturation utilise des crédits à l'usage : vous ne payez que vos requêtes.

Prêt à utiliser de meilleurs points de terminaison ?

Consultez nos tarifs ou contactez-nous si vous souhaitez que votre propre modèle soit déployé sur notre stack.