
Multimodal reasoner with 256K context, image and video input, function tools, structured JSON, and thinking on by default.
Multimodal reasoner with 256K context, image and video input, function tools, structured JSON, and thinking on by default.
Supports text, image, and video input, streaming, function tools, structured JSON output, seed control, and thinking mode on by default. Use reasoning_effort (xhigh, medium, low) or enable_thinking=false for direct answers. Automatic cache reads are billed at the cached-input rate when reported by the model service. Explicit cache controls are not supported. Open-weight context is 256K; hosted 1M context and official built-in tools are not included.
Também conhecido como Alibaba Cloud Qwen3.8 27B
qwen3-8-27b/v1/chat/completionsPOST/v1/responsesPOST/v1/messagesPOST/v1/completionsPOST/v1beta/models/qwen3-8-27b:generateContentqwen3.8-27bqwen/qwen3-8-27bqwen/qwen3.8-27bTarifas pay-as-you-go ao vivo do catálogo EmpirioLabs. Você paga só pelo que usa, sem mínimo mensal.
Qwen3.8 27B atende a API Chat Completions compatível com OpenAI. Aponte qualquer SDK OpenAI para https://api.empiriolabs.ai/v1 com sua chave de API EmpirioLabs e use o id de modelo qwen3-8-27b. Obtenha uma chave de API no painel EmpirioLabs.
curl https://api.empiriolabs.ai/v1/chat/completions \
-H "Authorization: Bearer $EMPIRIOLABS_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "qwen3-8-27b",
"messages": [
{"role": "user", "content": "Write a haiku about the ocean."}
]
}'from openai import OpenAI
client = OpenAI(
base_url="https://api.empiriolabs.ai/v1",
api_key="YOUR_EMPIRIOLABS_API_KEY",
)
response = client.chat.completions.create(
model="qwen3-8-27b",
messages=[{"role": "user", "content": "Write a haiku about the ocean."}],
)
print(response.choices[0].message.content)Parâmetros de requisição suportados pela API Qwen3.8 27B na EmpirioLabs. Os padrões valem quando um campo é omitido.
| Parâmetro | Tipo | Padrão | Intervalo / valores | Descrição |
|---|---|---|---|---|
| temperature | number | 1 | 0 a 2 | Sampling temperature. 0 is deterministic and 2 is maximum randomness. |
| top_p | number | 0.95 | 0 a 1 | Nucleus sampling probability mass. Lower values make outputs more focused. |
| max_tokens | integer | 4096 | 1 a 32768 | Maximum output tokens. |
| stop | string | - | - | Up to 4 strings where the model will stop generating further tokens. |
| enable_thinking | boolean | true | - | Enable reasoning before answering. |
| reasoning_effort | enum | xhigh | none, low, medium, xhigh | Reasoning effort level. none disables thinking. low, medium, and xhigh select discrete thinking depths. There is no token budget control. |
| top_k | integer | 20 | 1 a 200 | Limit sampling to the top K candidate tokens when supported. |
| min_p | number | 0 | 0 a 1 | Minimum probability threshold for token sampling. |
| frequency_penalty | number | 0 | -2 a 2 | Penalty based on how often a token has already appeared. |
| presence_penalty | number | 0 | -2 a 2 | Penalty for tokens that already appeared in the generated text. |
| repetition_penalty | number | 1 | 0.1 a 2 | Penalty used by SGLang to reduce repeated text. |
| seed | integer | - | 0 a 2147483647 | Optional random seed for reproducible sampling. |
| logprobs | boolean | false | - | Return token log probabilities when supported. |
| top_logprobs | integer | - | 0 a 20 | Return up to this many top token log probabilities. |
Supports text, image, and video input, streaming, function tools, structured JSON output, seed control, and thinking mode on by default. Use reasoning_effort (xhigh, medium, low) or enable_thinking=false for direct answers. Automatic cache reads are billed at the cached-input rate when reported by the model service. Explicit cache controls are not supported. Open-weight context is 256K; hosted 1M context and official built-in tools are not included.
Na EmpirioLabs, Qwen3.8 27B é cobrado por uso. A tabela de tarifas ao vivo desta página sempre corresponde ao que a API cobra.
Qwen3.8 27B suporta uma janela de contexto de 256K tokens com até 32.768 tokens de saída por resposta.
Sim. Qwen3.8 27B atende a API Chat Completions compatível com OpenAI, então SDKs OpenAI existentes funcionam apontando base_url para https://api.empiriolabs.ai/v1 e definindo o id de modelo qwen3-8-27b.
Sim. O playground da EmpirioLabs executa Qwen3.8 27B no navegador com os mesmos parâmetros que a API expõe, para você testar prompts antes de escrever código.
Crie uma conta EmpirioLabs e gere uma chave em API Keys no painel. A cobrança usa créditos pay-as-you-go, então você paga apenas pelas requisições que faz.
Confira nossos preços ou entre em contato se quiser que seu próprio modelo seja implementado em nossa pilha.