
Multimodal reasoner with 256K context, image and video input, function tools, structured JSON, and thinking on by default.
Multimodal reasoner with 256K context, image and video input, function tools, structured JSON, and thinking on by default.
Supports text, image, and video input, streaming, function tools, structured JSON output, seed control, and thinking mode on by default. Use reasoning_effort (xhigh, medium, low) or enable_thinking=false for direct answers. Automatic cache reads are billed at the cached-input rate when reported by the model service. Explicit cache controls are not supported. Open-weight context is 256K; hosted 1M context and official built-in tools are not included.
También conocido como Alibaba Cloud Qwen3.8 27B
qwen3-8-27b/v1/chat/completionsPOST/v1/responsesPOST/v1/messagesPOST/v1/completionsPOST/v1beta/models/qwen3-8-27b:generateContentqwen3.8-27bqwen/qwen3-8-27bqwen/qwen3.8-27bTarifas de pago por uso en vivo del catálogo de EmpirioLabs. Solo pagas por lo que usas, sin mínimo mensual.
Qwen3.8 27B sirve la API de Chat Completions compatible con OpenAI. Apunta cualquier SDK de OpenAI a https://api.empiriolabs.ai/v1 con tu clave de API de EmpirioLabs y usa el id de modelo qwen3-8-27b. Consigue una clave de API en el panel de EmpirioLabs.
curl https://api.empiriolabs.ai/v1/chat/completions \
-H "Authorization: Bearer $EMPIRIOLABS_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "qwen3-8-27b",
"messages": [
{"role": "user", "content": "Write a haiku about the ocean."}
]
}'from openai import OpenAI
client = OpenAI(
base_url="https://api.empiriolabs.ai/v1",
api_key="YOUR_EMPIRIOLABS_API_KEY",
)
response = client.chat.completions.create(
model="qwen3-8-27b",
messages=[{"role": "user", "content": "Write a haiku about the ocean."}],
)
print(response.choices[0].message.content)Parámetros de solicitud compatibles con la API de Qwen3.8 27B en EmpirioLabs. Los valores por defecto se aplican cuando se omite un campo.
| Parámetro | Tipo | Predeterminado | Rango / valores | Descripción |
|---|---|---|---|---|
| temperature | number | 1 | 0 a 2 | Sampling temperature. 0 is deterministic and 2 is maximum randomness. |
| top_p | number | 0.95 | 0 a 1 | Nucleus sampling probability mass. Lower values make outputs more focused. |
| max_tokens | integer | 4096 | 1 a 32768 | Maximum output tokens. |
| stop | string | - | - | Up to 4 strings where the model will stop generating further tokens. |
| enable_thinking | boolean | true | - | Enable reasoning before answering. |
| reasoning_effort | enum | xhigh | none, low, medium, xhigh | Reasoning effort level. none disables thinking. low, medium, and xhigh select discrete thinking depths. There is no token budget control. |
| top_k | integer | 20 | 1 a 200 | Limit sampling to the top K candidate tokens when supported. |
| min_p | number | 0 | 0 a 1 | Minimum probability threshold for token sampling. |
| frequency_penalty | number | 0 | -2 a 2 | Penalty based on how often a token has already appeared. |
| presence_penalty | number | 0 | -2 a 2 | Penalty for tokens that already appeared in the generated text. |
| repetition_penalty | number | 1 | 0.1 a 2 | Penalty used by SGLang to reduce repeated text. |
| seed | integer | - | 0 a 2147483647 | Optional random seed for reproducible sampling. |
| logprobs | boolean | false | - | Return token log probabilities when supported. |
| top_logprobs | integer | - | 0 a 20 | Return up to this many top token log probabilities. |
Supports text, image, and video input, streaming, function tools, structured JSON output, seed control, and thinking mode on by default. Use reasoning_effort (xhigh, medium, low) or enable_thinking=false for direct answers. Automatic cache reads are billed at the cached-input rate when reported by the model service. Explicit cache controls are not supported. Open-weight context is 256K; hosted 1M context and official built-in tools are not included.
En EmpirioLabs, Qwen3.8 27B se factura por uso. La tabla de tarifas en vivo de esta página siempre coincide con lo que cobra la API.
Qwen3.8 27B admite una ventana de contexto de 256K tokens con hasta 32.768 tokens de salida por respuesta.
Sí. Qwen3.8 27B sirve la API de Chat Completions compatible con OpenAI, así que los SDKs de OpenAI existentes funcionan apuntando base_url a https://api.empiriolabs.ai/v1 y usando el id de modelo qwen3-8-27b.
Sí. El playground de EmpirioLabs ejecuta Qwen3.8 27B en el navegador con los mismos parámetros que expone la API, para que pruebes prompts antes de escribir código.
Crea una cuenta de EmpirioLabs y genera una clave en API Keys en el panel. La facturación es con créditos de pago por uso, así que solo pagas por las solicitudes que haces.
Consulta nuestros precios o contacta con nosotros si quieres que tu propio modelo se implemente en nuestra pila.