使用适应性工具(搜索,内存,代码解释器)和测试时间缩放以提升复杂任务的精度的理性模型.
也称为 Alibaba Cloud Qwen3 Max Thinking, Qwen3-Max-Thinking
qwen3-max-thinking/v1/chat/completionsPOST/v1/responsesPOST/v1/messagesPOST/v1beta/models/qwen3-max-thinking:generateContent来自 EmpirioLabs 目录的实时按量计费价格。只为实际用量付费,没有月度最低消费。
Qwen3 Max Thinking 提供 OpenAI 兼容的 Chat Completions API。用你的 EmpirioLabs API 密钥把任意 OpenAI SDK 指向 https://api.empiriolabs.ai/v1,并使用模型 ID qwen3-max-thinking。 在EmpirioLabs 控制台获取 API 密钥。
curl https://api.empiriolabs.ai/v1/chat/completions \
-H "Authorization: Bearer $EMPIRIOLABS_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "qwen3-max-thinking",
"messages": [
{"role": "user", "content": "Write a haiku about the ocean."}
]
}'from openai import OpenAI
client = OpenAI(
base_url="https://api.empiriolabs.ai/v1",
api_key="YOUR_EMPIRIOLABS_API_KEY",
)
response = client.chat.completions.create(
model="qwen3-max-thinking",
messages=[{"role": "user", "content": "Write a haiku about the ocean."}],
)
print(response.choices[0].message.content)EmpirioLabs 上 Qwen3 Max Thinking API 支持的请求参数。省略字段时使用默认值。
| 参数 | 类型 | 默认值 | 范围 / 值 | 描述 |
|---|---|---|---|---|
| temperature | number | 0.7 | 0 到 2 | 采样温度。0 = 确定性,2 = 最大随机性。 |
| top_p | number | 0.9 | 0 到 1 | 核抽样概率质量。低 = 更专注。 |
| max_tokens | number | 4096 | 1 到 65536 | 回复中最多的代币。 |
| stop | string | - | - | 最多有4串字符串,模型会停止生成更多代币。 |
| enable_thinking | boolean | true | - | 启用扩展思考模式。虽然节奏较慢,但能提升推理性任务。 |
| tool_web_search | boolean | false | - | 允许模型在需要时进行网页搜索。 |
| web_search_mode | enum | standard | standard, thorough | 标准 = 单次搜索,彻底搜索 = 多次深入搜索。 |
| tool_code_interpreter | boolean | true | - | 允许模型在沙盒中执行Python代码来计算和分析数据。 |
| tool_web_extractor | boolean | true | - | 允许模型从发现的URL中获取和读取内容。 |
| response_format | enum | - | - | 返回输出为有效的 JSON 对象(JSON 模式)。描述你在提示词中想要的字段。 |
| disable_formatting | boolean | false | - | 跳过EmpirioLabs Markdown 格式(引用 [[N]](url)重写 + 使用网页搜索/工具时的引用块)。返回原始上游答案,带有纯[N]次引用。 |
网络搜索模式:标准(高效)或Thorough(综合,需要思考).
When this model invokes tools (web search, code interpreter, etc.) inside a single request, the response carries a normalized usage.tool_usage map alongside the token counts. The example below shows the shape — exact field names, units, and which tools appear can vary slightly per provider:
"usage": {
"prompt_tokens": 123,
"completion_tokens": 456,
"cost_usd": 0.0042,
"tool_usage": {"web_search": 3, "code_interpreter": 1}
}The tool counts are already factored into cost_usd — they are surfaced for transparency so you can audit per-tool billing. The field is omitted when no tools were invoked.
在 EmpirioLabs,Qwen3 Max Thinking 按量计费。本页的实时价格表始终与 API 的实际计费一致。
Qwen3 Max Thinking 支持 256K token 的上下文窗口,每次响应最多 65,536 个输出 token。
兼容。Qwen3 Max Thinking 提供 OpenAI 兼容的 Chat Completions API,现有 OpenAI SDK 只需把 base_url 指向 https://api.empiriolabs.ai/v1 并把模型 ID 设为 qwen3-max-thinking 即可使用。
可以。EmpirioLabs Playground在浏览器中以与 API 相同的参数运行 Qwen3 Max Thinking,你可以在写代码之前先测试提示词。
创建 EmpirioLabs 账户,然后在控制台的 API Keys生成密钥。计费使用按量付费的额度,只为实际发出的请求付费。