Explore our production ready AI Models

Browse the full catalog of models across text, image, audio, video, 3D, and more.

Model catalog

AI models on one OpenAI-compatible API.

Browse text, image, video, audio, 3D, search, and agent endpoints with pay-as-you-go pricing. The interactive catalog loads current availability from EmpirioLabs, and these model docs are crawlable without client JavaScript.

Open model docs

New & Featured

Save up to 84%

Qwen3.8 27B

Alibaba Cloud
Native InferenceNew

Multimodal reasoner with 256K context, image and video input, function tools, structured JSON, and thinking on by default.

Released Aug 14, 2026256K context
Save up to 10%
Native InferenceNew

Generates complete stereo songs from lyrics and a style prompt, with duration, seed, and format controls up to 5 minutes.

Released Aug 13, 2026
Proprietary EndpointNew

Official 0813 Pro release with major gains across coding, repository work, tool use, and agent tasks, plus hybrid thinking and a 1M context window.

SingaporeReleased Aug 13, 20261M context
Proprietary EndpointNew

Text-to-image plus multi-reference editing with up to three source images, selectable quality tiers, and 1K or 2K output resolution.

Released Aug 7, 2026
Save up to 47%
Native InferenceNew

Meta open 30B agentic model with image understanding, 128K context, tool calling, structured output, and controllable reasoning strength.

Released Aug 10, 2026128K context

Seedance 2.5

ByteDance
Proprietary EndpointNew

Long-form video model for coherent clips up to 30 seconds, with up to 50 reference images, videos, and audio clips, native audio, editing, and extension.

MalaysiaReleased Aug 7, 2026

Text Generation68

Save up to 21%
Proprietary Endpoint

Reasoning and coding model with a 1M token context, 128K output, adjustable reasoning effort, native web search, and tool calling.

SingaporeReleased Jun 16, 20261M context
GermanyReleased Jun 16, 20261M context

Kimi K3

Moonshot AI
Proprietary EndpointNew

Kimi K3 is Moonshot's flagship reasoning model with a 1M token context, always-on thinking, native web search, and text, image, and video inputs.

InternationalReleased Jul 15, 20261M context
Proprietary EndpointNew

Meta's updated frontier reasoning model with a 1,048,576-token context, image, video, audio, and PDF understanding, web search, and tool calling.

Released Aug 5, 20261M context
Save up to 7%

Kimi K2.7 Code

Moonshot AI
Proprietary Endpoint

Kimi K2.7 Code is Moonshot's trillion-parameter agentic coding model with 256K context, always-on reasoning, and text, image, and video inputs.

InternationalReleased Jun 16, 2026256K context
GermanyReleased Jun 16, 2026256K context
Proprietary EndpointNew

Meta frontier reasoning model with a 1,048,576-token context plus image, video, audio, and PDF understanding, web search, and tool calling.

Released Jul 9, 20261M context
Proprietary EndpointNew

Updated multi-agent conductor for hard reasoning, coding, and research, with distinct max effort, 1M context, image input, and web search.

Released Jul 23, 20261M context

Image Generation15

Save up to 39%

FLUX.2 Klein 4B

Black Forest Labs
Native Inference

Apache-licensed 4B FLUX.2 Klein image generation and editing model with text-to-image, reference-image editing, and creative workflow support.

Released Jan 15, 2026
Proprietary Endpoint

Image generation and editing model creating and modifying images from text or image inputs, with inpainting, virtual try-on, and style controls.

Released Dec 3, 2024
Proprietary Endpoint

Open-source text-to-image model on a multimodal Mixture-of-Experts architecture with photorealistic detail and strong multilingual text rendering.

Released Sep 28, 2025
Proprietary Endpoint

Autoregressive framework on the Janus Pro 7B model that unifies multimodal understanding and image generation in one architecture.

Released Jan 27, 2025

Qwen Image 2.0

Alibaba Cloud
Proprietary Endpoint

Unified image generation and editing model with class-leading complex Chinese/English text rendering, realistic textures, and multi-image fusion.

SingaporeReleased Mar 3, 2026

Qwen Image 3.0

Alibaba Cloud
Proprietary EndpointNew

Qwen image generation and editing with Base and Pro variants, multilingual typography, multi-image references, and controllable 1K or 2K output.

SingaporeReleased Jul 21, 2026

Video Generation28

Proprietary Endpoint

Text-to-video and image-to-video with synchronized native audio, at 720p or 1080p for 3 to 15 seconds, with aspect ratio and prompt control.

Released Jun 17, 2026
Proprietary Endpoint

Video generation model producing up to 2-minute multi-shot videos from text and optional image prompts with improved quality and consistency.

Released Apr 7, 2025

HappyHorse 1.0

Alibaba Cloud
Proprietary Endpoint

Video model offering Text-to-Video, Image-to-Video, Reference-to-Video, and Video Edit modes with high-fidelity, motion-smooth output.

SingaporeReleased May 6, 2026
Save up to 19%
Native Inference

8.3B-parameter video model with native 720p output (upscalable to 1080p), strong motion coherence, and bilingual prompt understanding up to 10s.

Released Nov 20, 2025

Kling O3

Kling AI
Proprietary Endpoint

Video model in Standard or Pro modes with Text-to-Video, Image-to-Video, Reference-to-Video, editing, native sound, and multi-scene transitions.

Released Feb 5, 2026
Proprietary Endpoint

Kling 3.0 model that transfers motion from a reference video onto a character from a reference image, with Standard 720p and Pro 1080p tiers.

Audio Generation15

Save up to 17%
Native Inference

Open-source music generation model for text-to-song and lyric-guided audio, with fast 8-step XL Turbo inference for controllable song iteration.

Released Apr 2, 2026
Save up to 12%

TTS 2

Inworld
Proprietary EndpointNew

Realtime voice model with plain-English voice direction, one voice identity across 100+ languages, and sub-200ms streaming time-to-first-audio.

Released May 5, 2026
Save up to 30%
Proprietary Endpoint

Sub-130ms TTFB voice synthesis with 271+ voices across 15 languages, expressive prosody, and real-time SSE streaming for low-latency voice agents.

Released Jan 21, 2026
Save up to 15%
Proprietary Endpoint

Broadcast-quality voice synthesis with rich expressive prosody, 271+ voices across 15 languages, and real-time SSE streaming with per-word timestamps.

Released Jan 21, 2026
Save up to 10%
Native InferenceNew

Generates complete stereo songs from lyrics and a style prompt, with duration, seed, and format controls up to 5 minutes.

Released Aug 13, 2026
Proprietary Endpoint

Low-latency text-to-speech with single- and multi-speaker voices and controllable style, accent, and expressive tone for production apps.

Released May 20, 2025

Transcription4

Proprietary Endpoint

Speech-to-text transcription using the Nova-3 model with multi-language support and advanced customizable settings for production workloads.

Released Feb 12, 2025
Proprietary Endpoint

Whisper-1 speech-to-text transcription trained on multilingual supervised audio, with a 25 MB upload limit per file.

Released Sep 21, 2022
Save up to 17%
Native Inference

Controlled Whisper Large v3 Turbo transcription with multilingual ASR, translation, VAD, timestamps, subtitles, hotwords, and decoder controls.

Released Oct 1, 2024
Proprietary EndpointNew

StepFun streaming speech recognition model for Chinese and English audio transcription.

InternationalReleased Apr 24, 2026

Research & Search13

Proprietary Endpoint

Quick LLM-style answer to a natural-language question, grounded in fresh Exa web search results with inline citations and source links.

Proprietary Endpoint

AI-powered web search with detailed overviews and answers, faster than Deep Search. Ranks #1 on OpenAI SimpleQA benchmark.

100K context
Proprietary Endpoint

Institutional-grade research powered by Claude Opus 4.6 reasoning, with maximum depth, enhanced tool access, and extensive source coverage.

Proprietary Endpoint

Research model for multi-step retrieval, synthesis, and reasoning, autonomously searching, reading, and evaluating sources across complex topics.

128K context

3D Generation1

Save up to 90%

TRELLIS.2 4B

Microsoft
Native Inference

TRELLIS.2 image-to-3D model that turns a reference image into a textured GLB asset with resolution, seed, mesh, texture, and export controls.

Embeddings3

Text Embedding v4

Alibaba Cloud
Proprietary Endpoint

Multilingual text embedding with selectable output dimensions (64–2048). Up to 8,192 tokens per input.

SingaporeReleased Jun 4, 20258K context
Proprietary Endpoint

Speed-optimised multimodal embedding, same shape as Vision-Plus, 3× cheaper image/video tokens.

SingaporeReleased Sep 23, 20251K context
Proprietary Endpoint

Multimodal embedding producing independent vectors for text, image, and video inputs.

SingaporeReleased Sep 23, 20251K context

Rerankers1

Qwen3 Rerank

Alibaba Cloud
Proprietary Endpoint

Semantic document reranker. Sorts up to 500 candidates per query by relevance, supports 100+ languages, and accepts a custom sorting instruction.

SingaporeReleased Jun 5, 20254K context

Tools & Agents2

GPTZero

GPTZero
Proprietary Endpoint

Deep-learning detector that flags portions of text likely generated by AI versus human, classifying content as entirely human, AI, or mixed.

Manus

Manus
Proprietary Endpoint

Autonomous AI agent that turns a high-level prompt into subtasks, calls tools and APIs, and delivers end-to-end results without manual orchestration.

No items found.
No items found.
No items found.
No items found.
No items found.
No items found.
No items found.
No items found.
No items found.
No items found.
No items found.
No items found.
No items found.

Ready to use better endpoints?

Check out our pricing or reach out if you want your own model deployed on our stack.