Models

AIHubMix's in-house automatic model router. Not every request needs the strongest, most expensive model — drawing on internal evals and public leaderboards, we bring together a range of today's mainstream models of varying capability and route each request automatically through a small model we trained. Set model to auto and requests are dispatched to the right model based on their content, lowering cost. Currently in beta; we keep updating and improving it.

GLM-5.3-Flash is a high-efficiency multimodal model from Z.AI. It supports a context window of roughly 1 million tokens, along with text, image, and video inputs, and includes tool-calling capabilities. It is primarily designed for coding agents, complex reasoning, and long-horizon software engineering tasks. Built on the existing GLM technology stack, the model has been further post-trained and optimized to deliver strong performance while placing greater emphasis on inference efficiency, responsiveness, and cost.The model is offered at a limited-time 50% discount; users are welcome to try it.

GLM-5.3 is Z.AI’s coding and agentic reasoning model, built for complex software engineering, long-running agent tasks, vulnerability analysis, and other demanding workloads. Building on GLM-5.2, it incorporates further post-training improvements to deliver stronger coding performance, better task execution, and greater token efficiency. We currently offer the production-ready GLM-5.3 API with unlimited concurrency, making it well suited for high-throughput workloads, coding agents, and large-scale automation. For a limited time, GLM-5.3 is available at 10% off.

coding-glm-5.3-flash-free is the open and free version of coding-glm-5.3-flash. To ensure stable service performance, usage limits are in place: up to 5 requests per minute, 500 requests per day, and a daily token allowance of 1 million.
Hy4 Preview is Tencent Hunyuan’s large language model for agents, coding, office automation, and complex tool use.Hy4 preview has a total of 770B parameters and 49B active parameters, and is primarily optimized for agent, coding, and productivity scenarios. It strengthens understanding, planning, tool invocation, and sustained execution capabilities for complex tasks. Compared with the previous generation, Hy4 preview further improves multi-step agents, code development, and productivity tasks, offering better task decomposition, context continuity, instruction following, and long-horizon execution. In coding scenarios, it further enhances code understanding, generation, modification, and the handling of complex engineering tasks; in productivity scenarios, it focuses on improving document processing, information analysis, office automation, game development, webpage generation, and cross-tool collaboration. Hy4 preview is suitable for coding agents, complex tool invocation, and various agent workflows that require multi-step planning and sustained execution, providing more reliable task completion for complex real-world business scenarios.

Coding GLM 5.3 Flash is a dedicated version of GLM 5.3 Flash built for AI coding and Coding Agent workflows. It is designed for code understanding, generation, editing, repository-level development, and automated software engineering tasks. With support for context windows of up to approximately 1 million tokens, it can handle large codebases and extended development sessions. The model also supports text, image, and video inputs, along with tool use, making it well suited for AI coding tools such as Claude Code, OpenCode, Cline, and other agentic development environments.

coding-glm-5.3-free is the open and free version of coding-glm-5.3. To ensure stable service performance, usage limits are in place: up to 5 requests per minute, 500 requests per day, and a daily token allowance of 1 million.
The Hy3 official version is honed for real-world business scenarios, using a Mixture-of-Experts (MoE) architecture with 295B total parameters and 21B activated parameters. It natively supports a 256K context window and offers multiple thinking modes: no_think (ultra-fast response), think_low (quick thinking), and think_high (deep reasoning), balancing ultra-fast responses, complex reasoning, and invocation cost. Compared with the Preview version, Hy3—based on real business feedback from Tencent Yuanbao, WorkBuddy, ima, Marvis, and others—focuses on improving the Coding Agent, long-form understanding, multi-turn context continuity, search QA, and complex task execution, performing more stably in reducing hallucinations, improving task completion, and engineering usability. It is better suited to practical scenarios such as frontend tasks, cross-file code development, long-document analysis, office automation, and multi-step Agent workflows.
MiniMax-M3 is a versatile multimodal foundation model developed by MiniMax that supports text, image, and video inputs to generate text outputs. With a massive context window of 1,048,576 tokens, it is capable of processing and understanding vast amounts of information. This model is highly optimized for complex tasks, making it exceptionally well-suited for coding and long-horizon agentic workflows.
Qwen3.8 Flash is Alibaba Cloud Qwen’s flagship native vision-language model for coding, office tasks, long-context reasoning, and agent workflows. It supports a 1M-token context, 128K output, web access, and tool calling. Compared with Qwen3.7-Plus, Qwen3.8-Flash significantly reduces training and inference costs—the training overhead is only about one-ninth of the former—while offering stronger capabilities on coding and office tasks.
- Input: $ 0.75 /M
- Output: $ 3.75 /M
- Web Search: $0.014/request
- Cache Storage: $1/h/M tokens
- Input Audio: $1/M tokens
- Input Video: $1/M tokens
Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web development, and knowledge work. It supports a 1M-token context window and adjustable thinking levels. Compared with Gemini 3.6 Flash, it improves coding, tool use, multi-step planning, and instruction following.
Developed by Minimax, MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Featuring a generous 196,608-token context window, the model integrates advanced agentic capabilities through multi-agent workflows. It is uniquely built to actively participate in its own evolution, delivering highly adaptable and intelligent performance.

This model actually points to glm-5.3-flash; if you need to use it in production, you can directly call the model name "glm-5.3-flash".

Dots3-Note Preview is an open-weight mixture-of-experts model developed by Dots Studio, featuring 16B active parameters out of 280B total. As the lightest model in the Dots 3 family, it is designed for efficient performance while supporting an expansive context length of 512,000 tokens. This preview version provides an accessible way to experience the capabilities of the Dots 3 architecture.
- Input: $ 0 /M
- Output: $ 0 /M
- Web Search: $0.014/request
- Cache Storage: $1/h/M tokens
- Input Audio: $2/M tokens
- Input Video: $1/M tokens
Gemini 3.7 Flash free version: Free model resources are limited and provided only for trial use; stability cannot be guaranteed, and you may encounter 429 errors during use. If you need to use it in a production environment and require unlimited concurrency with absolute stability, please choose the official version: gemini-3.7-flash

GLM-5.2 is Z.ai’s flagship model for the era of long-horizon tasks. With a truly usable 1M-token context window, it can handle project-level engineering context, execute long-running tasks more reliably, follow engineering standards more consistently, and complete the full development workflow from requirements to multi-platform deployment in a single task.
DeepSeek-V4-Flash-0731(deepseek-v4-flash-0731) is an open-source MoE large language model developed by the Chinese AI company DeepSeek, with support for a million-token context window. It is designed for coding, complex reasoning, tool use, agentic workflows, and long-document processing. Its advantages include strong performance with fewer active parameters and improved efficiency through DSpark speculative decoding. Compared with DeepSeek V4-Flash Preview, it offers significantly stronger coding and agent capabilities, while outperforming DeepSeek V4-Pro Preview on several benchmarks with fewer active parameters.
- Input: $ 0.142 /M
- Output: $ 0.284 /M
- Web Search: $0.00056/request
DeepSeek’s officially released new multimodal visual-understanding model, DeepSeek‑V4‑Flash‑Vision‑Exp, is experimental in nature and supports multimodal inputs. In pure-text capabilities (agents, reasoning, world knowledge, etc.), DeepSeek‑V4‑Flash‑Vision‑Exp is on par with the official DeepSeek‑V4‑Flash release.
DeepSeek V4 Pro 0813 is DeepSeek’s high-performance general-purpose reasoning and agent model, designed for complex reasoning, coding, long-document analysis, and agentic workflows. It supports thinking and non-thinking modes, a 1M-token context window, up to 384K output, tool calling, and the Responses API. Compared with V4 Flash 0731, Pro prioritizes capability on complex tasks, while Flash focuses on speed, cost efficiency, and high concurrency.
- Input: $ 2 /M
- Output: $ 6 /M
Grok 4.6 is xAI’s (SpaceXAI) flagship multimodal reasoning model for coding, long-running agents, knowledge work, and interactive application development. It supports image understanding, a 500K context window, tool calling, and structured outputs. Compared with Grok 4.5, it offers stronger multi-step execution, self-verification, coding, and visual project generation.
Popular models
Qwen3.8 Max Preview · Kimi K3 · Qwen3.8 Max · Qwen3.7 Flash · GLM 5.2 · Grok 4.5 · Claude Opus 5 · Claude Sonnet 5 · GPT 5.6 Luna · Gemini 3.6 Flash · DeepSeek V4 Flash · GPT 5.5 · Gemini 3.1 Pro Preview
Browse by model author
OpenAI (133) · Anthropic (28) · Google (83) · Grok (26) · Qwen (143) · DeepSeek (39) · Z.AI (70) · ByteDance (46) · Llama (50) · AI21 (2) · Microsoft (14) · Cohere (19) · Mistral (10) · Moonshot AI (34) · StepFun (4) · Nvidia (16) · Minimax (33) · Ideogram (9) · Jina AI (13) · Stable diffusion (1) · Hunyuan (7) · Baidu (23) · Flux (5) · Meituan (2) · Xiaomi (14) · InclusionAI (7) · BAAI (5) · Agnes (4) · KLing (2) · Meta (2) · Poolside (2) · Dots Studio (1) · Liquid (1)
55 free models — no credit card required →
All models (863)
Auto · GLM 5.3 Flash · GLM 5.3 · Coding GLM 5.3 · Coding GLM 5.3 Flash (free) · Hy4 Preview · Coding GLM 5.3 Flash · Coding GLM 5.3 (free) · Hy3 (free) · MiniMax M3 (free) · Qwen3.8 Flash · Gemini 3.7 Flash · MiniMax M2.7 (free) · Ox Alpha · Dots 3 Note Preview (free) · Gemini 3.7 Flash (free) · GLM 5.2 · DeepSeek V4 Flash 0731 · DeepSeek V4 Flash Vision Exp · DeepSeek V4 Pro 0813 · Grok 4.6 · DeepSeek V4 Flash 0731 Fast · Mai Thinking 1 · Wan3.0 Video · Wan3.0 Video Prime · GPT 5.6 Sol Disc · Doubao Seedance 2.5 260628 · GPT 5.6 Luna · GPT 5.6 Sol · GPT 5.6 Terra · Agnes 2.5 Flash · Agnes 2.5 Pro · Agnes 2.5 Pro Alpha · Agnes Image 2.1 Flash · Grok 4.5 · Lfm 2.5 2.6b (free) · MiniMax H3 · Qwen3.8 2.4t A95B · Claude Opus 5 · Gemini 3.6 Flash · Ling 3.0 Tiny (free) · Nemotron 3.5 Lightning (free) · Qwen Image 3.0 · Qwen Image 3.0 Pro · Qwen3.8 Max · Claude Sonnet 5 · Doubao Deepseek V4 Flash 0731 · Kimi K3 · Ling 3.0 Flash (free) · Muse Spark 1.2 · Qwen3.8 Max Preview · Gemini 3.1 Flash Lite Image · Gemini 3.5 Flash Lite · Gemini 3.5 Flash Lite (free) · Gemini 3.6 Flash (free) · GLM 5.2 Fast Preview · Muse Spark 1.1 · Qwen Audio 3.0 Tts Flash · Qwen Audio 3.0 Tts Plus · Claude Fable 5 · Jina Reranker V3.5 · Claude Opus 4.8 · Hy3 · Doubao Seed 2.1 Pro · Doubao Seed 2.1 Turbo · Mai Image 2.5 Pro · Gemini 3.5 Flash · Grok Build 0.1 · Mai Image 2.5 · Mai Image 2.5 Flash · Coding Kimi K3 · Happyhorse 1.1 I2v · Happyhorse 1.1 R2v · Happyhorse 1.1 T2v · Coding GLM 5.2 (free) · Coding Kimi K3 (free) · Gemini 3.1 Flash Image · GPT Oss 20B (free) · Kimi K2.7 Code · Kimi K2.7 Code Highspeed · Gemini 3 Pro Image · GPT 4o Transcribe Diarize · GPT Audio 1.5 · Hy 3d 3.1 · Kling V3 Omni · Kling Video O1 · Longcat 2.0 · Nemotron Nano 9B V2 (free) · Hy3 Preview · MiniMax M3 · Nemotron Nano 12B V2 VL (free) · Qwen3.7 Flash · Qwen3.7 Plus · Step 3.7 Flash · Claude Opus 4.8 Thinking · Nemotron 3 Super 120B A12B (free) · Nemotron 3 Nano Omni 30B A3B (reasoning) (free) · Nemotron 3 Ultra 550B A55B (free) · Qwen3.7 Max · GPT Image 2 · Nemotron 3.5 Content Safety (free) · Coding GLM 5.2 · ERNIE 5.1 · Gemini 3.1 Flash Lite · gemini-3.1-flash-lite-nothink · Grok 4.3 · Happyhorse 1.0 I2v · Happyhorse 1.0 R2v · Happyhorse 1.0 T2v · Happyhorse 1.0 Video Edit · North Mini Code (free) · GPT 5.5 · GPT 5.5 Pro · Laguna Xs 2.1 (free) · DeepSeek V4 Flash · DeepSeek V4 Pro · Gemma 4 31B It (free) · Command A Plus 05 2026 · Doubao Seedream 5.0 Pro · ERNIE 5.0 · Kimi K2.6 · Laguna S 2.1 (free) · Qwen3.6 Max Preview · Xiaomi Mimo V2.5 · Xiaomi Mimo V2.5 Pro · Claude Opus 4.7 · Claude Opus 4.7 Thinking · GPT Chat · Nemotron 3 Nano 30B A3B (free) · Qwen3.6 27B · Qwen3.6 35B A3B · Qwen3.6 Flash · Cohere Rerank V4.0 Fast · Cohere Rerank V4.0 Pro · Gemma 4 26B A4B It (free) · grok-4-20-non-reasoning · Grok 4 20 (reasoning) · Qwen Image 2.0 · Qwen Image 2.0 Pro · Coding MiniMax M3 (free) · Doubao Seedance 2.0 260128 · Doubao Seedance 2.0 Fast 260128 · Doubao Seedance 2.0 Mini 260615 · GLM 5.1 · GLM Image · Qwen3.6 Plus · Wan2.7 I2v · Wan2.7 R2v · Wan2.7 T2v · Wan2.7 Videoedit · CC K2.6 Code Preview · Gemma 4 26B A4B It · Gemma 4 31B It · GPT 5.4 · Wan2.7 Image · Wan2.7 Image Pro · Claude Sonnet 4.6 · Coding Xiaomi Mimo V2.5 · Coding Xiaomi Mimo V2.5 Pro · Doubao Seed 2.0 Lite 260428 · Doubao Seed 2.0 Mini 260428 · Gemini 3.1 Flash Image Preview · Gemini 3.1 Pro Preview · Gemini 3.1 Pro Preview Customtools · Gemini 3.1 Pro Preview Search · GPT 5.4 Mini · GPT 5.4 Nano · GPT 5.5 (free) · Qwen3.5 Plus · Claude Sonnet 4.6 Thinking · Coding Xiaomi Mimo V2 Omni · Coding Xiaomi Mimo V2 Pro · GPT 5.3 Chat · GPT-5.3-Codex · GPT Image 2 (free) · Qwen3.5 122B A10B · Qwen3.5 27B · Qwen3.5 35B A3B · Qwen3.5 397B A17B · Qwen3.5 Flash · Coding GLM 5.1 · Doubao Seed 2.0 Pro · GPT 5.4 High · GPT 5.4 Low · GPT 5.4 Pro · Qwen3 Coder Next · Xiaomi Mimo V2 Omni (free) · Xiaomi Mimo V2 Pro (free) · Xiaomi Mimo V2.5 (free) · Xiaomi Mimo V2.5 Pro (free) · Claude Opus 4.6 · Coding GLM 5.1 (free) · Coding MiniMax M2.7 (free) · GLM 5 · GLM 5 Vision Turbo · MiniMax M2.7 · Claude Opus 4.6 Thinking · Coding GLM 5 (free) · Coding GLM 5 Turbo (free) · Coding MiniMax M2.5 (free) · Doubao Seed 2.0 Code Preview · Doubao Seed 2.0 Lite 260215 · Doubao Seed 2.0 Mini · Gemini 3 Flash Preview · Gemini 3 Flash Preview Search · GLM 5 Turbo · CC GLM 5.1 · Claude Opus 4.5 · Claude Opus 4.5 Thinking · Embed V 4.0 · ERNIE Image Turbo · Gemini 3.1 Flash Image Preview (free) · MiMo V2 Omni · MiMo V2 Pro · Cohere Command A · Gemini 3 Flash Preview (free) · CC MiniMax M3 · Coding MiniMax M3 · GPT 4.1 (free) · GPT 4.1 Mini (free) · GPT 4.1 Nano (free) · GPT 4o (free) · Coding GLM 5 · Coding GLM 5 Turbo · GLM 4.7 · Veo 3.1 Lite Generate Preview · GLM 4.7 Flash (free) · Coding GLM 4.7 (free) · Doubao Seedance 1.5 Pro 251215 · Doubao Seedance 1.0 Pro 250528 · Doubao Seedance 1.0 Pro Fast 251015 · Gemini 3 Pro Image Preview · Gemini Embedding 2 · Deepinfra Gemma 4 26B A4B It · GPT-5.2-Codex · Doubao Seedream 5.0 Lite · GPT Image 1.5 · Baidu DeepSeek V4 Pro 0813 · GPT 5.2 · GPT 5.2 Chat · GPT 5.2 High · GPT 5.2 Low · GPT 5.2 Pro · GPT 5.1 · GPT-5.1-Codex Max · Doubao Seed 1.8 · GPT 5.1 Chat · GPT-5.1-Codex · GPT-5.1-Codex Mini · Claude Haiku 4.5 · Claude Sonnet 4.5 · Claude Sonnet 4.5 Thinking · Grok 4.20 Multi Agent 0309 · Mistral Large 3 · CC GLM 5 · CC GLM 5 Turbo · Cloudflare Glm 5.2 · Gemini 2.5 Flash Image · grok-4-1-fast-non-reasoning · Grok 4.1 Fast (reasoning) · Grok Code Fast 1 · K2.6 Code Preview (free) · MiMo V2 Flash · Musesteamer Air Image · Qwen3.6 Plus Preview (free) · Zai Glm 5 Turbo · GPT 5 · DeepSeek V3.2 · DeepSeek V3.2 Thinking · GPT-5-Codex · DeepSeek V3.1 Terminus · DeepSeek V3.1 Thinking · GPT 5 Pro · GPT 5 Mini · GPT 5 Nano · GPT 5 Chat · Claude Opus 4.1 · O3 Deep Research · Kimi K2.5 · Qwen3 Max 2026 01-23 · Qwen3 VL Flash · Qwen3 VL Flash 2026 01-22 · Qwen3 VL Plus · CC MiniMax M2.7 · CC MiniMax M2.7 Highspeed · MiniMax M2.5 · MiniMax M2.5 Highspeed · Mm Minimax M2.7 Highspeed · Coding MiniMax M2.7 · Coding MiniMax M2.7 Highspeed · CC MiniMax M2.5 · CC MiniMax M2.5 Highspeed · Coding MiniMax M2.5 · Coding MiniMax M2.5 Highspeed · Doubao Seedream 4.5 · Sora 2 · Sora 2 Pro · CC GLM 4.7 · CC MiniMax M2.1 · Coding GLM 4.7 · Coding MiniMax M2.1 · Coding MiniMax M2.1 (free) · GPT 4o Audio Preview · GPT 4o Mini Audio Preview · MiniMax M2.1 · O3 · Wan2.6 I2v · Wan2.6 T2v · CC GLM 4.6 · Coding GLM 4.6 · Coding GLM 4.6 (free) · Coding MiniMax M2 · Coding MiniMax M2 (free) · Flux 2 Flex · Flux 2 Pro · Gemini 2.5 Pro · GLM 4.6 · GLM 4.6 Vision · GLM Ocr · Kimi For Coding (free) · O3 Pro · Qianfan Ocr · Qianfan Ocr Fast · Step 3.5 Flash · Wan2.2 I2v Plus · Wan2.5 I2v Preview · Wan2.5 T2v Preview · Gemini 2.5 Pro Search · Kimi K2 Thinking · Gemini 2.5 Flash · Gemini 2.5 Flash Preview 09 2025 · GLM 4.5 Vision · Gemini 2.5 Flash Lite · gemini-2.5-flash-lite-nothink · Gemini 2.5 Flash Lite Preview 09 2025 · gemini-2.5-flash-lite-preview-09-2025-nothink · gemini-2.5-flash-nothink · Gemini 2.5 Flash Search · gemini-2.5-flash-preview-05-20-nothink · Gemini 2.5 Flash Preview 05-20 Search · DeepSeek V3 Fast · Imagen 4.0 · Imagen 4.0 Fast Generate 001 · Imagen 4.0 Generate 001 · Imagen 4.0 Ultra Generate 001 · Imagen 4.0 Ultra · GPT Image 1 · GPT Image 1 Mini · O4 Mini · DeepSeek-OCR · Alicloud Kimi K2 Instruct · DeepSeek Ocr · ERNIE 5.0 Thinking Exp · Flux Kontext Max · Gemini 2.5 Flash Image Preview · GLM 4.5 · GPT 4.1 · Grok 4 · grok-4-fast-non-reasoning · Grok 4 Fast (reasoning) · Kimi K2 0711 · Kimi K2 Instruct · Kimi K2 Turbo Preview · Paddleocr VL 0.9b · Pp Structurev3 · Qwen3 VL 235B A22B Instruct · Qwen3 VL 235B A22B Thinking · Qwen3 VL 30B A3B Instruct · Qwen3 VL 30B A3B Thinking · Veo 3.0 Generate Preview · Veo 3.1 Fast Generate Preview · Veo 3.1 Generate Preview · AIHubMix Router · GPT 4.1 Mini · GPT 4.1 Nano · Gemini 2.5 Pro Preview 05-06 · Gemini 2.5 Pro Preview 03-25 · Gemini 2.5 Pro Preview 05-06 Search · Gemini 2.5 Pro Preview 03-25 Search · Qwen3 Max Preview · Qwen3 Max · Qwen3 Next 80B A3B Instruct · Qwen3 Next 80B A3B Thinking · Qwen3 235B A22B Instruct 2507 · Qwen3 235B A22B Thinking 2507 · Qwen3 Coder 30B A3B Instruct · Qwen3 Coder 480B A35B Instruct · DeepSeek V3 · LongCat-Flash-Chat · Gemini 2.5 Pro Preview 06-05 Search · Jina Embeddings V5 Text Nano · Jina Embeddings V5 Text Small · Qwen3 235B A22B · Qwen3 Coder Flash · Qwen3 Coder Plus · Qwen3 Coder Plus 2025 07-22 · ERNIE 5.0 Thinking Preview · inclusionAI/Ling-1T · inclusionAI/Ring-1T · Bce Reranker Base · Codex Mini · Doubao Seedream 4.0 · Embedding V1 · ERNIE 4.5 Turbo · GLM 4.5 X · Gme Qwen2 VL 2B Instruct · Gte Rerank V2 · inclusionAI/Ling-flash-2.0 · inclusionAI/Ling-mini-2.0 · inclusionAI/Ring-flash-2.0 · Jina Deepsearch V1 · Jina Embeddings V4 · Jina Reranker V3 · Llama 4 Maverick · Llama 4 Scout · Qwen Image · Qwen Image Edit · Qwen Image Max · Qwen Mt Plus · Qwen Mt Turbo · Qwen3 Embedding 0.6b · Qwen3 Embedding 4B · Qwen3 Embedding 8B · Qwen3 Reranker 0.6b · Qwen3 Reranker 4B · Qwen3 Reranker 8B · Tao 8K · Jina Clip V2 · Jina Reranker M0 · Jina Colbert V2 · GPT 4o Search Preview · GPT 4o Mini Search Preview · Jina Embeddings V3 · Claude 3.7 Sonnet · ERNIE 4.5 · ERNIE 4.5 Turbo VL · MiMo V2 Flash (free) · FLUX-1.1-pro · O3 Mini · Doubao Seed 1.6 · Doubao Seed 1.6 Flash · Doubao Seed 1.6 Lite · Doubao Seed 1.6 Thinking · Qwen3 30B A3B Instruct 2507 · Qwen3 30B A3B Thinking 2507 · Qwen2 VL 72B Instruct · Qwen2 VL 7B Instruct · CC Kimi For Coding · Gemini Embedding 001 · gpt-oss-120b · Qwen 3 235B A22B Thinking 2507 · Qwen/Qwen3-30B-A3B · Qwen/Qwen3-32B · Qwen3 32B · Qwen/Qwen3-14B · Qwen/Qwen3-8B · Embedding 2 · Embedding 3 · Gemini 2.5 Pro Preview 06-05 · Qwen/Qwen2.5-VL-72B-Instruct · O1 · O1 Pro · ByteDance-Seed/Seed-OSS-36B-Instruct · Doubao Seed 1.6 250615 · Doubao Seed 1.6 Flash 250615 · Doubao Seed 1.6 Thinking 250615 · Doubao Seed 1.6 Vision 250815 · Doubao 1.5 Thinking Pro · CC MiniMax M2 · deepseek-ai/DeepSeek-Prover-V2-671B · Gemini 2.5 Flash Preview Tts · Gemini 2.5 Pro Preview Tts · Gemma 3 12B It · Gemma 3 27B It · Gemma 3 4B It · Gemma 3n E4B It · Gemma 3 1B It · DeepSeek R1 Distill Llama 70B · GPT 4o Mini Tts · tngtech/DeepSeek-R1T-Chimera · Veo 2.0 Generate 001 · O1 Preview · O1 Mini · GPT 4o 2024 11-20 · GPT 4o · GPT 4o Mini · AIHubMix Mistral Medium · ERNIE X1.1 Preview · Qwen/QwQ-32B · chutesai/Mistral-Small-3.1-24B-Instruct-2503 · ERNIE X1.1 Preview · MiniMax M2 · MiniMaxAI/MiniMax-M1-80k · Qwen/Qwen2.5-VL-32B-Instruct · baidu/ERNIE-4.5-300B-A47B · Bge Large En · Bge Large Zh · Codestral · ERNIE 4.5 0.3b · ERNIE 4.5 Turbo 128K Preview · ERNIE X1 Turbo · Kat Dev · Llama 3.3 70B · moonshotai/Kimi-Dev-72B · moonshotai/Moonlight-16B-A3B-Instruct · Nvidia Nemotron 3 Super 120B A12B · O1 Global · Qianfan Qi VL · Qwen2.5 VL 72B Instruct · tencent/Hunyuan-A13B-Instruct · unsloth/gemma-3-27b-it · gemini-exp-1206 · GPT 4o Zh · Qwen Qwq 32B · unsloth/gemma-3-12b-it · Qwen Max 0125 · BAAI/bge-large-en-v1.5 · BAAI/bge-large-zh-v1.5 · BAAI/bge-reranker-v2-m3 · tencent/Hunyuan-MT-7B · V3 · V_2 · V_2_TURBO · V_2A · V_2A_TURBO · V_1 · V_1_TURBO · Doubao Embedding Large Text 240915 · Kimi Thinking Preview · GPT 4o 2024 08-06 · Qwen Plus 2025 07-28 · Qwen Plus · Sonar · stepfun-ai/step3 · Text Embedding V4 · AIHubMix Phi 4 Mini (reasoning) · Qwen Turbo · Aihub Phi 4 Multimodal Instruct · Qwen3 30B A3B · Aihub Phi 4 Mini Instruct · Grok 3 · Aihub Phi 4 · Claude 3 Opus 20240229 · Dall E 3 · Doubao Embedding Text 240715 · Grok 3 Beta · Qwen3 14B · Grok 3 Fast · Qwen3 8B · deepseek-ai/DeepSeek-R1-Zero · Grok 3 Fast Beta · Grok 3 Mini · Qwen3 4B · Grok 3 Mini Beta · Qwen3 1.7b · Qwen3 0.6b · Alicloud Glm 5 · Command A 03 2025 · Grok 3 Mini Fast Beta · Qwen 3 32B · Qwen Turbo 2025 04-28 · Qwen Plus 2025 04-28 · THUDM/GLM-Z1-32B-0414 · THUDM/GLM-4.1V-9B-Thinking · Text Embedding 004 · THUDM/GLM-4-32B-0414 · THUDM/GLM-Z1-9B-0414 · THUDM/GLM-4-9B-0414 · CC Doubao Seed Code Preview · Doubao Seed Code Preview · deepseek-ai/Janus-Pro-7B · GLM Zero Preview · Qwen 3 235B A22B Instruct 2507 · Coding GLM 4.5 Air · Deepinfra Nvidia Nemotron 3 Nano 30B A3b2 · GLM 4.5 Air · GPT 4 32K · Nvidia Llama 3.1 Nemotron 70B Instruct · Nvidia Llama 3.3 Nemotron Super 49B V1.5 · Nvidia Nemotron 3 Nano 30B A3B · Nvidia Nemotron Nano 12B V2 VL · Nvidia Nemotron Nano 9B V2 · O1 Preview 2024 09-12 · Qwen/QVQ-72B-Preview · Qwen/QwQ-32B-Preview · Llama 3.1 Sonar Huge 128K Online · AIHubMix Mistral Large 2411 · Llama 3.1 Sonar Large 128K Online · AIHubMix Mistral Large 2407 · Grok 2 1212 · Llama 3.1 70B · Wan2.6 T2i · DESCRIBE · UPSCALE · Bai Qwen3 VL 235B A22B Instruct · CC MiniMax M2 · CC DeepSeek V3 · CC DeepSeek V3.1 · CC ERNIE 4.5 300B A47B · CC Kimi Dev 72B · CC Kimi K2 Instruct · CC Kimi K2 Instruct 0905 · CC Kimi K2 Thinking · Computer Use Preview · GPT Image Test · grok-4.20-beta-0309-non-reasoning · Grok 4.20 Beta 0309 (reasoning) · Grok 4.20 Multi Agent Beta 0309 · Jina Reader · Jina Search · Llama3.1 8B · O1 2024 12-17 · Sf Kimi K2 Thinking · Baichuan3 Turbo · Baichuan3 Turbo 128K · Baichuan4 · Baichuan4 Air · Baichuan4 Turbo · DeepSeek V3 · Doubao 1.5 Lite 32K · Doubao 1.5 Pro 256K · Doubao 1.5 Pro 32K · Doubao 1.5 Vision Pro 32K · Doubao Lite 128K · Doubao Lite 32K · Doubao Lite 4K · Doubao Pro 128K · Doubao Pro 256K · Doubao Pro 32K · Doubao Pro 4K · GPT-OSS-20B · Gryphe/MythoMax-L2-13b · MiniMax Text 01 · Mistral Large 2407 · Qwen/Qwen2-1.5B-Instruct · Qwen/Qwen2-57B-A14B-Instruct · Qwen/Qwen2-72B-Instruct · Qwen/Qwen2-7B-Instruct · Qwen/Qwen2.5-32B-Instruct · Qwen/Qwen2.5-72B-Instruct · Qwen/Qwen2.5-72B-Instruct-128K · Qwen/Qwen2.5-7B-Instruct · Qwen/Qwen2.5-Coder-32B-Instruct · Qwen3 235B A22B Thinking 2507 · Stable Diffusion 3.5 Large · WizardLM/WizardCoder-Python-34B-V1.0 · AIHubMix Phi 3.5 MoE Instruct · AIHubMix Phi 3.5 Mini Instruct · AIHubMix Phi 3.5 Vision Instruct · AIHubMix Phi 3 Medium 128K · AIHubMix Phi 3 Medium 4K · AIHubMix Phi 3 Small 128K · AIHubMix Codestral 2501 · AIHubMix Cohere Command R · AIHubMix Jamba 1.5 Large · AIHubMix Llama 3.1 405B Instruct · AIHubMix Llama 3.1 70B Instruct · AIHubMix Llama 3.1 8B Instruct · AIHubMix Llama 3.2 11B Vision · AIHubMix Llama 3.2 90B Vision · AIHubMix Llama 3 70B Instruct · AIHubMix Mistral Large · AIHubMix Command R 08 2024 · AIHubMix Command R Plus · AIHubMix Command R Plus 08 2024 · Alicloud Deepseek V3.2 · Alicloud Glm 4.7 · Alicloud Kimi K2 Thinking · Alicloud Kimi K2.5 · Alicloud Minimax M2.5 · Anthropic Opus 4.6 · Azure Deepseek V3.2 · Azure Deepseek V3.2 Speciale · Azure Kimi K2.5 · Cbs Glm 4.7 · Cerebras Llama 3.3 70B · Chatglm_lite · Chatglm_pro · Chatglm_std · Chatglm_turbo · Claude 2 · Claude 2.0 · Claude 2.1 · Claude 3 Haiku 20240229 · Claude 3 Haiku 20240307 · Claude 3 Sonnet 20240229 · Claude Instant 1 · Claude Instant 1.2 · Code Davinci Edit 001 · Cogview 3 · Cogview 3 Plus · Command · Command Light · Command Light Nightly · Command Nightly · Command R · Command R 08 2024 · Command R Plus · Command R Plus 08 2024 · Dall E 2 · Davinci · Davinci 002 · Deepinfra Llama 3.1 8B Instant · Deepinfra Llama 3.3 70B Instant Turbo · Deepinfra Llama 4 Maverick 17B 128e Instruct · Deepinfra Llama 4 Scout 17B 16e Instruct · deepseek-ai/DeepSeek-Coder-V2-Instruct · deepseek-ai/DeepSeek-R1-Distill-Llama-70B · deepseek-ai/DeepSeek-R1-Distill-Llama-8B · deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B · deepseek-ai/DeepSeek-R1-Distill-Qwen-14B · deepseek-ai/DeepSeek-R1-Distill-Qwen-32B · deepseek-ai/DeepSeek-R1-Distill-Qwen-7B · deepseek-ai/DeepSeek-V2-Chat · deepseek-ai/DeepSeek-V2.5 · deepseek-ai/deepseek-llm-67b-chat · deepseek-ai/deepseek-vl2 · DeepSeek V3 · Distil Whisper Large V3 En · Doubao 1.5 Thinking Vision Pro 250428 · Fx Flux 2 Pro · gemini-2.5-pro-exp-03-25 · gemini-embedding-exp-03-07 · gemini-exp-1114 · gemini-exp-1121 · Gemini Pro · Gemini Pro Vision · Gemma 7B It · GLM 3 Turbo · GLM 4 · GLM 4 Flash · GLM 4 Plus · GLM 4.5 Airx · GLM 4 Vision · GLM 4 Vision Plus · Google Gemma 3 12B It · Google Gemma 3 27B It · Google Gemma 3 4B It · google/gemini-exp-1114 · google/gemma-2-27b-it · google/gemma-2-9b-it:free · GPT 3.5 Turbo · GPT 3.5 Turbo 0301 · GPT 3.5 Turbo 0613 · GPT 3.5 Turbo 1106 · GPT 3.5 Turbo 16K · GPT 3.5 Turbo 16K 0613 · GPT 3.5 Turbo Instruct · GPT 4 · GPT 4 0125 Preview · GPT 4 0314 · GPT 4 0613 · GPT 4 1106 Preview · GPT 4 32K 0314 · GPT 4 32K 0613 · GPT 4 Turbo · GPT 4 Turbo 2024 04-09 · GPT 4 Turbo Preview · GPT 4 Vision Preview · GPT 4o 2024 05-13 · GPT 4o Mini 2024 07-18 · gpt-oss-20b · Grok 2 Vision 1212 · Grok Vision Beta · Groq Llama 3.1 8B Instant · Groq Llama 3.3 70B Versatile · Groq Llama 4 Maverick 17B 128e Instruct · Groq Llama 4 Scout 17B 16e Instruct · Jina Embeddings V2 Base Code · Learnlm 1.5 Pro Experimental · Llama 3.1 405B Instruct · Llama 3.1 405B (reasoning) · Llama 3.1 70B Versatile · Llama 3.1 8B Instant · Llama 3.1 Sonar Small 128K Online · Llama 3.2 11B Vision Preview · Llama 3.2 1B Preview · Llama 3.2 3B Preview · Llama 3.2 90B Vision Preview · Llama2 70B 4096 · Llama2 70B 40960 · Llama2 7B 2048 · Llama3 70B 8192 · Llama3 8B 8192 · Llama3 Groq 70B 8192 Tool Use Preview · Llama3 Groq 8B 8192 Tool Use Preview · meta-llama/Llama-3.2-90B-Vision-Instruct · meta-llama/llama-3.1-405b-instruct:free · meta-llama/llama-3.1-70b-instruct:free · meta-llama/llama-3.1-8b-instruct:free · meta-llama/llama-3.2-11b-vision-instruct:free · meta-llama/llama-3.2-3b-instruct:free · meta/llama-3.1-405b-instruct · meta/llama3-8B-chat · mistralai/mistral-7b-instruct:free · Moonshot Kimi K2.5 · Moonshot V1 128K · Moonshot V1 128K Vision Preview · Moonshot V1 32K · Moonshot V1 32K Vision Preview · Moonshot V1 8K · Moonshot V1 8K Vision Preview · nvidia/Llama-3_1-Nemotron-Ultra-253B-v1 · O1 Mini 2024 09-12 · Omni Moderation · Qwen Flash · Qwen Flash 2025 07-28 · Qwen Long · Qwen Max · Qwen Max Longcontext · Qwen Plus · Qwen Turbo · Qwen Turbo 2024 11-01 · Qwen2.5 14B Instruct · Qwen2.5 32B Instruct · Qwen2.5 3B Instruct · Qwen2.5 72B Instruct · Qwen2.5 7B Instruct · Qwen2.5 Coder 1.5b Instruct · Qwen2.5 Coder 7B Instruct · Qwen2.5 Math 1.5b Instruct · Qwen2.5 Math 72B Instruct · Qwen2.5 Math 7B Instruct · Step 2 16K · Text Ada 001 · Text Babbage 001 · Text Curie 001 · Text Davinci 002 · Text Davinci 003 · Text Davinci Edit 001 · Text Embedding 3 Large · Text Embedding 3 Small · Text Embedding Ada 002 · Text Embedding V1 · Text Moderation 007 · Text Moderation · Text Moderation Stable · Text Search Ada Doc 001 · Tts 1 · Tts 1 1106 · Tts 1 Hd · Tts 1 Hd 1106 · Whisper 1 · Whisper Large V3 · Whisper Large V3 Turbo · Yi Large · Yi Large Rag · Yi Large Turbo · Yi Lightning · Yi Medium · Yi VL Plus · DeepSeek R1 Distill Qianfan Llama 8B · Doubao 1.5 Pro 256K 250115 · Doubao 1.5 Pro 32K 250115 · GPT 4o 2024 08-06 Global · GPT 4o Mini Global · Meta Llama 3 70B · Meta Llama 3 8B · O3 Global · O3 Mini Global · O3 Pro Global · Qianfan Chinese Llama 2 13B · Qianfan Llama VL 8B