Explore the models available in OmnyChat
Compare language, image and video models by their capabilities, generation options and plan requirements.
Language models (LLMs)
GPT· OpenAI
GPT-6 Astra
OpenAI
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work.
GPT-5.6 Luna Pro
OpenAI
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
GPT-5.6 Luna
OpenAI
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series.
GPT-5.6 Terra Pro
OpenAI
GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
GPT-5.6 Terra
OpenAI
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier.
GPT-5.6 Sol Pro
OpenAI
GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
GPT-5.6 Sol
OpenAI
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series.
GPT-5.5 Pro
OpenAI
GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads.
GPT-5.5
OpenAI
GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks.
GPT-5.4 Nano
OpenAI
GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks.
GPT-5.4 Mini
OpenAI
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads.
GPT-5.4 Pro
OpenAI
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks.
GPT-5.4
OpenAI
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system.
GPT-5.3-Codex
OpenAI
GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2.
GPT-5.2-Codex
OpenAI
GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows.
GPT-5.2 Chat
OpenAI
GPT-5.2 Chat (AKA Instant) is the fast, lightweight member of the 5.2 family, optimized for low-latency chat while retaining strong general intelligence.
GPT-5.2 Pro
OpenAI
GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro.
GPT-5.2
OpenAI
GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1.
GPT-5.1-Codex-Max
OpenAI
GPT-5.1-Codex-Max is OpenAI’s latest agentic coding model, designed for long-running, high-context software development tasks.
GPT-5.1
OpenAI
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5.
GPT-5.1-Codex
OpenAI
GPT-5.1-Codex is a specialized version of GPT-5.1 optimized for software engineering and coding workflows.
GPT-5.1-Codex-Mini
OpenAI
GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex
GPT-5 Image Mini
OpenAI
GPT-5 Image Mini combines OpenAI's advanced language capabilities, powered by [GPT-5 Mini](https://openrouter.ai/openai/gpt-5-mini), with GPT Image 1 Mini for efficient image generation.
GPT-5 Image
OpenAI
[GPT-5](https://openrouter.ai/openai/gpt-5) Image combines OpenAI's GPT-5 model with state-of-the-art image generation capabilities.
GPT-5 Pro
OpenAI
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience.
GPT-5
OpenAI
GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience.
GPT-5 Mini
OpenAI
GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks.
GPT-5 Nano
OpenAI
GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments.
gpt-oss-20b
OpenAI
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license.
o3 Pro
OpenAI
The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning.
o3
OpenAI
o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks.
o4 Mini
OpenAI
OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities.
GPT-4.1
OpenAI
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning.
GPT-4.1 Mini
OpenAI
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost.
GPT-4.1 Nano
OpenAI
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series.
o3 Mini
OpenAI
OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding.
GPT-4o-mini
OpenAI
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs.
GPT-4o
OpenAI
GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs.
Claude· Anthropic
Claude Fable 5.1
Anthropic
Claude Fable 5.1 is a Anthropic language model available for chat in OmnyChat.
Claude Opus 5
Anthropic
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work.
Claude Sonnet 5
Anthropic
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work.
Claude Opus 4.8
Anthropic
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family.
Claude Opus 4.7
Anthropic
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents.
Claude Sonnet 4.6
Anthropic
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work.
Claude Opus 4.6
Anthropic
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks.
Claude Opus 4.5
Anthropic
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use.
Claude Haiku 4.5
Anthropic
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models.
Claude Sonnet 4.5
Anthropic
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows.
Claude Opus 4.1
Anthropic
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks.
Claude Opus 4
Anthropic
Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows.
Claude Sonnet 4
Anthropic
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability.
Gemini· Google
Gemini 3.8 Flash
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Gemini 3.7 Flash
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning.
Gemini 3.6 Flash
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development.
Gemini 3.5 Flash
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed.
Gemini 3.1 Pro Preview
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows.
Gemini 3 Flash Preview
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance.
Gemini 2.5 Flash Lite
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency.
Gemini 2.5 Flash
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks.
Gemini 2.5 Pro
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks.
Grok· xAI
Grok 4.6
xAI
Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.
Grok 4.5
xAI
Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.
Grok 4.3
xAI
Grok 4.3 is a reasoning model from SpaceXAI.
Grok 4.20
xAI
Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities.
DeepSeek· DeepSeek
DeepSeek V4 Pro 0423
DeepSeek
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window.
DeepSeek V4 Flash 0423
DeepSeek
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window.
DeepSeek V3.2
DeepSeek
DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance.
DeepSeek V3.2 Exp
DeepSeek
DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures.
DeepSeek V3.1 Terminus
DeepSeek
DeepSeek V3.1 Terminus is a DeepSeek language model available for chat in OmnyChat.
DeepSeek V3.1
DeepSeek
DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates.
R1
DeepSeek
DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens.
DeepSeek V3
DeepSeek
DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous versions.
GLM· Z-AI
GLM 5.3 Flash
Z-AI
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks.
GLM 5.3
Z-AI
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks.
GLM 5.2
Z-AI
GLM 5.2 is a large-scale reasoning model from Z.ai.
GLM 5.1
Z-AI
GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks.
GLM 5V Turbo
Z-AI
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks.
GLM 5 Turbo
Z-AI
GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios.
GLM 5
Z-AI
GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows.
GLM 4.7 Flash
Z-AI
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency.
GLM 4.7
Z-AI
GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution.
GLM 4.6
Z-AI
GLM 4.6 is a Z-AI language model available for chat in OmnyChat.
Kimi· MoonshotAI
Kimi K3
MoonshotAI
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI.
Kimi K2.6
MoonshotAI
Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration.
Kimi K2.5
MoonshotAI
Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm.
Kimi K2 Thinking
MoonshotAI
Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning.
Kimi K2 0711
MoonshotAI
Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass.
Qwen· Qwen
Qwen3.8 Flash
Qwen
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.
Qwen3.8 27B
Qwen
Qwen3.8 27B is an open-weight dense vision-language model from Qwen.
Qwen3.7 Flash
Qwen
Qwen3.7 Flash is a vision-language reasoning model from Alibaba.
Qwen3.7 Plus
Qwen
Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series.
Qwen3.7 Max
Qwen
Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series.
Qwen3.6 Flash
Qwen
Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window.
Qwen3.6 35B A3B
Qwen
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token.
Qwen3.6 Max Preview
Qwen
Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters.
Qwen3.6 27B
Qwen
Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026.
Qwen3.6 Plus
Qwen
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference.
Qwen3.5-9B
Qwen
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture.
Qwen3.5-35B-A3B
Qwen
The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency.
Qwen3.5-27B
Qwen
The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance.
Qwen3.5-122B-A10B
Qwen
The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency.
Qwen3.5 397B A17B
Qwen
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency.
Qwen3 Max Thinking
Qwen
Qwen3-Max-Thinking is the flagship reasoning model in the Qwen3 series, designed for high-stakes cognitive tasks that require deep, multi-step reasoning.
Qwen3 Coder Next
Qwen
Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows.
Qwen3 Max
Qwen
Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version.
Qwen3 Coder 480B A35B
Qwen
Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team.
Qwen3 32B
Qwen
Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue.
Qwen3 235B A22B
Qwen
Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass.
Qwen2.5 Coder 32B Instruct
Qwen
Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen).
Llama· Meta
Llama 4 Maverick
Meta
Llama 4 Maverick is a Meta language model available for chat in OmnyChat.
Llama 4 Scout
Meta
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B.
Llama 3.3 70B Instruct
Meta
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out).
Mistral· Mistral
Mistral Medium 3.5
Mistral
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI.
Mistral Small 4
Mistral
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system.
Devstral 2 2512
Mistral
Devstral 2 is a state-of-the-art open-source model by Mistral AI specializing in agentic coding. It is a 123B-parameter dense transformer model supporting a 256K context window.
Mistral Large 3 2512
Mistral
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Mistral Medium 3
Mistral
Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost.
Mistral Large 2407
Mistral
This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more.
MiniMax· MiniMax
MiniMax M2.7
MiniMax
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement.
MiniMax M2.5
MiniMax
MiniMax-M2.5 is a SOTA large language model designed for real-world productivity.
MiniMax M2.1
MiniMax
MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development.
MiniMax M2
MiniMax
MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows.
Image models
View allGPT Image 2.5 Sunburst
OpenAI
Detailed creative work with precise prompt following. Allow extra generation time.
Flux 2
Black Forest Labs
Lightweight, fast generation. Great for quick iterations.
Flux 2 Pro
Black Forest Labs
Professional-grade creative generation with fine detail.
Nano Banana
Balanced speed and quality for everyday use.
Seedream v4.5
ByteDance
High-quality generation from ByteDance.
GPT Image 1.5
OpenAI
OpenAI image generation with strong prompt adherence.
GPT Image 2
OpenAI
Text rendering and photorealistic image generation.
Nano Banana 2
Latest nano banana with balanced speed and quality.
Nano Banana Pro
Fast, high-quality generation. Best-in-class output.
GPT Image 2.5 Sunburst Edit
OpenAI
Focused edits with reference images for subject and composition consistency.
Nano Banana Pro (Identity)
Soul-aware scene generation. Accepts multiple ref images via image_urls.
Flux Kontext Pro
Black Forest Labs
Character-consistent scene generation from a single reference image.
PuLID
Tencent
Strongest face-identity lock. Up to 4 ref images. Best for Headshot Studio.
Video models
View allSeedance 2.0 Fast
ByteDance
Fast ByteDance video generation with optional audio.
Seedance 2.0
ByteDance
ByteDance's flagship video model with synchronized audio.
Kling Standard
Kuaishou
Kling 3.0 Standard. fluid motion with optional audio.
Kling Pro
Kuaishou
Kling 3.0 Pro. cinematic visuals, native audio, top-tier quality.
Veo 3.1 Fast
Google Veo 3.1 Fast. quick generations up to 4K with optional audio.
Veo 3.1
Google DeepMind Veo 3.1. flagship cinematic video with audio.
FLUX 3 Video
Black Forest Labs
Black Forest Labs' multimodal flagship with native audio and multi-shot direction.
Grok Imagine
xAI
xAI's budget-friendly video model with baked-in audio.