Model access depends on your plan. Standard models are available on all plans. Premium (rate-limited) models require Pro or above. See Plans and Pricing for quota details.
Model Types
Text / Chat Models
Text models handle all conversation-based tasks: reasoning, writing, coding, analysis, research, summarization, translation, and more. ZeroTwo hosts text models from 17+ providers.- Reasoning models — step-by-step logical thinking, math, complex coding
- Writing models — long-form documents, tone control, nuance
- Coding models — code generation, debugging, refactoring, explanation
- Analysis models — document understanding, data interpretation, research
Image Generation Models
Image models are available in the Studio at/studio/images and via the Image tool pill in chat. Choose from photorealistic, artistic, multilingual, and specialized generation styles.
Video Generation Models
Video models are available at/studio/video. Generate short clips from text prompts or animate images. Video generation runs as a background job; outputs are saved to your files library.
Audio Models
Audio models are available at/studio/audio and as file attachments in chat. Generate speech, music, and sound effects; automatically transcribe spoken audio via Whisper.
Provider Overview: Text Models
Image Generation Models
ZeroTwo’s image studio at/studio/images supports multiple generation models with different style characteristics and quality levels.
Reasoning Models
Several models in ZeroTwo support extended reasoning — they work through a problem step-by-step before producing a final answer. When you select a reasoning-capable model, ZeroTwo shows a thinking level slider in the prompt bar. Reasoning-capable models:- OpenAI o3 — OpenAI’s strongest reasoning model
- OpenAI o4-mini — faster, more efficient reasoning
- DeepSeek Reasoner — strong math and logical reasoning
- Claude Sonnet 4.6, Claude Opus 4.6 — extended thinking mode available
Higher thinking levels use more tokens and take longer, but produce more accurate and thorough answers for difficult problems. On plans with premium quotas, high-level reasoning responses consume more of your quota.
When to use reasoning models:
- Complex math, statistics, and proofs
- Algorithmic coding challenges and debugging multi-file logic
- Structured decision-making and logical analysis
- Research synthesis requiring careful integration of multiple sources
Premium vs. Standard Models
Premium (rate-limited) Models
These are the highest-capability, most in-demand models. Sending a message with a premium model counts against your monthly quota on Pro and Pro 2x plans:- Claude Opus 4.6 variants, Claude Sonnet 4.6 variants (Anthropic)
- GPT-5, GPT-4o (OpenAI)
- Gemini 2.5 Pro, Gemini 3.1 Pro (Google)
- Grok-4 (xAI)
- Cohere Command A (Cohere)
- Mistral Large (Mistral)
- Qwen Max (Qwen)
- Perplexity Sonar Pro (Perplexity)
Standard (non-rate-limited) Models
Standard models do not count against your premium quota and are available in unlimited quantities on all paid plans. They include:- GPT-5-mini, GPT-4o-mini, GPT-4.1 (OpenAI)
- Claude Haiku 4.5 (Anthropic)
- Gemini 2.5 Flash, Gemini Flash Lite (Google)
- DeepSeek Chat, DeepSeek Coder (DeepSeek)
- Grok-3, Grok Code Fast (xAI)
- Mistral Small, Magistral, Nemo (Mistral)
- Command R, Command R+, Command R7b (Cohere)
- Qwen Plus, Qwen Turbo, Qwen Flash, Qwen3 variants (Qwen)
- All Groq-hosted models
- All Kimi K2, Venice, TheSys, ZAI, Inception, ByteDance models
Fallback Models
When your premium quota is exhausted on Pro or Pro 2x, ZeroTwo automatically routes messages to a fallback standard model: GPT-5-mini, GPT-4o-mini, Gemini Flash Lite, Mistral Small, or Grok 4 Fast.Not Sure Which Model to Use?
For writing and analysis
For writing and analysis
Claude Sonnet 4.6 and Claude Opus 4.6 are the strongest choices for nuanced long-form writing, document analysis, and tasks requiring careful reasoning about tone and context. Claude’s 200k token context window makes it ideal for large documents. Gemini 2.5 Pro handles even longer documents (1M token context).
For coding
For coding
GPT-5 and DeepSeek Coder are strong for code generation, debugging, and refactoring. o3 and o4-mini are ideal for algorithmic problems requiring step-by-step reasoning. DeepSeek Coder is a standard model — excellent for developers on Free or Pro plans who want to preserve premium quota.
For real-time information
For real-time information
Perplexity Sonar Pro and Sonar are built for web-grounded answers with citations. Alternatively, enable Web Search with any tool-capable model (GPT-5, Claude Sonnet 4.6, Gemini 2.5 Pro) to ground responses in live data.
For speed
For speed
Groq-hosted models offer ultra-low latency. Gemini Flash Lite, GPT-4o-mini, and Mistral Small are fast standard models suitable for quick iteration, brainstorming, and high-volume tasks.
For multilingual tasks
For multilingual tasks
Qwen models (Qwen Max, Qwen3) have strong Chinese-language and multilingual capabilities. Mistral models perform well across European languages. Claude and GPT-5 handle a broad range of languages well.
For very long documents
For very long documents
Gemini 2.5 Pro with its 1M token context window is the best choice for extremely long documents, large codebases, or extended research sessions. Claude Sonnet 4.6 (200k tokens) is a strong alternative with excellent comprehension.
Related Pages
- Model Picker — how to select, search, and switch models in the chat interface
- Plans and Pricing — which models count as premium and quota details
- Plan Availability Matrix — full feature-by-plan comparison
- Answer Quality and Limitations — choosing the right model for accuracy

