Models
73 models from 12 providers — every one of them a click away inside EVA.
Anthropic
Claude Haiku 4.5
StandardAnthropic's smallest and fastest model in the Claude 4.5 lineup, optimized for low cost, low latency, and real-time responsiveness, while delivering near-frontier performance on coding, agentic workflows, and extended reasoning tasks.
Claude Sonnet 4.5
PlusAnthropic's most advanced Claude 4.5 model designed for coding, agentic AI workflows, and complex reasoning.
Claude Sonnet 4
PlusA balanced, faster sibling to Opus with strong reasoning and creative abilities.
Claude Sonnet 4.6
PlusAnthropic's advanced Claude 4.6 Sonnet model optimized for coding, complex reasoning, and agentic task execution.
Claude Opus 4.5
PremiumAnthropic's most advanced flagship model in the Claude 4.5 family.
Claude Opus 4.1
PremiumAnthropic's upgraded flagship Claude model from the 4-series, focused on enhanced coding, reasoning, and agentic performance.
Claude Opus 4
PremiumMost capable model of Anthropic —state‑of‑the‑art for coding, extended reasoning, and sustained workflows.
Claude Opus 4.6
PremiumAnthropic's flagship Claude 4.6 Opus model delivering top-tier intelligence for the most demanding tasks.
Claude Opus 4.7
PremiumAnthropic's latest and most capable model with industry-leading intelligence for complex reasoning and autonomous task execution.
DeepSeek
DeepSeek V4 Flash
BudgetDeepSeek's fast and ultra cost-efficient flagship with 1M context and optional thinking mode — the most affordable frontier-class model available.
DeepSeek V4 Pro
BasicDeepSeek's most capable V4 model with frontier-level reasoning, 1M context, and deep optimization for agentic coding workflows.
Gemini 3.1 Flash-Lite
BudgetGoogle's most cost-efficient multimodal model in the Gemini 3.1 family, offering the fastest performance for high-frequency, lightweight tasks with a 1-million-token context window and native thinking support.
Gemini 2.5 Flash-Lite
BudgetUltra-fast lightweight preview focused on speed for simple tasks.
Gemini 3 Flash
BasicGoogle DeepMind's next-generation fast and efficient variant of the Gemini 3 family.
Gemini 2.5 Flash
BasicGoogle's fast and cost-efficient reasoning model with a 1-million-token context window, native multimodal support, and adaptive thinking for a wide range of tasks.
Gemini 2.5 Pro
PlusPowerful multimodal reasoning model with excellent STEM capabilities.
Gemini 3.1 Pro
PremiumRefined Gemini 3 Pro with better thinking, improved token efficiency, and enhanced factual consistency. Optimized for software engineering behavior, agentic workflows, and precise multi-step execution across real-world domains.
Meta
MiniMax
MiniMax M2.5
BudgetMiniMax M2.5 available free via OpenRouter — SOTA real-world productivity model scoring 80.2% on SWE-Bench Verified. Fluent in office automation, agentic workflows, and cross-software coordination.
MiniMax M2.7
StandardMiniMax latest flagship agentic model designed for autonomous real-world productivity, featuring multi-agent collaboration, live debugging, financial modelling, and full document generation across Word, Excel, and PowerPoint.
Mistral
Ministral 3
BudgetSmall, dense multimodal model designed for efficient text and vision understanding with deployment-friendly performance.
Mistral Small 4
BudgetMistral Small 4 unifies reasoning, multimodal understanding, and agentic coding capabilities into a single 24B model — the most capable Mistral Small yet.
Mistral Large 3
BasicState-of-the-art open-weight multimodal Mixture-of-Experts (MoE) LLM with enormous context support, strong reasoning, and enterprise-grade performance.
Magistral Small 1.2
BasicOpen-weight small reasoning LLM from Mistral AI.
Mistral Medium 3.1
BasicFrontier-class multimodal model from Mistral AI.
Magistral Medium 1.2
StandardMistral AI's frontier-class multimodal reasoning model.
Moonshot
OpenAI
authz-1785857475782-group-a
Chats HTTP Test Model
GM Test Model
GM Test Model
Priced Test Model
Chats HTTP Test Model
Priced Test Model
authz-1785857475782-group-b
authz-1785857475782-model
GPT-5.4 nano
BudgetOpenAI's smallest and most cost-efficient GPT-5.4-class model, optimized for high-volume simple tasks with native reasoning support and a 400k context window.
GPT-5 nano
BudgetOpenAI's smallest and fastest GPT-5 variant, optimized for ultra-low latency, high throughput, and cost-efficient text and basic multimodal tasks.
GPT-5.4 mini
BasicOpenAI's strongest mini model for coding, computer use, and subagents. Brings the strengths of GPT-5.4 to a faster, more efficient model designed for high-volume workloads.
GPT-5 mini
BasicA compact version of GPT-5, optimized for speed and efficiency while retaining strong reasoning and basic multimodal support.
GPT-4.1 mini
BasicCompact version of GPT‑4.1 for faster yet capable responses.
o1-mini
BasicA cost-efficient reasoning-optimized variant of OpenAI's o1 family.
GPT-5
StandardOpenAI's next-generation flagship model, offering unparalleled intelligence, reasoning, and multimodal capabilities.
o3
StandardEnhanced reasoning and depth—ideal for heavy-duty logic and research.
GPT-4.1
StandardOpenAI's latest flagship instruction-following model with a 1-million-token context window, strong coding, vision, and agentic task capabilities.
GPT-5.4
PlusOpenAI's frontier model for complex professional work. The first mainline reasoning model incorporating GPT-5.3-Codex capabilities, with native computer use, 1M context, and state-of-the-art token efficiency.
GPT-5.2
PlusOpenAI's newest flagship GPT-5 series model.
GPT-5.5
PremiumOpenAI's newest frontier model representing a new class of intelligence. More intelligent and more token-efficient than GPT-5.4, with native computer use, coding capabilities, and a 1M context window.
GPT-5.5 Pro
PremiumMaximum-performance version of GPT-5.5 using more compute to think harder and provide consistently better answers. The most capable and accurate model in the GPT-5.5 family.
GPT-5.4 Pro
PremiumMaximum-performance version of GPT-5.4 using more compute to think harder. Designed for the hardest professional tasks requiring the highest accuracy.
GPT-5.2 pro
PremiumOpenAI's most advanced professional variant of the GPT-5.2 family.
GPT-5 pro
PremiumOpenAI's advanced high-reasoning version of the GPT-5 series, designed to produce smarter, more precise, and in-depth answers.
o3-pro
PremiumOpenAI's professional-grade reasoning model that uses significantly more compute per query to deliver higher accuracy on the most demanding tasks.
Perplexity
Sonar
BasicLightweight, fast, and real-time search-enabled model for quick, citation-backed web answers.
Sonar Reasoning Pro
StandardPremium reasoning-tier with chain-of-thought logic, large context (127k tokens), and extensive citation.
Sonar Deep Research
StandardExpert-level research model that autonomously performs dozens of web searches, synthesizes sources, and generates comprehensive reports (typically 2–4 minutes per task).
Sonar Reasoning
StandardReal-time reasoning model with multi-step logic and citation sourcing.
Sonar Pro
PlusEnhanced version with larger context and richer citation output for in-depth multi-step answers.
Qwen
Qwen3.5 Flash
BudgetAlibaba's ultra cost-efficient Qwen3.5 Flash model — fast inference with strong reasoning and 1M context at minimal cost.
Qwen3.6 Plus
BasicQwen3.6 Plus builds on a hybrid architecture combining efficient linear attention with sparse MoE routing — major gains in agentic coding, front-end development, and reasoning over the 3.5 series.
Qwen3.5 Plus
BasicHigh-capability Qwen3.5 native vision-language model with hybrid linear attention and sparse MoE architecture, delivering state-of-the-art performance across reasoning, coding, and multimodal tasks.
xAI
Grok 4.3
PlusxAI's newest flagship model with strong agentic tool calling, instruction following, low hallucination rate, and configurable reasoning.
Grok 4.1 Fast
BudgetxAI's high-performance variant of Grok 4.1 optimized for low-latency, tool-calling, and agentic workflows with extremely large context support and multimodal capabilities.
Grok 4.1 Fast (Non-Reasoning)
BudgetA variant of xAI's Grok 4.1 Fast optimized for low-latency, instantaneous responses by skipping internal reasoning tokens; best for straightforward, high-throughput text completion with huge context windows.
Grok 4 Fast
BudgetSpeed- and cost-optimized variant of xAI's flagship Grok 4 model.
Grok 4 Fast (Non-Reasoning)
BudgetA cost- and speed-optimized variant of xAI's Grok 4 Fast.
Grok 4
PlusxAI’s flagship model—strong coding, reasoning, and conversational skills.
Grok 4.20
PlusxAI's newest flagship model with industry-leading speed, agentic tool calling, and optional reasoning. Combines low hallucination rate with strict prompt adherence.
Grok 4.20 (Non-Reasoning)
PlusxAI's flagship Grok 4.20 in fast non-reasoning mode — same frontier capabilities with lower latency when chain-of-thought is not needed.
Zhipu GLM
GLM-5
BasicZ.ai's flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Delivers production-grade performance on large-scale programming tasks with advanced agentic planning and iterative self-correction.
GLM-5.1
BasicGLM-5.1 delivers a major leap in long-horizon coding capability — capable of working independently for 8+ hours on a single task, autonomously planning, executing, and self-improving to deliver complete engineering-grade results.