Models

Access the world's most powerful language models from leading providers. MCP, reasoning, web search, and more.

Capabilities

MCP Support

Connect external tools & APIs

Thinking + Web Search

Advanced reasoning & web access

File Handling

Images, documents & PDFs

Browse by Provider

Anthropic

Anthropic

18 models

OpenAI

OpenAI

29 models

Gemini

Gemini

9 models

xAI

xAI

14 models

OpenRouter

OpenRouter

42 models

All Models

Anthropic

Anthropic

18 models available

Documentation

Claude Fable 5

Anthropic's most capable publicly-available model. Designed for ambitious long-running agentic tasks — software engineering, scientific research, and complex knowledge work. Features mandatory adaptive thinking, 1M token context window, 128K output capacity, and all effort levels including xhigh and max. Priced at $10/$50 per million tokens.

View details

Claude Opus 5

Anthropic's latest Opus model for complex agentic coding and enterprise work — approaching Claude Fable 5's capability at half the price ($5/$25 per million tokens). Features a 1M token context window, 128K output capacity, adaptive thinking on by default with full effort control (low through max), and stronger self-verification and error recovery so it needs less back-and-forth. Knowledge cutoff May 2026.

View details

Claude Opus 4.8

Previous-generation Opus model. ~4x more reliable code review, top Super-Agent benchmark (only model to complete every case end-to-end), and 84% on Online-Mind2Web computer use. Features 1M token context window, 128K output capacity, adaptive thinking with effort control, and priority service tier. Anthropic recommends Claude Opus 5 for new work.

View details

Claude Opus 4.7

Previous-generation Opus model. Features 1M token context window with new tokenizer, 128K output capacity, adaptive thinking (only supported thinking mode), high-resolution image support (2576px / 3.75MP), task budgets, and xhigh effort level. Strong at complex reasoning, agentic coding, knowledge work, and vision tasks; Anthropic recommends Claude Opus 5 for new work.

View details

Claude Sonnet 4.6

Anthropic's flagship balanced model with near-Opus level performance at mid-tier pricing. Features 1M token context window, 128K output capacity, exceptional coding capabilities (79.6% on SWE-bench), and advanced computer use (72.5% on OSWorld). Ideal for complex agents, long-context reasoning, and production workflows.

View details

Claude Sonnet 5

Anthropic's latest balanced model with adaptive thinking on by default. Features 1M token context window, 128K output capacity, stronger coding and agentic capabilities, and improved vision. Uses a new tokenizer and the effort parameter for reasoning control. Introductory pricing at $2/$10 per million tokens through August 2026.

View details

Claude Sonnet 4.5

Anthropic's smart model for complex agents and coding tasks. Offers the best balance of intelligence, speed, and cost for most use cases, with exceptional performance in coding and agentic workflows.

View details

Claude Opus 4.6

Anthropic's most intelligent model with breakthrough capabilities. Features 1M token context window, 128K output capacity, adaptive thinking for optimal reasoning, and agent teams for complex workflows. Best for enterprise agents, production coding, and sophisticated problem-solving.

View details

Claude Opus 4.5

Anthropic's premium model combining maximum intelligence with practical performance. Delivers state-of-the-art coding capabilities and frontier performance for production code, sophisticated agents, and complex tasks.

View details

Claude Haiku 4.5

Anthropic's fastest model with near-frontier intelligence. Matches Sonnet 4's performance on coding, computer use, and agent tasks while being optimized for speed and cost-efficiency.

View details

Claude Opus 4.1

Legacy model. Anthropic recommends migrating to Claude 4.5 models for improved performance.

View details

Claude Sonnet 4

Legacy model. Anthropic recommends migrating to Claude 4.5 models for improved performance.

View details

Claude Opus 4

Legacy model. Anthropic recommends migrating to Claude 4.5 models for improved performance.

View details

Claude Sonnet 3.7

Deprecated

Deprecated model scheduled for retirement. Anthropic recommends migrating to Claude 4.5 models for improved performance.

View details

Claude Sonnet 3.5

Deprecated

⚠️ Retired model no longer available via API. This model was retired on October 28, 2025. Please migrate to Claude Sonnet 4.5.

View details

Claude Haiku 3.5

Fast and efficient model from Anthropic's Claude 3.5 series, designed for quick, straightforward tasks.

View details

Claude Opus3

Deprecated

⚠️ Retired model no longer available via API. This model was retired on January 5, 2026. Please migrate to Claude Opus 4.5.

View details

Claude Haiku 3

Anthropic's fastest, most compact Claude 3 model for near-instant responsiveness. Answers simple queries and requests with unmatched speed.

View details
OpenAI

OpenAI

29 models available

Documentation

GPT-5.6 Sol

OpenAI's flagship frontier model for complex professional work. Official API docs position Sol as the default starting point for complex reasoning and coding. Features a 1.05M token context window, 128K max output, image understanding, and max reasoning effort. The OpenAI API also exposes the alias 'gpt-5.6' for this model family default. OpenAI documents additional native tools for this family, but Aisle currently exposes the verified web search path only.

View details

GPT-5.6 Terra

OpenAI's GPT-5.6 model that balances intelligence and cost. Features a 1.05M token context window, 128K max output, image understanding, and max reasoning effort. OpenAI documents additional native tools for this family, but Aisle currently exposes the verified web search path only.

View details

GPT-5.6 Luna

OpenAI's GPT-5.6 model optimized for cost-sensitive, high-volume workloads. Features a 1.05M token context window, 128K max output, image understanding, and max reasoning effort. OpenAI documents additional native tools for this family, but Aisle currently exposes the verified web search path only.

View details

GPT-5.4

OpenAI's most capable frontier model. Features 1.05M token context window, native compaction for longer agent trajectories, 33% fewer errors than GPT-5.2, and 18% fewer response-level errors. First mainline model with computer-use capabilities.

View details

GPT-5.4 Pro

OpenAI's maximum-capability model for the most demanding professional workloads. Features 1.05M context window with restricted reasoning options (medium through xhigh) for consistently deep analysis.

View details

GPT-5.5

OpenAI's first fully retrained base model since GPT-4.5. Features 1.05M token context window, ~40% improved token efficiency, and reasoning defaulted to 'low' (rather than 'none') for stronger out-of-the-box answers. Same reasoning controls as GPT-5.4 — 'none' through 'xhigh' available.

View details

GPT-5.5 Pro

OpenAI's maximum-capability model in the GPT-5.5 family. Features 1.05M token context window with restricted reasoning options (medium through xhigh) for consistently deep analysis. Same parameter shape as GPT-5.4 Pro.

View details

GPT-5.2 Pro

OpenAI's most capable model for professional knowledge work, combining maximum reasoning capabilities with strong performance on complex, multi-step tasks across spreadsheets, presentations, coding, and long contexts.

View details

GPT-5.2

OpenAI's advanced model for complex knowledge work. Excels at building spreadsheets and presentations, writing code, interpreting images, and working with long contexts. Performs at or above human expert level on well-specified knowledge work tasks.

View details

GPT-5.2 Instant

Fast variant of GPT-5.2 optimized for responsive professional tasks. Delivers GPT-5.2's capabilities with reduced latency for real-time applications.

View details

GPT-5.3 Instant

Fast variant of GPT-5.3 optimized for everyday tasks. Delivers responsive, low-latency performance for real-time applications and conversational workflows.

View details

GPT-5.3 Codex

OpenAI's most capable agentic coding model. Combines Codex and GPT-5 training stacks for best-in-class code generation, reasoning, and general-purpose intelligence. 25% faster than its predecessor. Optimized for long-running agentic coding tasks.

View details

GPT-5

OpenAI's breakthrough model with significant improvements in math, coding, visual perception, and health. 94.6% on AIME 2025, 74.9% on SWE-bench Verified. ~45% less likely to hallucinate than GPT-4o.

View details

GPT-5 mini

Efficient variant of GPT-5 for cost-effective intelligent tasks. Balances strong performance with reduced computational requirements.

View details

GPT-5 nano

Ultra-efficient GPT-5 model optimized for high-volume, cost-sensitive applications. The fastest and most affordable option in the GPT-5 series.

View details

GPT-5.1

Iterative improvement of GPT-5 with more conversational style, improved instruction following, and adaptive reasoning capabilities.

View details

GPT-4.1

OpenAI model with major improvements in coding and instruction following. Completes 54.6% of SWE-bench Verified tasks vs 33.2% for GPT-4o. Features 1M token context window and improved long-context comprehension.

View details

GPT-4.1 mini

Balanced small model matching GPT-4o intelligence at nearly half the latency and 83% lower cost. Features 1M token context window with strong coding and instruction-following performance.

View details

GPT-4.1 nano

OpenAI's fastest and cheapest model with 1M token context window. Scores 80.1% on MMLU. Optimized for high-volume, low-latency applications requiring basic intelligence.

View details

GPT-4o

Deprecated

⚠️ Retired model no longer available via API. This model was retired on February 17, 2026. OpenAI recommends using GPT-5.x or GPT-4.1 models instead.

View details

GPT-4o with Web Search

Deprecated

⚠️ Retired model no longer available via API. GPT-4o was retired alongside the GPT-4o family. OpenAI recommends using GPT-5.x or GPT-4.1 models instead.

View details

GPT-4.5 Preview

Deprecated

⚠️ Retired model no longer available via API. This model was retired on July 14, 2025. OpenAI recommends using GPT-4.1 or o-series models instead.

View details

GPT-4o mini

Fast, affordable small model for focused tasks. Accepts text and image inputs with structured outputs. Optimized for cost-effective intelligent processing.

View details

o1

OpenAI reasoning model trained with reinforcement learning to perform complex reasoning tasks. Designed to think deeply before responding.

View details

o3

OpenAI's most powerful reasoning model. Pushes the frontier across coding, math, science, and visual perception. 20% fewer major errors than o1 on real-world tasks. First reasoning model with agentic tool use including web search and image understanding.

View details

o4-mini

Fast, cost-efficient reasoning model optimized for high-volume STEM tasks. Best-performing benchmarked model on AIME 2024 and 2025. Supports higher usage limits than o3 for throughput-sensitive workloads.

View details

o3-mini

Small reasoning model optimized for science, math, and coding tasks. Offers efficient complex reasoning at lower computational cost.

View details

DALL-E 3

OpenAI's most advanced image generation model. Creates high-quality images from text prompts with improved accuracy and detail. Supports 1024x1024 resolution.

View details

GPT Image 1

OpenAI's latest image generation model with improved quality and instruction following.

View details
Gemini

Gemini

9 models available

Documentation

Gemini 3.1 Pro Preview

Google's most capable model for multimodal understanding and advanced agentic tasks. Successor to Gemini 3 Pro Preview with improved reasoning across text, images, video, audio, and PDFs.

View details

Gemini 3 Pro Preview

Deprecated

Google's best model for multimodal understanding and the most powerful agentic model yet. Delivers richer visuals and deeper interactivity with state-of-the-art reasoning capabilities across text, images, video, audio, and PDFs.

View details

Gemini 3 Flash Preview

Google's most balanced model built for speed, scale, and frontier intelligence. Default model across several Google surfaces with strong coding and state-of-the-art reasoning. Supports thinking, structured outputs, and search grounding across text, images, video, audio, and PDFs.

View details

Gemini 3.1 Flash Lite Preview

Google's fastest and most cost-efficient Gemini 3.1 model. Optimized for high-throughput, low-latency applications with strong reasoning across text, images, video, audio, and PDFs.

View details

Gemini 2.5 Pro

Google's state-of-the-art thinking model, capable of reasoning over complex problems in code, math, and STEM. Excels at analyzing large datasets, codebases, and documents using long context windows. Scheduled for retirement October 16th, 2026.

View details

Gemini 2.5 Flash

Google's best model in terms of price-performance, offering well-rounded capabilities. Optimized for large-scale processing and low-latency, high-volume tasks with thinking capabilities. Scheduled for retirement October 16th, 2026.

View details

Gemini 2.5 Flash Lite

Google's fastest flash model optimized for cost-efficiency and high throughput. Ultra-fast processing designed for maximum efficiency in high-volume applications. Scheduled for retirement October 16th, 2026.

View details

Imagen 4

Google's latest image generation model.

View details

Imagen 4 Ultra

Google's highest quality image generation model.

View details
xAI

xAI

14 models available

Documentation

Grok 4.20 Reasoning

xAI's latest flagship reasoning model (March 2025). Strong at complex multi-step tasks, coding, and analysis. Supports X search and web search. 256K context.

View details

Grok 4.20

xAI's latest flagship model without reasoning overhead (March 2025). Faster responses for direct tasks. Supports X search and web search. 256K context.

View details

Grok 4.20 Multi-Agent

xAI's Grok 4.20 optimised for multi-agent orchestration (March 2025). Best suited for tasks involving coordination across multiple AI agents.

View details

Grok 4.1 Fast Reasoning

xAI's fast reasoning model. Balances depth with low latency for production reasoning workloads. Supports X search and web search. 256K context.

View details

Grok 4.1 Fast

xAI's fastest grok-4 generation model. Minimal latency, direct responses without reasoning overhead. Supports X search and web search. 256K context.

View details

Grok 4

xAI's most capable model with X search support

View details

Grok 4 (Jul 2025)

Grok 4 pinned to the July 2025 release

View details

Grok 3

xAI's most powerful model with deep reasoning capabilities. Excels at complex analysis, coding, and multi-step problem solving with a 1M token context window.

View details

Grok 3 Fast

xAI's high-speed model built for low-latency production workloads. Delivers rapid responses while maintaining strong reasoning capabilities. 1M token context window.

View details

Grok 3 Mini

xAI's compact, cost-efficient reasoning model. Supports reasoning_effort control for balancing speed vs depth. Ideal for high-volume tasks.

View details

Grok 3 Mini Fast

xAI's fastest compact model. Same capabilities as Grok 3 Mini with optimised infrastructure for minimal latency. Supports reasoning_effort control.

View details

Grok 2 Vision

xAI's multimodal model with image understanding. Accepts JPEG and PNG images up to 20MB alongside text for richer, context-aware responses.

View details

Grok Imagine

xAI's image generation model.

View details

Grok Imagine Pro

xAI's premium image generation model with enhanced quality.

View details
OpenRouter

OpenRouter

42 models available

Documentation

MoonshotAI Kimi K2.5

MoonshotAI's Kimi K2.5 model via OpenRouter. Features massive 262K token context window for handling very long documents and conversations.

View details

Writer Palmyra X5

Writer's Palmyra X5 model via OpenRouter. Enterprise AI agent model with 1.04M token context window. Features novel transformer architecture and hybrid attention mechanisms for faster inference and expanded memory, optimized for processing large enterprise data volumes.

View details

MiniMax M2.1

MiniMax's M2.1 model via OpenRouter. Consistently ranks in top-10 across benchmarks with strong performance in reasoning and coding tasks. Features 196K context window.

View details

Xiaomi MiMo V2 Flash

Xiaomi's MiMo V2 Flash model via OpenRouter. Ranked #1 on SWE-bench benchmark with 309B total parameters and 15B active parameters. Features hybrid-thinking toggle for optimal performance in coding tasks.

View details

Devstral 2

Mistral's Devstral 2 model via OpenRouter. 123B-parameter dense transformer model specializing in agentic coding across multiple files. Features 262K context window for handling complex codebases.

View details

Amazon Nova 2 Lite

Amazon's Nova 2 Lite model via OpenRouter. Multimodal model supporting text, image, video, and file inputs. Excels at document processing, video information extraction, code generation, and multi-step agentic workflows. Features 1M token context window.

View details

xAI Grok 4.1 Fast

xAI's best agentic tool calling model via OpenRouter. Features 2 million token context window, exceptional for customer support, research, and complex agentic workflows. Supports reasoning toggle and web search.

View details

Amazon Nova Premier

Amazon's Nova Premier model via OpenRouter. Flagship AWS model with 1M token context window designed for complex reasoning tasks. Multimodal support for text and image inputs. Can be used as a teacher model for custom distillation.

View details

Mistral Medium 3.1

Mistral AI's mid-tier model via OpenRouter. Balances performance and cost, suitable for a wide range of tasks requiring solid reasoning and generation capabilities.

View details

Codestral

Mistral AI's specialized coding model via OpenRouter. Optimized for code generation, completion, and technical tasks with exceptional programming capabilities across multiple languages.

View details

Gemini 2.5 Pro

Google's state-of-the-art thinking model via OpenRouter. Capable of reasoning over complex problems in code, math, and STEM with long context support. Scheduled for retirement June 17, 2026.

View details

Qwen3 235B

Alibaba's largest open-source Qwen3 model via OpenRouter with 235B parameters. Apache 2.0 licensed with exceptional multilingual capabilities and strong performance across diverse tasks.

View details

Qwen3 32B

Alibaba's balanced Qwen3 model via OpenRouter with 32B parameters. Apache 2.0 licensed, offering strong performance with efficient resource usage and multilingual support.

View details

Qwen3 14B

Alibaba's efficient Qwen3 model via OpenRouter with 14B parameters. Apache 2.0 licensed, optimized for speed and cost-effectiveness while maintaining quality across multiple languages.

View details

xAI Grok 3 Beta

xAI's Grok 3 model available via OpenRouter. Beta version of xAI's latest language model with enhanced capabilities.

View details

Claude 3.7 Sonnet

Deprecated

Deprecated model scheduled for retirement, available via OpenRouter. Anthropic recommends migrating to Claude 4.5 models for improved performance.

View details

Gemini Flash 2.0

Google's fast, efficient Gemini 2.0 model via OpenRouter. Optimized for responsive multimodal tasks with strong performance. Scheduled for retirement March 31, 2026.

View details

o3-mini

Small reasoning model optimized for science, math, and coding tasks via OpenRouter. Offers efficient complex reasoning at lower computational cost.

View details

Perplexity Sonar

Perplexity AI's Sonar model via OpenRouter with native web search integration. Provides real-time information by searching the web and synthesizing current results into responses.

View details

Liquid LFM 2.5 1.2B Thinking (Free)

LiquidAI's lightweight thinking model via OpenRouter. Compact 1.2B parameter model optimized for edge devices and resource-constrained environments. Excels at agentic tasks, data extraction, and RAG applications. Free tier available.

View details

Qwen3 VL 32B

Alibaba's Qwen3 VL multimodal model via OpenRouter. Supports text, images, and video inputs with 262K context window. Features strong visual tool use capabilities and agentic interaction for complex multimodal tasks.

View details

o1

OpenAI reasoning model trained with reinforcement learning via OpenRouter. Designed to think deeply before responding on complex reasoning tasks.

View details

Llama 3.3

Meta's open-source Llama 3.3 model (70B parameters) via OpenRouter. Instruction-tuned for helpful, high-quality responses.

View details

OpenRouter Free

OpenRouter's intelligent free model router via OpenRouter. Dynamically selects from available free models and smartly filters for required features like image understanding, tool calling, and structured outputs. Perfect for cost-conscious applications.

View details

StepFun Step 3.5 Flash (Free)

StepFun's Step 3.5 Flash free model via OpenRouter. 196B sparse MoE with 11B active parameters optimized for reasoning tasks. Features excellent speed and efficiency across long contexts. Free tier available.

View details

Arcee Trinity Large Preview (Free)

Arcee AI's Trinity Large free model via OpenRouter. 400B-parameter sparse MoE excelling in creative writing and agentic navigation. Suitable for real-world applications with free tier access.

View details

Relace Search

Relace's agentic search model via OpenRouter. Specialized for multi-step codebase exploration with parallel tool usage. Claims 4x faster performance than frontier models for code search tasks.

View details

Claude 3.5 Haiku

Fast and efficient model from Anthropic's Claude 3.5 series via OpenRouter. Designed for quick, straightforward tasks.

View details

MoonshotAI Kimi K2 Thinking

MoonshotAI's Kimi K2 with thinking mode via OpenRouter. Trillion-parameter MoE with 32B active parameters supporting persistent multi-turn reasoning. Capable of 200-300+ tool calls with step-by-step thought process.

View details

Perplexity Sonar Pro Search

Perplexity AI's Sonar Pro with advanced search capabilities via OpenRouter. Features autonomous multi-step reasoning for planning and executing entire research workflows. Combines web search with deep reasoning for complex research tasks.

View details

Claude 3.5 Sonnet

Anthropic model that raises the industry bar for intelligence, available via OpenRouter. Outperforms competitor models with the speed and cost of mid-tier models.

View details

Command R+

Cohere's premium enterprise model via OpenRouter. Optimized for retrieval-augmented generation (RAG), document processing, and business applications with exceptional instruction-following.

View details

Command R

Cohere's balanced enterprise model via OpenRouter. Provides strong performance for RAG and document tasks with efficient resource usage and reliable instruction-following.

View details

GPT-4o

OpenAI's versatile, high-intelligence flagship model via OpenRouter. Accepts text and image inputs with balanced performance across diverse tasks.

View details

Llama 3.1 405B

Meta's largest open-source Llama 3.1 model via OpenRouter with 405B parameters. Delivers frontier-level performance with instruction tuning for complex reasoning and generation tasks.

View details

Llama 3.1 70B

Meta's high-performance Llama 3.1 model via OpenRouter with 70B parameters. Balances quality and efficiency with instruction tuning for diverse applications.

View details

Llama 3.1 8B

Meta's ultra-efficient Llama 3.1 model via OpenRouter with 8B parameters. Optimized for ultra-fast inference and minimal costs while maintaining reliable performance.

View details

Mistral Nemo

Mistral AI's efficient smaller model via OpenRouter. Provides fast inference and lower costs while maintaining good performance for everyday tasks.

View details

Mistral 7B

Mistral AI's ultra-efficient 7B parameter model via OpenRouter. Optimized for speed and cost-effectiveness while delivering reliable performance for straightforward tasks.

View details

Mixtral 8x22B

Mistral AI's mixture-of-experts model via OpenRouter with 8 experts and 22B parameters each. Delivers excellent quality across diverse tasks with efficient parameter usage.

View details

Mistral Large

Mistral AI's flagship model via OpenRouter. Delivers strong performance across reasoning, coding, and multilingual tasks. Provides a compelling European alternative to frontier models.

View details

Perplexity Sonar Reasoning Pro

Perplexity AI's reasoning-enhanced Sonar model via OpenRouter. Combines deep reasoning capabilities with web search integration for complex research and analysis tasks.

View details