AllModelModel catalogue
Model catalogue
112models13vendors
OpenAI
18GPT 5.6 Sol
Latest frontier agentic coding model — the new Codex default.
GPT 5.5
OpenAI's flagship for coding, agents and research-grade reasoning.
GPT 5.6 Luna
The fast, light GPT-5.6 — cheap agents and high-volume coding.
GPT 5.6 Terra
Balanced agentic coding model for everyday work.
GPT 5.4
Production workhorse — near-flagship quality at half the price.
GPT 5.4 (mini)
The strongest mini — cheap sub-agents and high-volume tasks.
GPT OSS 120b (medium)
OpenAI's open 120B model at medium reasoning depth
GPT-OSS 20B
OpenAI's smaller open model — high-volume work at no charge
DALL·E 3
OpenAI's previous image generation — follows the prompt closely
GPT 4o Mini TTS
OpenAI's cheap speech — bot replies and notifications
GPT Image 1
OpenAI images: generation and edits from a description
GPT Image 1.5
OpenAI images, a generation newer than the first
GPT Image 2
OpenAI's latest images — text in frame and prompt-based edits
Text Embedding 3 Large
OpenAI vectors for semantic search — the larger of the two
Text Embedding 3 Small
OpenAI vectors at a lower rate — big corpora and frequent reindexing
Text Embedding Ada 002
Legacy OpenAI vectors — compatibility with older indexes
TTS 1
Fast OpenAI speech — streaming with little delay
TTS 1 HD
A step up in OpenAI speech — recordings and voice-over
Anthropic
10Claude Opus 5
The newest experimental Opus, served through Kiro.
Claude Opus 4.8
The current Opus — Anthropic's agentic-coding flagship.
Claude Sonnet 5
The new mid-tier Claude 5 — intro-priced by the vendor until September.
Claude Opus 4.7
Previous Opus — between 4.8 and 4.6 in the metering ladder.
Claude Opus 4.6
Previous-generation Opus — flagship quality, smaller multiplier.
Claude Sonnet 4.6
The balanced Claude — fast, smart, affordable.
Claude Haiku 4.5
Cheapest, fastest Claude for high-volume tasks.
Claude Fable 5
Anthropic's Mythos-class tier above Opus — #1 on our AI Arena.
Claude Opus 4.5
Previous-generation flagship Claude — long analysis and code
Claude Sonnet 4.5
Previous-generation workhorse Claude — balanced price and quality
Moonshot
2Zhipu
6GLM 5.2
Zhipu's MIT-licensed flagship — a coding-agent favourite.
GLM 4.6v
GLM with vision: images, screenshots and documents
GLM 4.7
Previous-generation GLM — inexpensive everyday text
GLM 5
Fifth-generation GLM — code, agents, long tasks
GLM 5.1
GLM 5.1 — the same generation, a step above the base five
GLM 5.3
The top GLM of the line — the family's heaviest tasks
xAI
11Grok 4.6
xAI's current frontier model — its own docs' default pick for code.
Grok 4.20 Multi Agent 0309
Grok's coordinated multi-agent mode for deep research.
Grok 4.20.0309 Reasoning
Grok with reasoning on — multi-step problems
Grok 4.3
Grok 4.3 — an earlier generation at a lower rate than the top models
Grok 4.5
Grok 4.5 — the line's general-purpose pick
Grok Build 0.1
Grok tuned for building code — project edits and agent runs
Grok 4.20.0309 Non Reasoning
The same Grok without reasoning — fast short answers
Grok Imagine Image
xAI images — fast generation from a description
Grok Imagine Image Quality
The same xAI images in quality mode — slower per frame
Grok Imagine Video
xAI video — short clips from a description
Grok Imagine Video 1.5
xAI video, a generation newer
DeepSeek
2Gemini 3.5 Flash (high)
Google's newest default — frontier smarts at a Flash price.
Gemini 3.1 Flash Lite
The fastest, cheapest Gemini of the 3rd series.
Gemma 4 31B IT
Google's open-weight model — near-free tokens for bulk work.
Gemini 3 Flash
Fast Gemini 3 — cheap request volume and simple agents
Gemini 3.6 Flash (high)
Gemini 3.6 Flash at high reasoning — multi-step problems
Gemini 3.7 Flash (high)
Gemini 3.7 Flash at high reasoning — harder analysis
Gemini Pro Agent
Gemini Pro for agents — long tool-using runs
DiffusionGemma 26B A4B IT
An open 26B Gemma built on diffusion
Gemini 2.5 Flash
Gemini 2.5 Flash — earlier generation, cheap everyday text
Gemini 2.5 Flash Lite
The lightest Gemini 2.5 — simple tasks at volume
Gemini 2.5 Pro
Gemini 2.5 Pro — long context and document analysis
Gemini 3 Flash Agent
The same Gemini 3 Flash tuned for tool-using agent runs
Gemini 3.1 Flash Image
Gemini images — generate and edit a frame from a description
Gemini 3.1 Pro (low)
Gemini 3.1 Pro at low reasoning — cheaper than the full run
Gemini 3.5 Flash (extra low)
Gemini 3.5 Flash at minimal reasoning — the cheapest run
Gemini 3.5 Flash (low)
Gemini 3.5 Flash at low reasoning — quick answers
Gemini 3.6 Flash (low)
Gemini 3.6 Flash, low reasoning — request volume
Gemini 3.6 Flash (medium)
Gemini 3.6 Flash at medium reasoning — the working middle
Gemini 3.7 Flash (low)
Latest Gemini 3.7 Flash, low reasoning — cheap volume
Gemini 3.7 Flash (medium)
Gemini 3.7 Flash at medium reasoning — everyday work
Imagen 3.0 Generate 002
Imagen 3 — Google's photoreal generation
Nano Banana
Google's prompt-based image editing — fast and cheap
Nano Banana Lite
Lighter image editing — drafts and bulk runs
Nano Banana Pro
Google's top image editing — clean detail and text
Omni Video
Google video — a scene generated from a description
Veo 3.1
Veo 3.1 — video from a description, with sound
Veo 3.1 Lite
Lighter Veo — draft clips at lower cost
MiniMax
1NVIDIA
19Nemotron 3 Ultra
NVIDIA's open reasoning flagship, available free through AnyModel.
Llama 3.1 Nemotron Nano VL 8B v1
A small vision Llama from NVIDIA — images and documents
Mistral Nemotron
Mistral, tuned by NVIDIA — everyday text and tools
Nemotron 3 Nano 30B A3B
Nemotron 3 Nano — a light 30B model at no charge
Nemotron 3 Nano Omni 30B A3B Reasoning
Nemotron Nano Omni with reasoning — text, images and audio together
Nemotron 3 Super 120B A12B
The top 120B Nemotron — the line's heaviest tasks
Nemotron 3.5 Content Safety
NVIDIA content safety — flags disallowed material in text
Nemotron 3.5 Lightning 30B A3B
The line's fastest Nemotron — short answers at speed
Nemotron Nano 12B v2 VL
A 12B vision Nemotron — screenshots and tables
NVIDIA Nemotron Nano 9B v2
The lightest 9B Nemotron — bulk runs at no charge
Riva Translate 4B Instruct v2
NVIDIA Riva translator — streaming text translation
Llama Nemotron Embed 1B v2
NVIDIA vectors built on Llama — search over your own corpus
Llama Nemotron Embed VL 1b V2
NVIDIA vision vectors on the gateway's second rate
Llama Nemotron Embed VL 1B v2
NVIDIA vision vectors — search across images and documents
Nemotron 3 Embed 1B
1B Nemotron vectors — text indexing at no charge
Nemotron 3 Nano Omni 30B A3B Reasoning (Speech to Text)
Speech transcription on Nemotron Omni — audio to text
NV Embed v1
NVIDIA's classic vectors — general semantic search
NV EmbedCode 7B v1
NVIDIA vectors for code — search across a repository
NV EmbedQA E5 v5
NVIDIA vectors for question answering — retrieval for RAG
Alibaba
2Deepgram
4Nova
Deepgram transcription — streaming speech to text
Nova 2
Nova 2 — a generation newer than the first
Nova 3
Third-generation Nova — Deepgram's current transcription
Whisper Large
Whisper Large via Deepgram — multilingual transcription
Black Forest Labs
3Perplexity
2Other
5Free Models Auto Router
Automatically selects a currently available free model.
Inkling
A small gateway text model — simple tasks at no charge
Laguna XS 2.1
A compact Laguna model — short answers at no charge
Llama 3.2 11B Vision Instruct
11B vision Llama 3.2 — an open model for images
Step 3.7 Flash
Step 3.7 Flash — a fast model for request volume