AllModel
Sign inSign up

AllModelModel catalogue

Model catalogue

112models13vendors

OpenAI

18

GPT 5.6 Sol

Latest frontier agentic coding model — the new Codex default.

Text modelcontext 372K
$0.15/ 1M×3

GPT 5.5

OpenAI's flagship for coding, agents and research-grade reasoning.

Text modelcontext 400K
$0.10/ 1M×2

GPT 5.6 Luna

The fast, light GPT-5.6 — cheap agents and high-volume coding.

Text modelcontext 272K
$0.05/ 1M×1

GPT 5.6 Terra

Balanced agentic coding model for everyday work.

Text modelcontext 272K
$0.10/ 1M×2

GPT 5.4

Production workhorse — near-flagship quality at half the price.

Text modelcontext 400K
$0.075/ 1M×1.5

GPT 5.4 (mini)

The strongest mini — cheap sub-agents and high-volume tasks.

Text modelcontext 400K
$0.05/ 1M×1

GPT OSS 120b (medium)

OpenAI's open 120B model at medium reasoning depth

Text modelcontext 128K
$0.025/ 1M×0.5

GPT-OSS 20B

OpenAI's smaller open model — high-volume work at no charge

Text modelcontext 131K
free

DALL·E 3

OpenAI's previous image generation — follows the prompt closely

Image generation
$0.052/ image×10.4

GPT 4o Mini TTS

OpenAI's cheap speech — bot replies and notifications

Text to speech
$0.0025/ 1K ch.×1

GPT Image 1

OpenAI images: generation and edits from a description

Image generation
$0.026/ image×5.2

GPT Image 1.5

OpenAI images, a generation newer than the first

Image generation
$0.0037/ image×0.75

GPT Image 2

OpenAI's latest images — text in frame and prompt-based edits

Image generation
$0.005/ image×1

Text Embedding 3 Large

OpenAI vectors for semantic search — the larger of the two

Embeddings
$0.05/ 1M×1

Text Embedding 3 Small

OpenAI vectors at a lower rate — big corpora and frequent reindexing

Embeddings
$0.05/ 1M×1

Text Embedding Ada 002

Legacy OpenAI vectors — compatibility with older indexes

Embeddings
$0.05/ 1M×1

TTS 1

Fast OpenAI speech — streaming with little delay

Text to speech
$0.0025/ 1K ch.×1

TTS 1 HD

A step up in OpenAI speech — recordings and voice-over

Text to speech
$0.0025/ 1K ch.×1
More10

Anthropic

10
More2

Moonshot

2

Zhipu

6

xAI

11
More3

DeepSeek

2

Google

27

Gemini 3.5 Flash (high)

Google's newest default — frontier smarts at a Flash price.

Text modelcontext 1M
$0.03/ 1M×0.6

Gemini 3.1 Flash Lite

The fastest, cheapest Gemini of the 3rd series.

Text modelcontext 1M
$0.03/ 1M×0.6down

Gemma 4 31B IT

Google's open-weight model — near-free tokens for bulk work.

Text modelcontext 131K
freeunstable

Gemini 3 Flash

Fast Gemini 3 — cheap request volume and simple agents

Text modelcontext 1M
$0.03/ 1M×0.6

Gemini 3.6 Flash (high)

Gemini 3.6 Flash at high reasoning — multi-step problems

Text modelcontext 1M
$0.03/ 1M×0.6

Gemini 3.7 Flash (high)

Gemini 3.7 Flash at high reasoning — harder analysis

Text modelcontext 1M
$0.03/ 1M×0.6

Gemini Pro Agent

Gemini Pro for agents — long tool-using runs

Text modelcontext 1M
$0.085/ 1M×1.7

DiffusionGemma 26B A4B IT

An open 26B Gemma built on diffusion

Text modelcontext 131K
free

Gemini 2.5 Flash

Gemini 2.5 Flash — earlier generation, cheap everyday text

Text modelcontext 1M
$0.03/ 1M×0.6

Gemini 2.5 Flash Lite

The lightest Gemini 2.5 — simple tasks at volume

Text modelcontext 1M
$0.03/ 1M×0.6down

Gemini 2.5 Pro

Gemini 2.5 Pro — long context and document analysis

Text modelcontext 1M
$0.085/ 1M×1.7

Gemini 3 Flash Agent

The same Gemini 3 Flash tuned for tool-using agent runs

Text modelcontext 1M
$0.03/ 1M×0.6

Gemini 3.1 Flash Image

Gemini images — generate and edit a frame from a description

Image generation
$0.005/ image×1

Gemini 3.1 Pro (low)

Gemini 3.1 Pro at low reasoning — cheaper than the full run

Text modelcontext 1M
$0.085/ 1M×1.7

Gemini 3.5 Flash (extra low)

Gemini 3.5 Flash at minimal reasoning — the cheapest run

Text modelcontext 1M
$0.03/ 1M×0.6

Gemini 3.5 Flash (low)

Gemini 3.5 Flash at low reasoning — quick answers

Text modelcontext 1M
$0.03/ 1M×0.6

Gemini 3.6 Flash (low)

Gemini 3.6 Flash, low reasoning — request volume

Text modelcontext 1M
$0.03/ 1M×0.6

Gemini 3.6 Flash (medium)

Gemini 3.6 Flash at medium reasoning — the working middle

Text modelcontext 1M
$0.03/ 1M×0.6

Gemini 3.7 Flash (low)

Latest Gemini 3.7 Flash, low reasoning — cheap volume

Text modelcontext 1M
$0.03/ 1M×0.6

Gemini 3.7 Flash (medium)

Gemini 3.7 Flash at medium reasoning — everyday work

Text modelcontext 1M
$0.03/ 1M×0.6

Imagen 3.0 Generate 002

Imagen 3 — Google's photoreal generation

Image generation
$0.052/ image×10.4

Nano Banana

Google's prompt-based image editing — fast and cheap

Image generation
$0.0075/ image×1.5

Nano Banana Lite

Lighter image editing — drafts and bulk runs

Image generation
$0.0037/ image×0.75

Nano Banana Pro

Google's top image editing — clean detail and text

Image generation
$0.015/ image×3

Omni Video

Google video — a scene generated from a description

Video generation
by coefficient×0.15

Veo 3.1

Veo 3.1 — video from a description, with sound

Video generation
by coefficient×0.75

Veo 3.1 Lite

Lighter Veo — draft clips at lower cost

Video generation
by coefficient×0.1
More19

MiniMax

1

NVIDIA

19

Nemotron 3 Ultra

NVIDIA's open reasoning flagship, available free through AnyModel.

Text modelcontext 262K
free

Llama 3.1 Nemotron Nano VL 8B v1

A small vision Llama from NVIDIA — images and documents

Text modelcontext 131K
free

Mistral Nemotron

Mistral, tuned by NVIDIA — everyday text and tools

Text modelcontext 131K
free

Nemotron 3 Nano 30B A3B

Nemotron 3 Nano — a light 30B model at no charge

Text modelcontext 262K
free

Nemotron 3 Nano Omni 30B A3B Reasoning

Nemotron Nano Omni with reasoning — text, images and audio together

Text modelcontext 262K
free

Nemotron 3 Super 120B A12B

The top 120B Nemotron — the line's heaviest tasks

Text modelcontext 262K
free

Nemotron 3.5 Content Safety

NVIDIA content safety — flags disallowed material in text

Text modelcontext 33K
free

Nemotron 3.5 Lightning 30B A3B

The line's fastest Nemotron — short answers at speed

Text modelcontext 262K
free

Nemotron Nano 12B v2 VL

A 12B vision Nemotron — screenshots and tables

Text modelcontext 131K
free

NVIDIA Nemotron Nano 9B v2

The lightest 9B Nemotron — bulk runs at no charge

Text modelcontext 131K
free

Riva Translate 4B Instruct v2

NVIDIA Riva translator — streaming text translation

Text modelcontext 33K
free

Llama Nemotron Embed 1B v2

NVIDIA vectors built on Llama — search over your own corpus

Embeddings
free

Llama Nemotron Embed VL 1b V2

NVIDIA vision vectors on the gateway's second rate

Embeddings
$0.005/ 1M×0.1

Llama Nemotron Embed VL 1B v2

NVIDIA vision vectors — search across images and documents

Embeddings
free

Nemotron 3 Embed 1B

1B Nemotron vectors — text indexing at no charge

Embeddings
free

Nemotron 3 Nano Omni 30B A3B Reasoning (Speech to Text)

Speech transcription on Nemotron Omni — audio to text

Speech to text
free

NV Embed v1

NVIDIA's classic vectors — general semantic search

Embeddings
free

NV EmbedCode 7B v1

NVIDIA vectors for code — search across a repository

Embeddings
free

NV EmbedQA E5 v5

NVIDIA vectors for question answering — retrieval for RAG

Embeddings
free
More11

Alibaba

2

Deepgram

4

Black Forest Labs

3

Perplexity

2

Other

5