Skip to main content

Model Hub

Discover the best AI models for code migration, analysis, and generation. Compare capabilities, pricing, and find the right model for your project.

Filter by provider:
StepFun
2026-10-08

Step-5 Preview

A hosted preview foundation model available via OpenRouter with a 1,000,000-token context window. The provided discovery data verifies it as a newly added long-context model.

text-generationlong-contextreasoning
Context
1.0M
License
Open
Anthropic
2026-10-07

Claude Haiku 5.5

The fastest and most efficient model in Anthropic's Claude 5.5 family, positioned for subagents and high-volume, cost-sensitive workloads. It is available through Amazon Bedrock, Claude Platform on AWS, and OpenRouter with a 1,000,000-token context window.

text-generationreasoningcode-generation+2
Context
1.0M
License
Open
OpenAI
2026-10-07

GPT-6

OpenAI's GPT-6 model is rolling out globally in ChatGPT with Intelligent UI, enabling faster responses with visuals and interactive experiences. OpenAI also published a model guide for the GPT-6 family during the tracked date range.

text-generationreasoningmultimodal+2
Context
1.1M
License
Open
Google
2026-10-06

Gemini Nano Banana 2.1

Google model newly added to OpenRouter with a 65,536-token context window. The provided discovery data verifies its provider, OpenRouter model ID, context length, and add date.

text-generationlong-context
Context
66K
License
Open
Mistral AI
2026-10-06

Mistral Large 4.0

Mistral AI large foundation model newly added to OpenRouter with a 524,288-token context window. It also appears in the provided Ollama library list as an open-weight model runnable locally.

text-generationlong-context
Context
524K
License
Open
InclusionAI
2026-10-02

Ling 3.1 Flash

Ling 3.1 Flash is a newly added hosted long-context model on OpenRouter, positioned as a fast Flash variant with a 262k-token context window.

text-generationreasoninglong-context+1
Context
262K
Input/1M
$0
OpenAI
2026-10-01

GPT-6 Astra Ultrafast

Fast-serving variant of GPT-6 Astra announced as available in the OpenAI API and for eligible ChatGPT Work and Codex users. NVIDIA reports it runs on Blackwell GPUs and offers up to 8x faster inference.

reasoningcode-generationcomputer-use+1
Context
N/A
License
Open
Apodex
2026-10-01

Apodex 1.1 Mini

OpenRouter-listed hosted model added during the check window with a 262K-token context window. The discovery data identifies it as a free Apodex 1.1 Mini endpoint.

text-generationlong-context
Context
262K
License
Open
Unbiased
2026-10-01

Pareto 26.10 Preview

OpenRouter-listed preview model from Unbiased added during the check window. The discovery data reports a 1,048,576-token context window.

text-generationlong-context
Context
1.0M
Input/1M
$0.8
OpenAI
2026-09-29

GPT-6.1 Sol

OpenAI model announced as bringing near-Astra intelligence to coding, computer use, and professional work at lower token prices than Astra. It is listed with a 1,050,000-token context window.

reasoningcode-generationcomputer-use+2
Context
1.1M
Input/1M
$2
OpenAI
2026-09-29

GPT-6.1 Sol Pro

Pro variant of OpenAI's GPT-6.1 Sol line, listed on OpenRouter with a 1,050,000-token context window. It targets demanding coding, computer-use, and professional work scenarios.

reasoningcode-generationcomputer-use+2
Context
1.1M
Input/1M
$2
Anthropic
2026-09-28

Claude Sonnet 5.5

Anthropic's Claude Sonnet 5.5 is described as a smarter, more efficient Sonnet model for focused coding and knowledge work. It is listed with a 1,000,000-token context window.

reasoningcode-generationlong-context+1
Context
1.0M
Input/1M
$2
Perceptron
2026-09-25

Perceptron MK1.5

OpenRouter-listed hosted model added during the check window. The discovery data reports a 36,864-token context window.

text-generation
Context
37K
Input/1M
$0.15
Alibaba
2026-09-25

Qwen3-VL-8B

Qwen3-VL-8B is a Qwen3 vision-language model referenced for multimodal reinforcement-learning post-training with GRPO on Amazon SageMaker HyperPod.

visionmultimodalimage-understanding+1
Context
N/A
License
Open
Alibaba
2026-09-25

Qwen3-TTS-12Hz-1.7B-Base

Qwen3-TTS-12Hz-1.7B-Base is a publicly available text-to-speech model for real-time personalized speech generation and cross-lingual voice cloning via Amazon SageMaker JumpStart.

text-to-speechvoice-cloningmultilingual+1
Context
N/A
License
Open
Fireworks AI
2026-09-24

Ember 1

Ember 1 is a long-context hosted foundation model added to OpenRouter with a 1,048,576-token context window. The discovery data verifies it as a newly added Fireworks model during the target week.

text-generationreasoningcode-generation+1
Context
1.0M
Input/1M
$3
Zhipu AI
2026-09-23

GLM-5.3 Prime

GLM-5.3 Prime is a Zhipu AI / Z.ai hosted foundation model added to OpenRouter with a 1,000,000-token context window. It appears to be a higher-tier Prime variant of the GLM-5.3 model family.

text-generationreasoningcode-generation+1
Context
1.0M
Input/1M
$2.8
Alibaba
2026-09-23

Qwen3.8 Max Prime

Qwen3.8 Max Prime is an Alibaba Qwen hosted foundation model added to OpenRouter with a 1,000,000-token context window. It is a Prime variant of the Qwen3.8 Max line for long-context general reasoning and generation.

text-generationreasoningcode-generation+2
Context
1.0M
Input/1M
$4
Aion Labs
2026-09-23

Aion 3.5

Aion 3.5 is a hosted foundation model added to OpenRouter with a 262,144-token context window. The discovery data identifies it as a new Aion Labs model in the target week.

text-generationreasoningcode-generation+1
Context
262K
Input/1M
$3
Aion Labs
2026-09-23

Aion 3.5 Mini

Aion 3.5 Mini is a smaller hosted variant in the Aion 3.5 family added to OpenRouter with a 262,144-token context window. It is positioned as a more compact model for general text and coding workflows.

text-generationreasoningcode-generation+1
Context
262K
Input/1M
$0.7
Upstage
2026-09-23

Solar Mini4

Solar Mini4 is an Upstage hosted foundation model added to OpenRouter with a 524,288-token context window. It is a compact Solar-family model for long-context text generation and reasoning.

text-generationreasoningcode-generation+1
Context
524K
Input/1M
$0.05
OpenAI
2026-09-22

GPT-6 Sol

GPT-6 Sol is an OpenAI frontier model announced alongside GPT-6 Luna, aimed at bringing high intelligence to everyday work with a different balance of capability and cost.

text-generationreasoninglong-context+1
Context
1.1M
Input/1M
$2
OpenAI
2026-09-22

GPT-6 Sol Pro

GPT-6 Sol Pro is the higher-capability hosted variant of GPT-6 Sol listed on OpenRouter, with a 1.05M-token context window for demanding long-context workloads.

text-generationreasoninglong-context+1
Context
1.1M
Input/1M
$2
OpenAI
2026-09-22

GPT-6 Luna

GPT-6 Luna is an OpenAI frontier model announced with GPT-6 Sol, positioned as a new option for everyday work with a distinct capability-cost balance.

text-generationreasoninglong-context+1
Context
1.1M
Input/1M
$0.1
OpenAI
2026-09-22

GPT-6 Luna Pro

GPT-6 Luna Pro is the higher-capability hosted variant of GPT-6 Luna listed on OpenRouter, offering a 1.05M-token context window for large-scale tasks.

text-generationreasoninglong-context+1
Context
1.1M
Input/1M
$0.1
Anthropic
2026-09-22

Claude Opus 5.5

Claude Opus 5.5 is Anthropic's most capable Opus model for agentic coding, knowledge work, and long-running tasks, now available through Amazon Bedrock and OpenRouter.

text-generationreasoningcode-generation+2
Context
1.0M
Input/1M
$4
Cohere
2026-09-22

Command A Plus

Command A Plus is a Cohere hosted language model added to OpenRouter with a 192K-token context window for long-context enterprise workloads.

text-generationreasoninglong-context
Context
192K
Input/1M
$0.3
xAI
2026-09-21

Grok 4.7

Grok 4.7 is an xAI hosted frontier model listed on OpenRouter with a 500K-token context window for long-context reasoning and general-purpose tasks.

text-generationreasoningcode-generation+1
Context
500K
Input/1M
$2
Alibaba
2026-09-21

Qwen3.8 Omni Flash

Qwen3.8 Omni Flash is an Alibaba Qwen model listed on OpenRouter with a 1M-token context window, positioned as a fast omni-capable model in the Qwen3.8 family.

text-generationmultimodalreasoning+1
Context
1.0M
Input/1M
$0.15
Xiaomi
2026-09-21

MiMo v2.6 Flash

MiMo v2.6 Flash is a Xiaomi model listed on OpenRouter with a 1,048,576-token context window, indicating a fast long-context variant of the MiMo v2.6 family.

text-generationreasoninglong-context
Context
1.0M
Input/1M
$0.14
Xiaomi
2026-09-21

MiMo v2.6 Pro

MiMo v2.6 Pro is a Xiaomi model listed on OpenRouter with a 1,048,576-token context window, representing the Pro tier of the MiMo v2.6 model family.

text-generationreasoninglong-context
Context
1.0M
Input/1M
$0.435
Xiaomi
2026-09-21

MiMo v2.6 Pro Ultraspeed

MiMo v2.6 Pro Ultraspeed is a Xiaomi model listed on OpenRouter with a 1,048,576-token context window, optimized as a faster Pro-tier MiMo v2.6 variant.

text-generationreasoninglong-context
Context
1.0M
Input/1M
$4.35
Prism ML
2026-09-18

Ternary Bonsai 2 27B

Ternary Bonsai 2 27B is a 27B-parameter model added to OpenRouter with a 262,144-token context window. The discovery data identifies it as a new Prism ML model during the target date range.

text-generationreasoningcode-generation+1
Context
262K
Input/1M
$0.075
Zhipu AI
2026-09-18

GLM-5.3-FlashX

GLM-5.3-FlashX is a Zhipu AI / Z.ai GLM-family hosted language model added to OpenRouter with a 1,048,576-token context window. It is suited for long-context text generation and reasoning workflows.

text-generationreasoninglong-context
Context
1.0M
Input/1M
$0.37
NVIDIA
2026-09-15

NVIDIA Nemotron 3 Super

NVIDIA Nemotron 3 Super is a foundation model referenced as the base for Salesforce's Koa CRM reasoning model and is available in the Ollama library for local use.

text-generationreasoninglocal-inference
Context
262K
Input/1M
$0.08
DeepSeek
2026-09-14

DeepSeek Pro Latest

A hosted DeepSeek Pro-class model endpoint added to OpenRouter with a 1,048,576-token context window. It is intended for long-context text generation and general-purpose assistant workloads.

text-generationlong-context
Context
1.0M
Input/1M
$0.1901
DeepSeek
2026-09-14

DeepSeek Flash Latest

A hosted DeepSeek Flash-class model endpoint added to OpenRouter with a 1,048,576-token context window. It targets long-context text generation in a faster Flash-tier model line.

text-generationlong-context
Context
1.0M
Input/1M
$0.003
Inference Net
2026-09-12

Schematron v2 Turbo

A 128k-context hosted language model in the Schematron v2 family, added to OpenRouter on 2026-09-12. The available data verifies its model ID and context length but does not provide benchmark, modality, or pricing details.

text-generationlong-context
Context
128K
Input/1M
$0.03
Inference Net
2026-09-12

Schematron v2 Small

A 128k-context hosted language model in the Schematron v2 family, added to OpenRouter on 2026-09-12. The available data verifies its model ID and context length but does not provide benchmark, modality, or pricing details.

text-generationlong-context
Context
128K
Input/1M
$0.05
Sakana AI
2026-09-11

Fugu Ultra v2

A Sakana AI hosted foundation model with a 1,000,000-token context window, added to OpenRouter on 2026-09-11. It appears to be a new v2 update in the Fugu Ultra line, with available data verifying context length but not pricing or benchmark details.

text-generationlong-context
Context
1.0M
Input/1M
$5
Sakana AI
2026-09-11

Fugu Max

A Sakana AI hosted foundation model with a 1,000,000-token context window, added to OpenRouter on 2026-09-11. The OpenRouter data verifies it as a new Fugu-family model with very long context support, but does not provide pricing or benchmark details.

text-generationlong-context
Context
1.0M
Input/1M
$2
OpenAI
2026-09-10

GPT-Live-1

GPT-Live-1 is OpenAI's real-time voice model for natural, full-duplex conversations in the API. It emphasizes stronger instruction following, custom voices, and telephony-oriented voice experiences.

real-time-audiospeech-to-speechinstruction-following+1
Context
N/A
License
Open
DeepSeek
2026-09-10

DeepSeek V4.1 Flash

DeepSeek V4.1 Flash is a long-context foundation model from DeepSeek, newly listed on OpenRouter and also present in the Ollama library. It is positioned as a fast general-purpose model with a 1,048,576-token context window.

text-generationreasoninglong-context
Context
1.0M
Input/1M
$0.3
InclusionAI
2026-09-10

Ling 3.0 Flash VL

Ling 3.0 Flash VL is a vision-language variant of InclusionAI's Ling 3.0 Flash family, newly added to OpenRouter. It supports multimodal use cases with a 262,144-token context window.

visionmultimodaltext-generation+1
Context
262K
Input/1M
$0.021
TwelveLabs
2026-09-10

Marengo Embed 3.0

Marengo Embed 3.0 is TwelveLabs' multimodal embedding model for semantic search over video, image, and audio content. It became generally available in Amazon Bedrock Knowledge Bases during the discovery window.

embeddingssemantic-searchvideo-understanding+2
Context
N/A
License
Open
Skild AI
2026-09-10

Skild S1

Skild S1 is a robot foundation model designed to teach robots previously unseen, long-horizon tasks from a single video. It was announced through NVIDIA's Physical AI ecosystem coverage.

roboticsvisionimitation-learning+1
Context
N/A
License
Open
IBM
2026-09-09

Granite Time Series PatchTST-FM-r2

Granite Time Series PatchTST-FM-r2 is IBM's time-series foundation model released with a commercial-friendly license. It targets forecasting and time-series analysis workloads.

time-series-forecastingtime-series-analysis
Context
N/A
License
commercial-friendly license
Open Weight
OpenAI
2026-09-08

ChatGPT Images 2.5

OpenAI image-generation and image-editing model for turning ideas, sketches, and reference photos into more personalized, polished images.

image-generationimage-editingmultimodal
Context
N/A
License
Open
Inception
2026-09-08

Mercury 2.5

Hosted language model added to OpenRouter with a 260k-token context window.

text-generationlong-context
Context
260K
Input/1M
$0.04
NEX AGI
2026-09-08

NEX N2.5 Pro

Free hosted NEX AGI language model added to OpenRouter with a 262,144-token context window.

text-generationlong-context
Context
262K
Input/1M
$0.075
NEX AGI
2026-09-08

NEX N2.5 Mini

Free hosted compact NEX AGI language model added to OpenRouter with a 262,144-token context window.

text-generationlong-context
Context
262K
Input/1M
$0.025
Pathway
2026-09-08

Baby Dragon Hatchling

Pathway's Baby Dragon Hatchling is a brain-inspired, post-transformer architecture that reasons in latent space instead of emitting chain-of-thought tokens.

reasoninglatent-space-reasoning
Context
N/A
License
Open
NVIDIA
2026-09-04

NVIDIA Cosmos 3

NVIDIA Cosmos 3 is used in Physical AI model-factory pipelines for synthetic data generation, post-training, and closed-loop evaluation.

physical-aisynthetic-data-generationmodel-evaluation
Context
N/A
License
Open
OpenAI
2026-09-04

GPT-6 Astra Pro

GPT-6 Astra Pro is a higher-tier GPT-6 Astra model variant listed on OpenRouter with a 1,050,000-token context window. It appears intended for demanding hosted frontier-model workloads requiring long-context reasoning and advanced agentic capabilities.

reasoningcode-generationcomputer-use+2
Context
1.1M
Input/1M
$10
OpenAI
2026-09-03

GPT-6 Astra

OpenAI's GPT-6 Astra is described as its most intelligent and aligned broadly deployed model, with state-of-the-art capabilities across computer use, coding, cybersecurity, and science. It is the first OpenAI model in the provided data to reach the Critical cybersecurity capability threshold under the Preparedness Framework.

reasoningcode-generationcomputer-use+2
Context
1.1M
Input/1M
$10
Meta
2026-09-02

Muse Spark 1.3

Muse Spark 1.3 is a Meta hosted foundation model newly added to OpenRouter with a 1,048,576-token context window. It appears to be the next base release in the Muse Spark series, suited for long-context general-purpose generation.

text-generationreasoninglong-context
Context
1.0M
Input/1M
$1.25
Google
2026-09-02

Gemini 3.8 Flash

Gemini 3.8 Flash is a Google Gemini Flash-series model newly added to OpenRouter with a 1,048,576-token context window. It is positioned as a fast, large-context general-purpose model.

text-generationreasoninglong-context
Context
1.0M
Input/1M
$0.75
OpenAI
2026-09-01

Astra

Astra is an OpenAI frontier model announced as the first OpenAI model to meet the Critical cybersecurity capability threshold under the Preparedness Framework. The announcement emphasizes stronger safeguards for release.

reasoningcybersecuritysafety
Context
N/A
License
Open
Anthropic
2026-09-01

Claude Fable 5.1

Claude Fable 5.1 is a new Anthropic Claude model available on Amazon Bedrock, Claude Platform on AWS, and OpenRouter. The release highlights model improvements and Enterprise Frontier Safeguards for controlled cloud deployments.

text-generationreasoninglong-context+1
Context
1.0M
Input/1M
$10
Inception
2026-08-31

Mercury 2.5 Preview

Mercury 2.5 Preview is a hosted preview model added to OpenRouter under the inception namespace. It provides a large 260k-token context window for long-context text tasks.

text-generationreasoninglong-context
Context
260K
Input/1M
$0.04
IBM
2026-08-31

Granite 4.2 8B

Granite 4.2 8B is an IBM Granite open-weight foundation model added to OpenRouter with a 131k-token context window. It is also present in the Ollama library as granite4.2 for local deployment.

text-generationreasoningcode-generation+1
Context
131K
Input/1M
$0.06
Tencent
2026-08-28

HY4 Preview

Tencent HY4 Preview is a newly added hosted foundation model on OpenRouter with a 1,048,576-token context window. It appears to target general long-context chat and text generation workloads.

text-generationchatlong-context
Context
1.0M
Input/1M
$0.834
InclusionAI
2026-08-27

Ling 3.0 Flash Fin

A newly added hosted Ling 3.0 Flash financial-domain model on OpenRouter with a 262,144-token context window. It appears intended for long-context finance-oriented text generation and analysis tasks.

text-generationreasoninglong-context+1
Context
262K
Input/1M
$0.042
Alibaba
2026-08-26

Qwen3.8 Flash

Qwen3.8 Flash is an Alibaba Qwen model newly added on OpenRouter with a 1,000,000-token context window. It is positioned as a fast, long-context general-purpose model.

text-generationchatlong-context
Context
1.0M
Input/1M
$0.15
Zhipu AI
2026-08-26

GLM-5.3-Flash

GLM-5.3-Flash is a Zhipu AI / Z.ai GLM model newly added on OpenRouter with a 1,310,720-token context window. It is also present in the Ollama library, indicating local open-weight availability.

text-generationchatlong-context
Context
1.3M
Input/1M
$0.15
DeepSeek
2026-08-21

DeepSeek V4 Flash Vision Exp

Experimental vision-capable variant of DeepSeek V4 Flash available on OpenRouter, adding multimodal image understanding to the Flash model line with a 1M-token context window.

text-generationvisionmultimodal+1
Context
1.0M
Input/1M
$0.2156
Tencent
2026-08-20

HY-MT2-1.8B

Tencent HY-MT2-1.8B is a newly listed hosted text model on OpenRouter with an 8K-token context window. It appears to be part of Tencent's HY-MT2 model family.

text-generationtranslation
Context
8K
Input/1M
$0.044
Tencent
2026-08-20

HY-MT2-30B-A3B

Tencent HY-MT2-30B-A3B is a newly listed hosted text model on OpenRouter with an 8K-token context window. It is the larger HY-MT2 variant surfaced in the discovery data.

text-generationtranslation
Context
8K
Input/1M
$0.074
Tencent
2026-08-19

HY-MT2 7B

Tencent's 7B-parameter HY-MT2 model made available on OpenRouter with an 8K-token context window. It appears to be part of Tencent's HY-MT2 model family, positioned for multilingual text and translation-style workloads.

text-generationtranslationmultilingual
Context
8K
Input/1M
$0.074
Zhipu AI
2026-08-18

GLM-5.3

GLM-5.3 is a Zhipu AI / Z.ai foundation model newly added on OpenRouter. The provided data verifies a very large 1,048,576-token context window, making it suitable for long-context text tasks.

text-generationreasoninglong-context
Context
1.0M
Input/1M
$1.4
Alibaba
2026-08-14

Qwen3.8-27B

Qwen3.8-27B is a Qwen 3.8 family language model from Alibaba, listed with a 262k-token context window. It appears to be a smaller open-weight member of the Qwen3.8 lineup suitable for general-purpose long-context language tasks.

text-generationreasoninglong-context+1
Context
262K
Input/1M
$0.425
Dots Studio
2026-08-14

Dots 3 Note Preview

OpenRouter-listed free preview model from Dots Studio with a 512k-token context window. The discovery data identifies it as a recent preview model suitable for long-context text workflows.

text-generationlong-contextnote-taking
Context
512K
License
Open
Google
2026-08-13

Gemini 3.7 Flash

Google's Gemini 3.7 Flash is a newly listed hosted model with a 1,048,576-token context window. It appears positioned as a fast long-context Gemini model for general-purpose AI workloads.

text-generationreasoninglong-context
Context
1.0M
Input/1M
$0.75
ByteDance Seed
2026-08-12

Seed 2.1 Turbo

Seed 2.1 Turbo is a newly listed ByteDance Seed hosted model on OpenRouter with a 262,144-token context window. The Turbo designation indicates a speed-oriented variant for general-purpose long-context workloads.

text-generationreasoninglong-context
Context
262K
Input/1M
$0.5
Alibaba
2026-08-12

Qwen3.8-2.4T-A95B

Qwen3.8-2.4T-A95B is a newly listed Alibaba Qwen model with a 1,010,000-token context window. The name indicates a large mixture-style Qwen model variant intended for high-capacity long-context reasoning and generation.

text-generationreasoninglong-context
Context
1.0M
Input/1M
$2
ByteDance Seed
2026-08-12

Seed 2.0 Code

Seed 2.0 Code is a newly listed ByteDance Seed code-focused model with a 262,144-token context window. It is positioned for software engineering and code-generation workflows.

code-generationcode-analysisreasoning+1
Context
262K
Input/1M
$0.5
xAI
2026-08-12

Grok 4.6

Grok 4.6 is a newly listed xAI hosted model with a 500,000-token context window. It appears as a new Grok-series general-purpose reasoning model.

text-generationreasoninglong-context
Context
500K
Input/1M
$2
OpenAI
2026-08-11

Daybreak Red

Daybreak Red is an OpenAI specialized cyber defense model made available to eligible customers on Amazon Bedrock. The discovery data describes it as supporting authorized vulnerability research, exploit validation, and security testing workflows.

cybersecurityvulnerability-researchexploit-validation+2
Context
N/A
License
Open
OpenAI
2026-08-11

Daybreak Blue

Daybreak Blue is an OpenAI specialized cyber defense model made available to eligible customers on Amazon Bedrock. The discovery data identifies it as part of OpenAI and AWS's Daybreak cyber defense model offering with secure zero-operator-access deployment.

cybersecuritycyber-defensesecurity-analysis+1
Context
N/A
License
Open
Liquid AI
2026-08-11

LFM 2.5 2.6B

Liquid AI's LFM 2.5 2.6B is a compact long-context foundation model listed on OpenRouter with a 128k-token context window.

text-generationlong-contextefficient-inference
Context
128K
License
Open
NVIDIA
2026-08-11

NVIDIA Nemotron 3.5 Lightning

NVIDIA's Nemotron 3.5 Lightning expands the Nemotron 3 family as an efficient open model for long-running agentic AI workloads.

agentic-aitext-generationlong-context+2
Context
262K
Input/1M
$0.06
Sakana AI
2026-08-11

Sakana Namazu

Sakana Namazu is a Sakana AI model newly listed on OpenRouter with a 262k-token context window.

text-generationlong-context
Context
262K
Input/1M
$0.95
OpenAI
2026-08-10

GPT-5.6-Cyber

OpenAI cybersecurity-specific model available through Daybreak Red for authorized vulnerability research, exploit validation, and security testing.

cyber-securityvulnerability-researchexploit-validation+3
Context
N/A
License
Open
Upstage
2026-08-10

Solar Pro4

Upstage Solar Pro4 is a newly listed long-context hosted foundation model on OpenRouter, supporting a 524k-token context window.

text-generationlong-context
Context
524K
Input/1M
$0.09
Meta
2026-08-09

Muse Glimmer 30B

Meta's Muse Glimmer 30B is a 30B-parameter open-weight model newly listed on OpenRouter with a 131k-token context window and also present in the Ollama library.

text-generationlong-contextopen-weight
Context
131K
Input/1M
$0.35
InclusionAI
2026-08-06

Ling 3.0 Tiny

A smaller Ling 3.0-series hosted foundation model added to OpenRouter with a 262,144-token context window. The OpenRouter listing identifies it as a free-access model variant.

text-generationlong-context
Context
262K
License
Open
Meta
2026-08-05

Muse Spark 1.2

A Meta hosted foundation model added to OpenRouter with a 1,048,576-token context window. It is a newer Muse Spark release than the existing 1.1 model.

text-generationlong-context
Context
1.0M
Input/1M
$1.25
NVIDIA
2026-08-04

NVIDIA Alpamayo 2 Super

An open frontier model for robotaxis and autonomous vehicles, announced as available for commercial use. It targets long-tail autonomous-driving scenarios that require physical-world understanding beyond standard perception and motion prediction.

physical-aiworld-modelingautonomous-driving
Context
N/A
License
Open
Alibaba
2026-08-03

Qwen3.8 Max

Alibaba’s Qwen3.8 Max is a hosted Qwen model added to OpenRouter with a 1,000,000-token context window. It is positioned for long-context general-purpose language tasks, reasoning, and generation.

text-generationreasoninglong-context
Context
1.0M
Input/1M
$2
Thinking Machines
2026-07-30

Inkling Small

Inkling Small is a hosted 524K-context model from Thinking Machines, added on OpenRouter as a smaller variant of the Inkling model family. It is suited for long-context text and reasoning workloads where a lighter model is preferred.

text-generationreasoninglong-context
Context
524K
Input/1M
$0.45
OpenAI
2026-07-30

GPT-Realtime

OpenAI model referenced in a 2026-07-30 OpenAI case study as powering avatarin's 24/7 multilingual retail agent. It is positioned for real-time conversational AI experiences such as customer support agents.

real-time-audiospeech-to-speechmultilingual+1
Context
N/A
License
Open
Alibaba
2026-07-27

Qwen3.7 Flash

Qwen3.7 Flash is a fast hosted Qwen model variant added to OpenRouter with a 1,000,000-token context window. It is suited for long-context text, reasoning, and general assistant workloads where lower latency is important.

text-generationreasoninglong-context+1
Context
1.0M
Input/1M
$0.03
Anthropic
2026-07-24

Claude Opus 5

Anthropic's most capable Opus model, positioned for advanced reasoning, agentic systems, and production inference workloads.

text-generationreasoningagentic-workflows
Context
1.0M
Input/1M
$5
Anthropic
2026-07-24

Claude Opus 5 Fast

A faster hosted variant of Claude Opus 5 with the same 1M-token context window, aimed at lower-latency use in advanced reasoning and agentic workflows.

text-generationreasoningagentic-workflows
Context
1.0M
Input/1M
$10
InclusionAI
2026-07-23

Ling 3.0 Flash

Ling 3.0 Flash is a hosted general-purpose language model from InclusionAI, newly listed on OpenRouter with a 262K-token context window. The Flash variant is positioned for fast, long-context chat and reasoning workloads.

text-generationchatreasoning+1
Context
262K
Input/1M
$0.021
Poolside
2026-07-21

Laguna S 2.1

Laguna S 2.1 is a long-context foundation model from Poolside, newly listed on OpenRouter and also present in the Ollama library. The OpenRouter listing reports a 1,048,576-token context window.

text-generationlong-context
Context
1.0M
Input/1M
$0.09
Google
2026-07-21

Gemini 3.6 Flash

Gemini 3.6 Flash is a Google Gemini model newly added to OpenRouter with a 1,048,576-token context window. It is positioned as a Flash-family hosted model for long-context text generation and reasoning workloads.

text-generationreasoninglong-context
Context
1.0M
Input/1M
$0.75
Google
2026-07-21

Gemini 3.5 Flash-Lite

Gemini 3.5 Flash-Lite is a lightweight Google Gemini hosted model newly added to OpenRouter with a 1,048,576-token context window. It is suited to cost- and latency-sensitive long-context text generation workloads.

text-generationlong-contextlow-latency
Context
1.0M
Input/1M
$0.3
Meituan
2026-07-20

LongCat 2.0

LongCat 2.0 is a Meituan long-context foundation model newly added to OpenRouter. The listing reports a 1,048,756-token context window for hosted text-generation workloads.

text-generationlong-context
Context
1.0M
Input/1M
$0.3
Thinking Machines
2026-07-17

Inkling

Inkling is a hosted model newly listed on OpenRouter with a 1,048,576-token context window. The discovery data verifies its availability and long-context capability, but does not provide further architectural or benchmark details.

text-generationlong-context
Context
1.0M
Input/1M
$0.95
Moonshot AI
2026-07-16

Kimi K3

Kimi K3 is a Moonshot AI long-context hosted language model added to OpenRouter with a 1,048,576-token context window. It is positioned for large-context text, reasoning, and coding workloads.

text-generationlong-contextreasoning+1
Context
1.0M
Input/1M
$0.67
Meta
2026-07-16

Muse Spark 1.1

Muse Spark 1.1 is a Meta model added to OpenRouter with a 1,048,576-token context window. The discovery data verifies it as a newly listed long-context hosted model.

text-generationlong-context
Context
1.0M
Input/1M
$1.25
KwaiPilot
2026-07-10

KAT Coder Air v2.5

KAT Coder Air v2.5 is a code-focused model from the KwaiPilot namespace added to OpenRouter with a 256,000-token context window. It appears to be the lighter Air variant of the KAT Coder v2.5 family.

code-generationlong-contexttext-generation
Context
256K
Input/1M
$0.15
KwaiPilot
2026-07-10

KAT Coder Pro v2.5

KAT Coder Pro v2.5 is a code-focused model from the KwaiPilot namespace added to OpenRouter with a 256,000-token context window. It appears to be the higher-capability Pro variant of the KAT Coder v2.5 family.

code-generationlong-contexttext-generation
Context
256K
Input/1M
$0.74
OpenAI
2026-07-09

GPT-5.6 Luna Pro

OpenRouter-listed GPT-5.6 hosted model variant with a 1,050,000-token context window. It is positioned as a long-context frontier text model for demanding reasoning and productivity workloads.

text-generationreasoningcode-generation+1
Context
1.1M
Input/1M
$0.2
OpenAI
2026-07-09

GPT-5.6 Luna

OpenRouter-listed GPT-5.6 hosted model variant with a 1,050,000-token context window. It is intended for broad long-context text, coding, and reasoning use cases.

text-generationreasoningcode-generation+1
Context
1.1M
Input/1M
$0.2
OpenAI
2026-07-09

GPT-5.6 Terra Pro

OpenRouter-listed GPT-5.6 hosted model variant with a 1,050,000-token context window. It is a proprietary long-context model for advanced text generation, reasoning, and coding workflows.

text-generationreasoningcode-generation+1
Context
1.1M
Input/1M
$2
OpenAI
2026-07-09

GPT-5.6 Terra

OpenRouter-listed GPT-5.6 hosted model variant with a 1,050,000-token context window. It targets long-context general-purpose assistance, reasoning, and code-generation tasks.

text-generationreasoningcode-generation+1
Context
1.1M
Input/1M
$2
OpenAI
2026-07-09

GPT-5.6 Sol Pro

OpenRouter-listed GPT-5.6 hosted model variant with a 1,050,000-token context window. It is a proprietary long-context model for high-end reasoning, coding, and productivity workloads.

text-generationreasoningcode-generation+1
Context
1.1M
Input/1M
$2
OpenAI
2026-07-09

GPT-5.6

OpenAI frontier general-purpose model announced as delivering more intelligence per token, stronger performance per dollar, and scalable capability for demanding work.

text-generationreasoningcode-generation+2
Context
1.1M
License
Open
OpenAI
2026-07-08

GPT-Live

OpenAI voice model generation for natural human-AI interaction, announced as powering ChatGPT Voice.

voicespeech-to-speechaudio-input+2
Context
N/A
License
Open
xAI
2026-07-08

Grok 4.5

xAI hosted Grok model newly added on OpenRouter with a 500k-token context window for long-context general AI tasks.

text-generationreasoningcode-generation+1
Context
500K
Input/1M
$2
Aion Labs
2026-07-07

Aion 3.0

Aion Labs hosted general-purpose model newly added on OpenRouter with a 131k-token context window.

text-generationlong-context
Context
131K
Input/1M
$3
Aion Labs
2026-07-07

Aion 3.0 Mini

Smaller Aion Labs hosted model newly added on OpenRouter with a 131k-token context window.

text-generationlong-context
Context
131K
Input/1M
$0.7
Tencent
2026-07-06

Tencent HY3

Tencent long-context hosted model newly added on OpenRouter with a 262k-token context window.

text-generationlong-context
Context
262K
Input/1M
$0.132
Poolside
2026-07-02

Laguna XS 2.1

Laguna XS 2.1 is a Poolside model added to OpenRouter with a 262K-token context window and also available in the Ollama library. It is suited to long-context coding and text workflows.

code-generationlong-contexttext-generation
Context
262K
Input/1M
$0.06
NVIDIA
2026-07-01

NVIDIA Nemotron 3 Nano

NVIDIA Nemotron 3 Nano is an open-weight Nemotron model family referenced as newly supported on Amazon Bedrock in AWS GovCloud, including Nano 9B v2, Nano 12B v2, and Nano 30B variants.

text-generationinstruction-following
Context
N/A
License
Open
Anthropic
2026-06-30

Claude Sonnet 5

Anthropic's latest-generation Sonnet model, described as its most capable Sonnet model and made available on Amazon Bedrock, Claude Platform on AWS, and OpenRouter.

text-generationreasoningcode-generation+1
Context
1.0M
Input/1M
$2
Google
2026-06-30

Gemini 3.1 Flash-Lite Image

A Google Gemini 3.1 Flash-Lite image model added to OpenRouter, providing image-focused multimodal capabilities with a 65,536-token context window.

visionimage-generationmultimodal
Context
66K
Input/1M
$0.25
Amazon
2026-06-29

Amazon Nova 2 Lite

Amazon Nova 2 Lite is a lightweight multimodal Nova model referenced for cost-optimized scanned document processing, where it handles native multimodal extraction before downstream Claude processing.

multimodalvisiondocument-understanding+1
Context
1.0M
Input/1M
$0.3
OpenAI
2026-06-26

GPT-5.6 Sol

GPT-5.6 Sol is a next-generation OpenAI model previewed with stronger capabilities in coding, science, and cybersecurity, paired with OpenAI's most advanced safety stack.

code-generationscientific-reasoningcybersecurity+2
Context
1.1M
Input/1M
$2
Sakana AI
2026-06-24

Fugu Ultra

Fugu Ultra is a Sakana model listed on OpenRouter with a 1M-token context window. It is positioned for long-context general-purpose reasoning and text generation.

long-contextreasoningtext-generation
Context
1.0M
Input/1M
$5
OpenAI
2026-06-22

GPT-5.5-Cyber

Cybersecurity-focused OpenAI model introduced with Daybreak tools to help organizations find, validate, and patch vulnerabilities at scale.

cybersecurityvulnerability-detectioncode-analysis+1
Context
N/A
License
proprietary
Google
2026-06-18

Gemini 3.1 Flash Image

Google Gemini image-focused model listed on OpenRouter with a 131,072-token context window. It is positioned as a Flash-tier multimodal/image model for lower-latency image-centric workloads.

multimodalvisionimage-generation
Context
131K
Input/1M
$0.5
Google
2026-06-18

Gemini 3 Pro Image

Google Gemini Pro-tier image-focused model listed on OpenRouter with a 65,536-token context window. It targets higher-capability multimodal and image-generation use cases than Flash-tier variants.

multimodalvisionimage-generation
Context
66K
Input/1M
$2
Zhipu AI / Z.ai
2026-06-16

GLM-5.2

GLM-5.2 is a Z.ai / Zhipu AI foundation model listed on OpenRouter with a 1,048,576-token context window and available in the Ollama library. It targets long-context reasoning, generation, and coding workloads.

reasoninglong-contexttext-generation+1
Context
1.0M
Input/1M
$0.0224
Moonshot AI
2026-06-12

Kimi K2.7 Code

Kimi K2.7 Code is a Moonshot AI coding-focused model listed on OpenRouter with a 262K-token context window and available in the Ollama library. It is aimed at software engineering and long-context code understanding tasks.

code-generationlong-contextreasoning
Context
262K
Input/1M
$0.6712
Google DeepMind
2026-06-10

DiffusionGemma

An experimental open model from Google DeepMind built for exceptionally fast text generation, with NVIDIA optimizations for local and accelerated inference across RTX, RTX PRO, and DGX Spark systems.

text-generationfast-inferencelocal-inference
Context
N/A
License
Open
Open Weight
Cohere
2026-06-09

North Mini Code

Cohere’s first model for developers, focused on coding and developer-assistance workflows.

code-generationdeveloper-assistancetext-generation
Context
N/A
License
Open
Open Weight
Anthropic
2026-06-09

Claude Fable 5

Claude Fable 5 is a newly listed Anthropic model on OpenRouter with a 1,000,000-token context window. The discovery data verifies it as a new Anthropic model added during the target date range.

text-generationreasoninglong-context
Context
1.0M
Input/1M
$10
NVIDIA
2026-06-04

NVIDIA Nemotron 3 Ultra 550B A55B

NVIDIA Nemotron 3 Ultra 550B A55B is a large open-weight Nemotron-family model listed on OpenRouter with a 1M-token context window and available in the Ollama library. It is intended for long-context reasoning and general text-generation workloads.

reasoninglong-contexttext-generation+1
Context
1.0M
Input/1M
$0.5
Alibaba
2026-06-03

Qwen3.7 Plus

A Qwen-family large language model added on OpenRouter with a 1,000,000-token context window for long-context general-purpose AI workloads.

text-generationlong-context
Context
1.0M
Input/1M
$0.32
Google
2026-05-28

Gemini Omni

Gemini Omni is a Google Gemini-family model announced at Google I/O 2026 and showcased in Google demos alongside Gemini 3.5. The discovery data verifies the announcement but does not provide context length, output limit, or pricing details.

multimodal
Context
N/A
License
Open
Anthropic
2026-05-27

Claude Opus 4.8

Anthropic's Claude Opus 4.8 is a proprietary frontier model listed on OpenRouter and announced as available on AWS. The provided data highlights its use for agentic systems and production inference workloads, with a 1,000,000-token context window.

text-generationreasoninglong-context+1
Context
1.0M
Input/1M
$5
Anthropic
2026-05-27

Claude Opus 4.8 Fast

Claude Opus 4.8 Fast is an Anthropic model variant added to OpenRouter with a 1,000,000-token context window. It is positioned as the fast variant of Claude Opus 4.8 for lower-latency agentic and production workloads.

text-generationreasoninglong-context+1
Context
1.0M
Input/1M
$10
Google
2026-05-19

Gemini 3.5

Google’s Gemini 3.5 is a new frontier model series focused on combining strong general intelligence with agentic action/tool use, announced at Google I/O 2026.

reasoningtool-useagentic-workflows+3
Context
N/A
License
Open
Google
2026-05-19

Gemini 3.5 Flash

A fast, efficient Gemini 3.5-series model variant listed on OpenRouter, intended for low-latency agentic and general assistant workloads with a very large context window.

reasoningtool-useagentic-workflows+2
Context
1.0M
Input/1M
$1.5
Anthropic
2026-05-12

Claude Opus 4.7 Fast

A latency-optimised variant of Claude Opus 4.7 with a one-million-token context window, designed for real-time agentic workflows, dependency auditing, and large codebase analysis where Opus-class reasoning is required at lower response times.

reasoninglong-contexttool-use+3
Context
1.0M
Input/1M
$30
OpenAI
2026-05-05

GPT-5.5 Instant

An updated default ChatGPT model focused on smarter, more accurate responses with reduced hallucinations and improved personalization controls.

reasoningtext-generationcode-generation+1
Context
N/A
License
Open
OpenAI
2026-05-05

gpt-chat-latest

A ChatGPT-aligned OpenAI model alias newly added to OpenRouter with a 400k token context window, intended for general conversational and assistant-style use.

text-generationlong-contextinstruction-following+1
Context
400K
Input/1M
$5
Mistral AI
2026-04-30

Mistral Medium 3.5

A Mistral AI foundation model newly listed on OpenRouter with a 262k token context window, positioned as a balanced medium-tier model for general purpose generation and reasoning tasks.

text-generationreasoninglong-context+1
Context
262K
Input/1M
$1.5
xAI
2026-04-30

Grok 4.3

A new Grok-series flagship model variant listed on OpenRouter with a 1M-token context window, aimed at high-context general reasoning and assistant use.

reasoninglong-contextchat+2
Context
1.0M
Input/1M
$1.25
Alibaba
2026-04-27

Qwen3.6 Max (Preview)

A preview flagship Qwen3.6 foundation model variant aimed at strong general-purpose reasoning and instruction following with a large context window.

reasoningtool-usecode-generation+2
Context
262K
Input/1M
$1.027
Alibaba
2026-04-27

Qwen3.6 Flash

A speed-optimized Qwen3.6 foundation model for low-latency chat and agent workloads while retaining a very large context window.

long-contextinstruction-followingtool-use+2
Context
1.0M
Input/1M
$0.1875
DeepSeek
2026-04-24

DeepSeek V4 Pro

DeepSeek’s V4 Pro foundation model listing with a 1M-token context window, intended for long-context reasoning and agentic workloads.

long-contextreasoningcode-generation+2
Context
1.0M
Input/1M
$0.2088
DeepSeek
2026-04-24

DeepSeek V4 Flash

DeepSeek’s V4 Flash foundation model listing with a 1M-token context window, optimized for lower-latency long-context tasks.

long-contextreasoningcode-generation+2
Context
1.0M
Input/1M
$0.03
OpenAI
2026-04-23

GPT-5.5

OpenAI’s flagship GPT-5.5 model, positioned as faster and more capable for complex tasks like coding, research, and data analysis across tools.

reasoningcode-generationtool-use+3
Context
1.1M
Input/1M
$5
OpenAI
2026-04-22

OpenAI Privacy Filter

An open-weight OpenAI model for detecting and redacting personally identifiable information (PII) in text, intended as a privacy/safety component in pipelines.

pii-detectiontext-redactioncompliance+1
Context
N/A
License
Open
OpenAI
2026-04-21

GPT-5.4 Image 2

An OpenAI multimodal model oriented around image understanding/generation workflows, listed on OpenRouter as a new GPT-5.4 image-capable offering with a large context window.

visionimage-generationmultimodal+2
Context
272K
Input/1M
$8
Google
2026-04-16

Nano Banana 2

An image generation model in the Gemini app that uses personal context and Google Photos to create more personalized images.

image-generationpersonalizationphoto-editing+1
Context
N/A
License
Open
OpenAI
2026-04-16

GPT-Rosalind

A frontier reasoning model for life sciences research, positioned to accelerate drug discovery workflows including genomics analysis and protein reasoning.

reasoninglife-sciencescode-generation+1
Context
N/A
License
Open
Anthropic
2026-04-16

Claude Opus 4.7

A new Claude Opus-series frontier model version listed on OpenRouter with a 1M-token context window, intended for high-end reasoning and long-context workloads.

reasoninglong-contexttool-use+3
Context
1.0M
Input/1M
$5
Google
2026-04-15

Gemini 3.1 Flash TTS

A text-to-speech model focused on next-generation expressive speech, now available across Google products.

text-to-speechspeech-generationexpressive-audio+1
Context
N/A
License
Open
OpenAI
2026-04-14

GPT-5.4-Cyber

A GPT-5.4-derived model introduced under OpenAI’s Trusted Access for Cyber program, intended for vetted cyber defenders with strengthened safeguards for cybersecurity use cases.

reasoningcybersecuritythreat-analysis+2
Context
N/A
License
Open
Anthropic
2026-04-07

Claude Opus 4.6 Fast

A faster variant of Claude Opus 4.6 exposed via OpenRouter, aimed at high-throughput production workloads while retaining the Opus-class capability profile.

reasoningtext-generationtool-use+2
Context
1.0M
License
Open
Google
2026-04-03

Gemma 4 26B A4B IT

An instruction-tuned Gemma 4 model listed on OpenRouter, positioned as a large open model for general-purpose chat and instruction following with a long context window.

text-generationinstruction-followingreasoning+2
Context
262K
Input/1M
$0.09
Alibaba
2026-04-02

Qwen3.6-Plus

A long-context Qwen model variant listed on OpenRouter, intended for general-purpose instruction following and long-document workloads.

text-generationinstruction-followinglong-context+2
Context
1.0M
Input/1M
$0.325
Google
2026-04-02

Gemma 4 31B IT

An instruction-tuned Gemma 4 family model offered via OpenRouter with a very large context window, aimed at general-purpose assistant and agentic workflows.

instruction-followingreasoningtool-use+2
Context
262K
Input/1M
$0.09
Google
2026-03-31

Veo 3.1 Lite

Cost-effective video generation model available in paid preview via the Gemini API and for testing in Google AI Studio.

video-generation
Context
N/A
License
Open
Google
2026-03-30

Lyria 3 Pro (Preview)

A preview Lyria 3 variant surfaced on OpenRouter, associated with Google’s Lyria music/audio generation stack for higher-end generation workflows.

audio-generationmusic-generationlong-context
Context
1.0M
Input/1M
$0
Google
2026-03-30

Lyria 3 CLIP (Preview)

A preview Lyria 3 variant listed on OpenRouter, likely intended for clip-based audio/music generation or related multimodal embedding workflows within the Lyria stack.

audio-generationmusic-generationclip-generation
Context
1.0M
Input/1M
$0
Alibaba
2026-03-30

Qwen3.6 Plus Preview

Preview release of Alibaba's Qwen 3.6 Plus model as listed on OpenRouter, offering a very large context window for general-purpose text tasks.

text-generationreasoninglong-context+1
Context
1.0M
License
Open
Google
2026-03-26

Gemini 3.1 Flash Live

A low-latency, live audio-capable Gemini Flash model designed for more natural, reliable real-time voice interactions across Google products.

audioreal-timemultimodal+2
Context
N/A
License
Open
Google
2026-03-25

Lyria 3

Google’s newest music generation model, available in paid preview through the Gemini API and for testing in Google AI Studio.

music-generationaudiocreative-generation
Context
N/A
License
Open
Mistral AI
2026-03-16

Mistral Small 2603

A new Mistral Small series release listed on OpenRouter with a 262k context window, positioned as a general-purpose foundation model for long-context workloads.

long-contextreasoningtext-generation+2
Context
262K
Input/1M
$0.15
xAI
2026-03-12

Grok 4.20 (Beta)

A Grok 4.20 beta model offering a very large (2M token) context window for long-context general-purpose chat and reasoning workloads.

long-contextreasoningchat+2
Context
2.0M
License
Open
xAI
2026-03-12

Grok 4.20 Multi-Agent (Beta)

A Grok 4.20 beta variant positioned for multi-agent workflows, with a 2M token context window for coordinating longer multi-step tasks.

long-contextreasoningagentic+2
Context
2.0M
License
Open
NVIDIA
2026-03-11

NVIDIA Nemotron 3 Super (120B, A12B)

An open model from NVIDIA designed for scalable agentic AI, described as a 120B-parameter model with 12B active parameters and optimized throughput.

reasoningagenticlong-context+2
Context
262K
Input/1M
$0.08
Alibaba
2026-03-10

Qwen3.5-9B

A 9B-parameter Qwen3.5 foundation model with a large (262k token) context window, positioned for general chat and reasoning with long-context inputs.

chatreasoninglong-context+2
Context
262K
Input/1M
$0.1
OpenAI
2026-03-05

GPT-5.4

OpenAI frontier foundation model positioned as more capable and efficient for professional work, with state-of-the-art coding, computer use, and tool search, plus a 1M-token context window.

reasoningcode-generationtool-use+3
Context
1.1M
Input/1M
$2.5
OpenAI
2026-03-05

GPT-5.4 Pro

Higher-tier GPT-5.4 offering listed by OpenRouter, providing a 1M-token context window for advanced professional and agentic workloads.

reasoningcode-generationtool-use+3
Context
1.1M
Input/1M
$30
OpenAI
2026-03-03

GPT-5.3 Instant

Conversation-focused GPT-5.3 variant announced by OpenAI for smoother, more useful everyday chat interactions.

chatreasoningsummarization+1
Context
128K
License
Open
Google
2026-03-03

Gemini 3.1 Flash-Lite

Google’s fastest and most cost-efficient Gemini 3 series model, built for intelligence at scale.

reasoningchattool-use+2
Context
1.0M
Input/1M
$0.25
Google
2026-02-26

Gemini 3.1 Flash Image (Preview)

Google's Flash-speed image generation and editing model referenced as "Nano Banana 2" and listed on OpenRouter as a Gemini 3.1 Flash Image preview.

image-generationimage-editingmultimodal
Context
66K
Input/1M
$0.5
Google
2026-02-25

Gemini 3.1 Pro Preview (Custom Tools)

A Gemini 3.1 Pro preview variant listed on OpenRouter that is explicitly labeled for custom tools, suggesting enhanced tool-use integration with a very large context window.

tool-usereasoninglong-context
Context
1.0M
Input/1M
$2
Alibaba
2026-02-25

Qwen3.5 Flash 02-23

A Qwen3.5 Flash model snapshot (02-23) newly listed on OpenRouter with a 1M-token context window, positioned for fast, long-context inference.

long-contextreasoning
Context
1.0M
Input/1M
$0.065
Alibaba
2026-02-25

Qwen3.5 122B A10B

A large Qwen3.5 Mixture-of-Experts-style model variant newly added on OpenRouter, offering a large 262k-token context window.

reasoninglong-context
Context
262K
Input/1M
$0.26
Alibaba
2026-02-25

Qwen3.5 35B A3B

A Qwen3.5 model variant newly listed on OpenRouter with a 262k-token context window, intended as a mid-sized foundation option in the Qwen3.5 family.

reasoninglong-context
Context
262K
Input/1M
$0.15
Alibaba
2026-02-25

Qwen3.5 27B

A Qwen3.5 27B foundation model newly added on OpenRouter, providing a 262k-token context window for general assistant workloads.

reasoninglong-context
Context
262K
Input/1M
$0.195
OpenAI
2026-02-24

GPT-5.3 Codex

A new Codex-branded GPT-5.3 model intended for code-centric use cases, listed as newly added on OpenRouter with a large context window.

code-generationreasoning
Context
400K
Input/1M
$1.75
Google
2026-02-19

Gemini 3.1 Pro Preview

Preview release of Google's Gemini 3.1 Pro model with a very large context window, aimed at advanced general-purpose reasoning and long-context workloads.

reasoninglong-contexttool-use
Context
1.0M
Input/1M
$2
Alibaba
2026-02-16

Qwen3.5-Plus-02-15

Alibaba Qwen 3.5 'Plus' model variant as listed on OpenRouter, featuring a 1M-token context window for long-context general-purpose generation and analysis.

reasoninglong-contextcode-generation
Context
1.0M
Input/1M
$0.26
Alibaba
2026-02-16

Qwen3.5-397B-A17B

Large-scale Qwen 3.5 model (397B with A17B MoE-style routing indicated by the name) added on OpenRouter, intended for high-end reasoning and generation with a 262K context window.

reasoninglong-contextcode-generation
Context
262K
Input/1M
$0.55
Google
2026-02-15

Gemini 3.1 Pro

Advanced intelligence with complex problem-solving, agentic and vibe coding capabilities

reasoningcode-generationagentic-coding+3
Context
2.0M
Input/1M
$3
xAI
2026-02-15

Grok 420

xAI's most advanced model with breakthrough capabilities (early access)

reasoningcode-generationagentic-tasks+2
Context
1.0M
Input/1M
$1.25
xAI
2026-02-15

Grok 420 Multi-Agent

Grok 420 variant optimized for multi-agent orchestration

reasoningmulti-agentagentic-tasks+1
Context
1.0M
Input/1M
$1.25
Google
2026-02-10

Gemini 3.0 Pro

Latest Gemini Pro with enhanced reasoning and coding capabilities across all modalities

reasoningcode-generationcode-review+3
Context
2.0M
Input/1M
$2
Anthropic
2026-02-01

Claude 4.6 Opus

Latest flagship Anthropic model with state-of-the-art reasoning, coding expertise, and agentic capabilities

reasoningcode-generationcode-review+5
Context
1.0M
Input/1M
$25
Anthropic
2026-02-01

Claude 4.6 Sonnet

Most advanced Claude Sonnet with exceptional coding and reasoning, ideal balance of capability and efficiency

code-generationcode-reviewanalysis+4
Context
500K
Input/1M
$5
Anthropic
2026-02-01

Claude 4.6 Haiku

Ultra-fast Claude 4.6 model for real-time applications and high-volume processing

code-generationquick-analysisfunction-calling+1
Context
200K
Input/1M
$0.6
OpenAI
2026-02-01

GPT-5.2 Pro

Most capable GPT-5.2 variant producing smarter and more precise responses

reasoningcode-generationagentic-tasks+3
Context
256K
Input/1M
$21
Google
2026-01-20

Gemini 3.0 Flash

Next-generation fast model with improved efficiency and multimodal capabilities

code-generationanalysisvision+2
Context
2.0M
Input/1M
$0.15
OpenAI
2026-01-20

GPT-5.2 Codex

Most intelligent coding model optimized for long-horizon agentic coding tasks

code-generationcode-refactoringagentic-coding+2
Context
256K
Input/1M
$1.75
Google
2026-01-20

Gemini 3 Pro

Google's state-of-the-art reasoning model with advanced multimodal understanding

reasoningcode-generationvision+2
Context
2.0M
Input/1M
$2.5
xAI
2026-01-15

Grok 4

Latest iteration of xAI's flagship model with breakthrough performance

reasoningcode-generationcomplex-analysis+2
Context
500K
Input/1M
$8
xAI
2026-01-15

Grok 4 Mini

Efficient version of Grok 4 optimized for speed and cost-effectiveness

code-generationanalysisfunction-calling
Context
256K
Input/1M
$1.5
OpenAI
2026-01-15

GPT-5.2

OpenAI's best model for coding and agentic tasks across industries

reasoningcode-generationagentic-tasks+3
Context
256K
Input/1M
$1.75
Anthropic
2026-01-15

Claude Opus 4.6

The most intelligent Claude model for building agents and coding with extended thinking

reasoningcode-generationagentic-tasks+3
Context
1.0M
Input/1M
$5
Anthropic
2026-01-15

Claude Sonnet 4.6

Best combination of speed and intelligence with extended thinking support

code-generationreasoningextended-thinking+3
Context
1.0M
Input/1M
$3
Google
2026-01-10

Gemini 3 Flash

Frontier-class performance rivaling larger models at a fraction of the cost

code-generationanalysisvision+2
Context
2.0M
Input/1M
$0.2
xAI
2026-01-01

Grok 4 Voice

Grok 4 with real-time voice conversation capabilities

voicereal-timeconversation+1
Context
256K
Input/1M
$0
Alibaba
2026-01-01

Qwen 3 Coder 235B

Alibaba's largest and most capable coding model

code-generationcode-reviewdebugging+2
Context
256K
License
Apache 2.0
OpenAI
2025-12-01

GPT-5

Next-generation GPT model (announced for 2025)

multimodalreasoningcode-generation+1
Context
256K
Input/1M
$1.25
Alibaba
2025-12-01

Qwen Coder 3 72B

Alibaba's latest flagship coding model with exceptional performance

code-generationcode-reviewdebugging+2
Context
256K
License
Apache 2.0
OpenAI
2025-12-01

GPT-OSS 120B

OpenAI's most powerful open-weight model, fits on H100 GPU

code-generationreasoninganalysis+1
Context
128K
Input/1M
$0.037
OpenAI
2025-12-01

GPT-OSS 20B

Medium-sized open-weight model for low latency

code-generationanalysis
Context
64K
Input/1M
$0.018
Google
2025-12-01

Gemini Deep Research

Agentic model for autonomous multi-step research across hundreds of sources

researchanalysissynthesis+2
Context
1.0M
Input/1M
$5
Meta
2025-12-01

Llama 4 Coder 405B

Meta's most capable code model based on Llama 4 architecture

code-generationcode-reviewdebugging+2
Context
256K
License
Llama 4 Community License
Meta
2025-12-01

Llama 4 Coder 70B

Efficient Llama 4 coding variant for production use

code-generationcode-reviewdebugging
Context
256K
License
Llama 4 Community License
Google
2025-11-15

Gemini 2.5 Ultra

Google's most powerful model for demanding enterprise tasks and complex reasoning

reasoningcode-generationcomplex-analysis+3
Context
2.0M
Input/1M
$10
OpenAI
2025-11-15

GPT-5.1 Codex Mini

Cost-effective smaller version of GPT-5.1 Codex

code-generationcode-completionquick-fixes
Context
128K
Input/1M
$0.25
xAI
2025-11-01

Grok 3.5

xAI's advanced model with improved reasoning and real-time knowledge integration

reasoningcode-generationreal-time-knowledge+1
Context
256K
Input/1M
$5
OpenAI
2025-11-01

GPT-5.1 Codex Max

GPT-5.1 Codex optimized for long-running coding tasks

code-generationagentic-codinglong-running-tasks+1
Context
512K
Input/1M
$1.25
Zhipu AI
2025-11-01

CodeGeeX 5

Latest multilingual code generation model with enhanced capabilities

code-generationcode-completionmulti-language+1
Context
66K
License
Apache 2.0
Anthropic
2025-10-15

Claude 4.5 Opus

Anthropic's most capable model with breakthrough reasoning, extended thinking, and exceptional coding abilities

reasoningcode-generationcode-review+4
Context
500K
Input/1M
$20
Anthropic
2025-10-15

Claude 4.5 Sonnet

High-performance Claude model balancing intelligence and speed, excels at code generation and analysis

code-generationcode-reviewanalysis+3
Context
500K
Input/1M
$4
Anthropic
2025-10-15

Claude 4.5 Haiku

Fastest Claude 4.5 model optimized for quick tasks and high-throughput applications

code-generationquick-analysisfunction-calling+1
Context
200K
Input/1M
$0.5
OpenAI
2025-10-01

GPT-5.1 Codex

GPT-5.1 optimized for agentic coding in Codex environment

code-generationagentic-codingcode-review+2
Context
256K
Input/1M
$1.25
Anthropic
2025-10-01

Claude Haiku 4.5

Fastest Claude model with near-frontier intelligence and extended thinking

code-generationextended-thinkingvision+1
Context
200K
Input/1M
$1
Google
2025-10-01

Gemini Computer Use

Specialized model for UI automation - clicking, typing, and navigating browser tasks

computer-useui-automationbrowser-control+1
Context
256K
Input/1M
$3
OpenAI
2025-09-15

GPT-5.1

Intelligent reasoning model for coding and agentic tasks with configurable reasoning effort

reasoningcode-generationagentic-tasks+2
Context
256K
Input/1M
$1.25
Google
2025-09-01

Gemini 2.5 Pro Thinking

Google's most advanced reasoning model with extended chain-of-thought capabilities

reasoningcode-generationcomplex-analysis+3
Context
2.0M
Input/1M
$3.5
BigCode
2025-09-01

StarCoder3 32B

Next-generation open-source code LLM with improved capabilities

code-generationcode-completionmulti-language
Context
33K
License
BigCode OpenRAIL-M
GitHub
2025-09-01

GitHub Copilot Workspace

Agentic AI for complex multi-file development tasks

agentic-codingmulti-file-editingcode-generation+1
Context
256K
License
Open
Google
2025-08-01

Gemini Code

Specialized coding model optimized for software development and code understanding

code-generationcode-reviewdebugging+2
Context
1.0M
Input/1M
$1.5
xAI
2025-08-01

Grok Vision

Multimodal Grok model with advanced image and document understanding

visioncode-generationdocument-analysis+1
Context
128K
Input/1M
$3
OpenAI
2025-08-01

GPT-5 Codex

GPT-5 optimized for agentic coding in Codex

code-generationagentic-codingdebugging+1
Context
256K
Input/1M
$8
Google
2025-08-01

Gemini 2.5 Flash-Lite

Fastest and most budget-friendly multimodal model in the Gemini 2.5 family

code-generationanalysisvision
Context
1.0M
Input/1M
$0.1
IBM
2025-08-01

Granite Code 3 34B

IBM's latest enterprise code model with enhanced security awareness

code-generationcode-reviewsecurity-analysis+1
Context
33K
License
Apache 2.0
OpenAI
2025-07-01

o4-mini

Next-generation compact reasoning model

reasoningcode-generationmath
Context
200K
Input/1M
$1.1
OpenAI
2025-07-01

GPT-5 Pro

GPT-5 variant producing smarter and more precise responses

reasoningcode-generationcomplex-analysis+1
Context
256K
Input/1M
$15
OpenAI
2025-07-01

GPT-5 Nano

Fastest, most cost-efficient version of GPT-5

code-generationquick-analysis
Context
64K
Input/1M
$0.05
Open Source
2025-07-01

OlympicCoder 32B

Competition-grade code model fine-tuned on competitive programming

code-generationalgorithm-designproblem-solving
Context
33K
License
Apache 2.0
OpenAI
2025-06-15

GPT-5 Mini

Faster, cost-efficient version of GPT-5 for well-defined tasks

code-generationanalysisfunction-calling
Context
128K
Input/1M
$0.25
DeepSeek
2025-06-01

DeepSeek Coder V3

Latest DeepSeek coding model with state-of-the-art code understanding

code-generationcode-reviewdebugging+2
Context
256K
License
DeepSeek License
Mistral AI
2025-06-01

Mistral Large Code

Mistral's flagship model optimized for enterprise coding tasks

code-generationcode-reviewarchitecture-design+1
Context
256K
Input/1M
$3
GitHub
2025-06-01

GitHub Copilot Chat

Conversational AI for coding powered by GPT-5

code-generationchatcode-explanation+1
Context
128K
License
Open
Anthropic
2025-05-22

Claude 4 Opus

Most capable Claude model with extended thinking

extended-thinkingcode-generationdeep-analysis+1
Context
200K
License
Open
Anthropic
2025-05-22

Claude 4 Sonnet

Balanced Claude 4 model with strong coding abilities

code-generationanalysisreasoning+1
Context
200K
License
Open
Anthropic
2025-05-22

Claude 4 Haiku

Fast and efficient Claude 4 model

code-generationanalysisfast-inference
Context
200K
License
Open
Anthropic
2025-05-22

Claude Opus 4

Latest flagship Claude model with superior reasoning

deep-reasoningcode-generationanalysis+1
Context
200K
Input/1M
$15
Anthropic
2025-05-22

Claude Sonnet 4

Balanced Claude 4 model optimized for coding

code-generationreasoninganalysis
Context
200K
Input/1M
$3
OpenAI
2025-05-15

o4-mini Deep Research

Cost-efficient deep research model

reasoningresearchanalysis+1
Context
500K
Input/1M
$5
Mistral AI
2025-05-01

Devstral

Mistral's agentic coding model for complex development tasks

code-generationagentic-codingdebugging+1
Context
256K
Input/1M
$2
Alibaba
2025-04-28

Qwen 3 235B

Latest flagship Qwen model with MoE architecture

code-generationreasoningmultilingual+1
Context
131K
License
Apache 2.0
Alibaba
2025-04-28

Qwen 3 32B

Balanced Qwen 3 model for diverse tasks

code-generationreasoninganalysis
Context
131K
Input/1M
$0.08
Alibaba
2025-04-28

Qwen 3 8B

Efficient Qwen 3 model for quick tasks

code-generationchatanalysis
Context
131K
Input/1M
$0.117
Google
2025-04-17

Gemini 2.5 Flash

Fast and efficient Gemini 2.5 model with thinking

reasoningcode-generationfast-inference+1
Context
1.0M
Input/1M
$0.3
OpenAI
2025-04-14

GPT-4.1

Optimized GPT-4 variant with improved coding and instruction following

code-generationlong-contextinstruction-following+1
Context
1.0M
Input/1M
$2
OpenAI
2025-04-14

GPT-4.1 mini

Cost-effective version of GPT-4.1 for everyday tasks

code-generationlong-contextanalysis
Context
1.0M
Input/1M
$0.4
OpenAI
2025-04-14

GPT-4.1 nano

Smallest and fastest GPT-4.1 variant for quick tasks

code-generationsimple-analysisfast-inference
Context
1.0M
Input/1M
$0.1
Meta
2025-04-05

Llama 4 Scout

Llama 4 variant optimized for efficient multi-turn tasks

long-contextcode-generationreasoning+1
Context
10.0M
Input/1M
$0.1
Meta
2025-04-05

Llama 4 Maverick

Llama 4 variant for complex reasoning and coding

code-generationreasoningplanning+1
Context
1.0M
Input/1M
$0.1875
OpenAI
2025-04-01

o3 Deep Research

o3 optimized for multi-step deep research tasks

reasoningresearchanalysis+2
Context
500K
Input/1M
$20
Meta
2025-04-01

Code Llama 3 70B

Meta's latest Code Llama based on Llama 3 architecture

code-generationcode-completiondebugging+1
Context
131K
License
Llama 3 Community License
Meta
2025-04-01

Code Llama 3 8B

Efficient Code Llama 3 for local development

code-generationcode-completion
Context
131K
License
Llama 3 Community License
Google
2025-03-25

Gemini 2.5 Pro

Latest Gemini model with enhanced thinking capabilities

deep-reasoningcode-generationmultimodal+1
Context
1.0M
Input/1M
$1.25
Meta
2025-03-15

Llama 3.3 Coder 70B

Meta's latest code-specialized Llama model with enhanced coding capabilities

code-generationcode-reviewdebugging+2
Context
131K
License
Llama 3.3 Community License
Cohere
2025-03-13

Command A

Latest flagship model optimized for enterprise tasks

ragcode-generationreasoning+2
Context
256K
Input/1M
$2.5
OpenAI
2025-03-01

o3 Pro

o3 with more compute for better, more thorough responses

reasoningcode-generationcomplex-analysis+1
Context
200K
Input/1M
$20
OpenAI
2025-02-27

GPT-4.5 Preview

Next-generation GPT model with enhanced reasoning and multimodal capabilities

reasoningcode-generationvision+2
Context
128K
Input/1M
$75
Microsoft
2025-02-26

Phi-4-mini

Compact Phi-4 for efficient deployment

code-generationreasoningfunction-calling
Context
128K
License
MIT
Anthropic
2025-02-24

Claude 3.5 Opus

Enhanced Opus model with superior reasoning

deep-reasoningcode-generationcomplex-analysis
Context
200K
License
Open
xAI
2025-02-17

Grok-3

Next-generation Grok with enhanced reasoning

deep-reasoningcode-generationanalysis+1
Context
131K
License
Open
xAI
2025-02-17

Grok-3 mini

Efficient Grok-3 with thinking capabilities

reasoningcode-generationthinking+1
Context
131K
License
Open
Mistral AI
2025-02-15

Codestral 25.02

Latest Mistral coding model with enhanced performance

code-generationcode-reviewdebugging+2
Context
256K
Input/1M
$1.2
Google
2025-02-05

Gemini 2.0 Pro

Advanced Gemini 2.0 model for complex reasoning tasks

reasoningcode-generationmultimodal+1
Context
2.0M
License
Open
Mistral AI
2025-02-03

Mistral Saba

Expert model for Middle Eastern and South Asian languages

multilingualcode-generationregional-expertise
Context
33K
Input/1M
$0.2
Amazon
2025-02-03

Amazon Nova Premier

Most capable Nova model for complex reasoning

deep-reasoningcode-generationanalysis+1
Context
1.0M
License
Open
DeepSeek
2025-02-01

DeepSeek R1 Coder

DeepSeek's reasoning model specialized for complex coding tasks

code-generationreasoningcode-review+2
Context
256K
License
DeepSeek License
OpenAI
2025-01-31

o3-mini

Next-generation reasoning model with improved efficiency (announced)

reasoningcode-generationplanning+1
Context
200K
Input/1M
$1.1
OpenAI
2025-01-31

o3

Full o3 reasoning model for frontier problem solving

deep-reasoningcode-generationmath+2
Context
200K
Input/1M
$2
OpenAI
2025-01-31

o3 High

High compute version of o3 for maximum reasoning depth

deep-reasoningcode-generationmath+1
Context
200K
License
Open
Mistral AI
2025-01-30

Mistral Small 3

Latest small model with enhanced capabilities

code-generationreasoningfunction-calling
Context
33K
Input/1M
$0.1
NVIDIA
2025-01-23

Llama 3.3 70B Nemotron

NVIDIA-optimized Llama 3.3 for enterprise

code-generationreasoninganalysis+1
Context
128K
License
Llama 3.3 Community
Google
2025-01-21

Gemini 2.0 Flash Thinking

Flash model with explicit reasoning for complex tasks

reasoningcode-generationanalysis+1
Context
1.0M
License
Open
DeepSeek
2025-01-20

DeepSeek R1

Reasoning model with chain-of-thought capabilities

deep-reasoningmathcode-generation+1
Context
128K
Input/1M
$0.7
DeepSeek
2025-01-20

DeepSeek R1 Distill Qwen 32B

Distilled R1 model based on Qwen for efficient reasoning

reasoningcode-generationmath
Context
128K
License
MIT
DeepSeek
2025-01-20

DeepSeek R1 Distill Llama 70B

Distilled R1 model based on Llama 70B

reasoningcode-generationanalysis
Context
128K
Input/1M
$0.8
DeepSeek
2025-01-20

DeepSeek Reasoner

API-accessible reasoning model based on R1

deep-reasoningmathcode-generation
Context
128K
Input/1M
$0.55
DeepSeek
2025-01-20

DeepSeek R1 Distill Qwen 7B

Compact distilled reasoning model

reasoningcode-generationmath
Context
128K
License
MIT
DeepSeek
2025-01-20

DeepSeek R1 Distill Qwen 1.5B

Ultra-compact reasoning model

reasoningcode-generation
Context
128K
License
MIT
DeepSeek
2025-01-20

DeepSeek R1 Distill Llama 8B

Efficient Llama-based reasoning model

reasoningcode-generation
Context
128K
License
MIT
Mistral AI
2025-01-15

Codestral 2501

Latest Mistral coding model with improved performance and longer context

code-generationcode-reviewdebugging+2
Context
256K
Input/1M
$1
DeepSeek
2024-12-25

DeepSeek V3

MoE model with 671B parameters achieving frontier performance

code-generationreasoningmath+1
Context
128K
Input/1M
$0.27
LG AI Research
2024-12-19

EXAONE 3.5 32B

Korean-English bilingual model from LG

bilingualcode-generationreasoning
Context
33K
License
EXAONE AI Model
LG AI Research
2024-12-19

EXAONE 3.5 7.8B

Efficient Korean-English model

bilingualchatcode-generation
Context
33K
License
EXAONE AI Model
Microsoft
2024-12-12

Phi-4

Latest Phi model with state-of-the-art reasoning

code-generationreasoningmath+1
Context
16K
Input/1M
$0.07
Google
2024-12-11

Gemini 2.0 Flash

Next-generation multimodal model with native tool use and agentic capabilities

code-generationmultimodaltool-use+1
Context
1.0M
Input/1M
$0.075
Technology Innovation Institute
2024-12-11

Falcon 3 10B

Latest Falcon 3 model for efficient deployment

code-generationreasoninganalysis
Context
33K
License
Falcon-3
Meta
2024-12-06

Llama 3.3 70B

Open-weight multilingual model matching Llama 3.1 405B performance

code-generationcode-translationmultilingual
Context
128K
License
Llama 3.3 Community License
OpenAI
2024-12-05

o1

Reasoning model designed to solve hard problems across domains using chain-of-thought

reasoningcode-generationcomplex-analysis+1
Context
200K
Input/1M
$15
OpenAI
2024-12-05

o1 Pro

Pro version of o1 with extended compute for harder problems

deep-reasoningcomplex-analysiscode-generation+1
Context
200K
Input/1M
$150
Amazon
2024-12-03

Amazon Nova Micro

Fastest and most cost-effective Nova model

text-generationclassificationchat
Context
128K
Input/1M
$0.035
Amazon
2024-12-03

Amazon Nova Lite

Multimodal Nova model for image and video understanding

multimodalvisionvideo-understanding+1
Context
300K
Input/1M
$0.06
Amazon
2024-12-03

Amazon Nova Pro

Balanced Nova model for most tasks

multimodalreasoningcode-generation+1
Context
300K
Input/1M
$0.8
Alibaba
2024-11-27

QwQ 32B

Reasoning-focused model from Qwen family

deep-reasoningmathcode-generation+1
Context
33K
License
Apache 2.0
Skywork
2024-11-25

Skywork o1 Open 8B

Open reasoning model following o1 methodology

reasoningmathcode-generation
Context
33K
License
Apache 2.0
Allen AI
2024-11-22

Tulu 3 405B

Fine-tuned Llama 3.1 405B for instruction following

code-generationreasoninginstruction-following
Context
128K
License
ODC-BY
Allen AI
2024-11-22

Tulu 3 70B

Efficient Tulu model for balanced tasks

code-generationreasoninganalysis
Context
128K
License
ODC-BY
Alibaba
2024-11-22

Marco-o1

Reasoning model inspired by o1 methodology

reasoningmathcode-generation+1
Context
33K
License
Apache 2.0
Mistral AI
2024-11-18

Pixtral Large

Large multimodal model for complex visual tasks

visioncode-generationreasoning+1
Context
128K
License
Open
Mistral AI
2024-11-18

Mistral Large 2411

Latest Mistral Large with system prompt improvements

code-generationreasoningfunction-calling+1
Context
128K
Input/1M
$2
Nexusflow
2024-11-12

Athene V2 Chat 72B

Qwen-based model optimized for chat and reasoning

chatreasoningcode-generation+1
Context
131K
License
Apache 2.0
Alibaba
2024-11-11

Qwen 2.5 Coder 32B

State-of-the-art open code model rivaling GPT-4o on coding tasks

code-generationcode-completioncode-reasoning
Context
131K
License
Apache 2.0
Alibaba
2024-11-11

Qwen 2.5 Coder 7B

Efficient coding model from Qwen 2.5 family

code-generationcode-completioncode-reasoning
Context
131K
License
Apache 2.0
Alibaba
2024-11-11

Qwen Coder 2.5 32B

Alibaba's specialized coding model with strong code understanding capabilities

code-generationcode-reviewdebugging+1
Context
131K
License
Apache 2.0
Alibaba
2024-11-11

Qwen Coder 2.5 14B

Balanced code model with strong performance and reasonable resource requirements

code-generationcode-reviewdebugging
Context
131K
License
Apache 2.0
Alibaba
2024-11-11

Qwen Coder 2.5 7B

Efficient code model for quick tasks and resource-constrained environments

code-generationcode-completion
Context
131K
License
Apache 2.0
Infinigence AI
2024-11-06

Megrez 3B

Efficient model designed for edge deployment

code-generationchatreasoning
Context
128K
License
Apache 2.0
Tencent
2024-11-05

Hunyuan-Large

Tencent's large MoE model

code-generationreasoninganalysis+1
Context
256K
License
Tencent Hunyuan
Anthropic
2024-11-04

Claude 3.5 Haiku

Fast and affordable model for high-volume tasks

code-generationcode-translationquick-analysis
Context
200K
Input/1M
$0.8
Allen AI
2024-11-04

OLMo 2 13B

Fully open model with training data available

text-generationreasoningresearch
Context
4K
License
Apache 2.0
Allen AI
2024-11-04

OLMo 2 7B

Efficient fully open model

text-generationanalysisresearch
Context
4K
License
Apache 2.0
Hugging Face
2024-11-01

SmolLM2 1.7B

Compact model for on-device deployment

chatcode-generationsummarization
Context
8K
License
Apache 2.0
Hugging Face
2024-11-01

SmolLM2 360M

Tiny model for ultra-constrained environments

chatclassificationsimple-tasks
Context
8K
License
Apache 2.0
Codeium
2024-11-01

Windsurf Cascade

Agentic AI for autonomous coding with deep codebase understanding

agentic-codingcodebase-understandingrefactoring
Context
100K
License
Open
Recraft
2024-10-29

Recraft V3

Professional image generation for design

image-generationvector-graphicsdesign
Context
N/A
License
Open
Cohere
2024-10-23

Aya Expanse 32B

Multilingual model supporting 23 languages

multilingualcode-generationtranslation
Context
128K
License
CC-BY-NC-4.0
Cohere
2024-10-23

Aya Expanse 8B

Efficient multilingual model

multilingualchattranslation
Context
128K
License
CC-BY-NC-4.0
Anthropic
2024-10-22

Claude 3.5 Sonnet

Most intelligent Claude model, excels at coding and complex reasoning

code-generationcode-translationanalysis+2
Context
200K
Input/1M
$3
Stability AI
2024-10-22

Stable Diffusion 3.5

Latest text-to-image generation model

image-generationtext-to-image
Context
N/A
License
Stability AI Community
Anthropic
2024-10-22

Claude Computer Use

Claude model specialized for computer control and automation

computer-useui-automationbrowser-control+1
Context
200K
Input/1M
$3
IBM
2024-10-21

Granite 3 8B

IBM's efficient enterprise model

code-generationreasoningenterprise+1
Context
128K
License
Apache 2.0
IBM
2024-10-21

Granite 3 2B

Compact IBM model for edge deployment

code-generationchatenterprise
Context
128K
License
Apache 2.0
Mistral AI
2024-10-16

Ministral 8B

Edge-focused model for on-device deployment

code-generationreasoningedge-ai
Context
128K
Input/1M
$0.11
Mistral AI
2024-10-16

Ministral 3B

Smallest Ministral for ultra-efficient tasks

chatclassificationsimple-tasks
Context
128K
License
Mistral Research
NVIDIA
2024-10-11

Llama 3.1 Nemotron 70B

NVIDIA-optimized Llama 3.1 for enterprise

code-generationreasoningenterprise
Context
128K
License
Llama 3.1 Community
Udio
2024-10-10

Udio v1.5

Music generation with high fidelity

music-generationaudio-creation
Context
N/A
License
Open
OpenAI
2024-10-01

Whisper Large v3 Turbo

Fast speech recognition model

speech-to-textfast-transcriptionmultilingual
Context
N/A
License
MIT
Black Forest Labs
2024-10-01

FLUX 1.1 Pro

High-quality image generation model

image-generationtext-to-imagefast-generation
Context
N/A
License
Open
Meta
2024-09-25

Llama 3.2 1B

Tiny Llama model for edge and mobile deployment

summarizationinstruction-followingchat
Context
128K
License
Llama 3.2 Community
Meta
2024-09-25

Llama 3.2 3B

Compact Llama model for efficient deployment

code-generationsummarizationchat
Context
128K
License
Llama 3.2 Community
Meta
2024-09-25

Llama 3.2 11B Vision

Multimodal Llama with vision capabilities

visioncode-generationanalysis
Context
128K
License
Llama 3.2 Community
Meta
2024-09-25

Llama 3.2 90B Vision

Large multimodal Llama with vision

visioncode-generationanalysis+1
Context
128K
License
Llama 3.2 Community
Allen AI
2024-09-25

Molmo 72B

Multimodal model for vision and language tasks

visionimage-understandinganalysis
Context
4K
License
Apache 2.0
Meta
2024-09-25

Llama 3.2 Vision (General)

Multimodal Llama with image understanding

visioncode-generationanalysis
Context
128K
License
Llama 3.2 Community
Alibaba
2024-09-19

Qwen 2.5 72B

Largest Qwen 2.5 model for complex tasks

code-generationreasoningmultilingual+1
Context
131K
License
Apache 2.0
Alibaba
2024-09-19

Qwen 2.5 7B

Efficient Qwen 2.5 for everyday tasks

code-generationchatanalysis
Context
131K
License
Apache 2.0
Alibaba
2024-09-19

Qwen 2.5 14B

Mid-size Qwen 2.5 for balanced tasks

code-generationreasoninganalysis
Context
131K
License
Apache 2.0
Alibaba
2024-09-19

Qwen 2.5 32B

Large Qwen 2.5 for complex tasks

code-generationreasoninganalysis+1
Context
131K
License
Apache 2.0
Voyage AI
2024-09-18

Voyage 3

State-of-the-art embedding model

embeddingssemantic-searchrag
Context
32K
License
Open
Voyage AI
2024-09-18

Voyage Code 3

Code-specialized embedding model

code-embeddingscode-searchcode-similarity
Context
32K
License
Open
Jina AI
2024-09-18

Jina Embeddings v3

Multi-task embedding model with matryoshka support

embeddingsmultilingualmatryoshka
Context
8K
License
CC-BY-NC-4.0
Mistral AI
2024-09-17

Pixtral 12B

Multimodal model with vision capabilities

visioncode-generationimage-understanding
Context
128K
License
Apache 2.0
OpenAI
2024-09-12

o1-mini

Fast reasoning model optimized for coding, math, and science

reasoningcode-generationdebugging
Context
128K
Input/1M
$3
OpenAI
2024-09-12

o1-preview

Preview version of OpenAI's reasoning model

reasoningcode-generationmath+1
Context
128K
Input/1M
$15
HyperWrite
2024-09-05

Reflection 70B

Self-correcting model trained on synthetic data

self-correctionreasoningcode-generation
Context
8K
License
Apache 2.0
01.AI
2024-09-04

Yi Coder 9B

Efficient open code model with strong multilingual support

code-generationcode-completionmulti-language
Context
131K
License
Apache 2.0
01.AI
2024-09-04

Yi Coder 1.5B

Ultra-efficient code model for edge deployment and quick tasks

code-generationcode-completion
Context
131K
License
Apache 2.0
01.AI
2024-09-01

Yi Lightning

Fast Yi model for quick responses

chatcode-generationanalysis
Context
16K
License
Open
Cohere
2024-09-01

Command Code

Cohere's enterprise code model for development tasks

code-generationcode-reviewdebugging
Context
128K
Input/1M
$1
AI21 Labs
2024-08-22

Jamba 1.5 Large

Hybrid SSM-Transformer for long context

long-contextcode-generationreasoning
Context
256K
License
Jamba Open Model
AI21 Labs
2024-08-22

Jamba 1.5 Mini

Efficient hybrid model for quick tasks

long-contextcode-generationchat
Context
256K
License
Jamba Open Model
Nous Research
2024-08-20

Hermes 3 Llama 3.1 405B

Fine-tuned Llama 3.1 405B for instruction following

code-generationreasoningfunction-calling+1
Context
128K
Input/1M
$1
Nous Research
2024-08-20

Hermes 3 Llama 3.1 70B

Fine-tuned Llama 3.1 70B with enhanced capabilities

code-generationreasoningfunction-calling
Context
128K
Input/1M
$0.7
Ideogram
2024-08-20

Ideogram 2

Image model with excellent text rendering

image-generationtext-in-imagelogos
Context
N/A
License
Open
Hugging Face
2024-08-15

Parler TTS Large

Open-source controllable TTS

text-to-speechcontrollable-generation
Context
N/A
License
Apache 2.0
xAI
2024-08-13

Grok-2

Latest Grok model with frontier capabilities

code-generationreasoningvision+1
Context
128K
License
Open
xAI
2024-08-13

Grok-2 mini

Efficient Grok-2 variant for faster inference

code-generationreasoningchat
Context
128K
License
Open
Google
2024-08-13

Imagen 3

Google's latest image generation model

image-generationphotorealistictext-to-image
Context
N/A
License
Open
Cohere
2024-08-01

c4ai-command-r-08-2024

Latest Command R with RAG optimizations

ragcode-generationtool-use+1
Context
128K
License
CC-BY-NC-4.0
Black Forest Labs
2024-08-01

FLUX.1 [dev]

Open-weight image model for development

image-generationtext-to-image
Context
N/A
License
FLUX.1 [dev] Non-Commercial
Mistral AI
2024-07-24

Mistral Large 2

Flagship model with 128k context and function calling

code-generationfunction-callingmultilingual+1
Context
128K
Input/1M
$2
Meta
2024-07-23

Llama 3.1 405B

Largest open-weight model with frontier-class capabilities

code-generationcode-translationcomplex-reasoning
Context
128K
License
Llama 3.1 Community License
Meta
2024-07-23

Llama 3.1 8B

Extended context Llama 3.1 8B model

code-generationchatanalysis
Context
128K
License
Llama 3.1 Community
Meta
2024-07-23

Llama 3.1 70B

Extended context Llama 3.1 70B model

code-generationreasoninganalysis
Context
128K
License
Llama 3.1 Community
OpenAI
2024-07-18

GPT-4o Mini

Affordable small model for fast, lightweight tasks

code-generationcode-translationanalysis+1
Context
128K
Input/1M
$0.15
Mistral AI
2024-07-18

Mistral Nemo

Small but capable model for efficient deployment

code-generationreasoningmultilingual
Context
128K
Input/1M
$0.019
Mistral AI
2024-07-16

Codestral Mamba

Mamba-architecture code model for unlimited context

code-generationcode-completionlong-context
Context
N/A
License
Apache 2.0
ElevenLabs
2024-07-15

ElevenLabs Turbo v2.5

Fast text-to-speech model

text-to-speechvoice-cloningfast-synthesis
Context
N/A
License
Open
Zhipu AI
2024-07-05

CodeGeeX 4

Open-source multilingual code generation model with strong performance

code-generationcode-completionmulti-language+1
Context
33K
License
Apache 2.0
Google
2024-06-27

Gemma 2 27B

Open-weight model for research and development

code-generationanalysisreasoning
Context
8K
License
Gemma License
Google
2024-06-27

Gemma 2 9B

Efficient open-weight model for various tasks

code-generationanalysischat
Context
8K
License
Gemma License
Alibaba
2024-06-22

GTE-Qwen2-7B-instruct

High-performance embedding model based on Qwen2

embeddingslong-contextmultilingual
Context
33K
License
Apache 2.0
Suno
2024-06-21

Suno v3.5

AI music generation model

music-generationaudio-creationvocals
Context
N/A
License
Open
DeepSeek
2024-06-17

DeepSeek Coder V2

Code-specialized MoE model supporting 300+ languages

code-generationcode-completionmath+1
Context
128K
License
MIT
NVIDIA
2024-06-14

Nemotron-4 70B

NVIDIA's flagship model for enterprise

code-generationreasoninganalysis
Context
33K
License
NVIDIA Open Model
NVIDIA
2024-06-14

Nemotron-4 340B

Largest NVIDIA model for enterprise tasks

code-generationreasoninganalysis+1
Context
4K
License
NVIDIA Open Model
Alibaba
2024-06-07

Qwen 2 72B

Previous generation large Qwen model

code-generationreasoningmultilingual
Context
131K
License
Apache 2.0
Zhipu AI
2024-06-05

GLM-4 9B

Efficient bilingual model from GLM family

code-generationreasoningbilingual+1
Context
128K
License
GLM-4
CMU
2024-06-01

PolyCoder 16B

Open-source polyglot code model trained on many programming languages

code-generationmulti-language
Context
16K
License
Apache 2.0
Mistral AI
2024-05-29

Codestral

Specialized code model trained on 80+ programming languages

code-generationcode-completionfill-in-the-middle
Context
32K
Input/1M
$0.2
Microsoft
2024-05-21

Phi-3-small

Balanced Phi-3 model for diverse tasks

code-generationreasoninganalysis+1
Context
128K
License
MIT
Microsoft
2024-05-21

Phi-3-medium

Largest Phi-3 for complex reasoning

code-generationcomplex-reasoninganalysis
Context
128K
License
MIT
Google
2024-05-14

Gemini 1.5 Pro

Production-ready model with massive context window for complex tasks

code-generationcode-translationlong-context+1
Context
2.0M
Input/1M
$1.25
Google
2024-05-14

Gemini 1.5 Flash

Fast and versatile model for diverse tasks at scale

code-generationcode-translationanalysis
Context
1.0M
Input/1M
$0.075
OpenAI
2024-05-13

GPT-4o

Multimodal flagship model with vision and audio capabilities, optimized for speed and cost

code-generationcode-translationanalysis+2
Context
128K
Input/1M
$2.5
01.AI
2024-05-13

Yi 1.5 34B Chat

Enhanced Yi chat model with extended context

chatcode-generationreasoning
Context
16K
License
Apache 2.0
01.AI
2024-05-13

Yi Large

Flagship Yi model via API

code-generationreasoninganalysis
Context
33K
License
Open
DeepSeek
2024-05-06

DeepSeek V2

Efficient MoE model with strong general capabilities

code-generationreasoningchat+1
Context
128K
License
DeepSeek License
DeepSeek
2024-05-06

DeepSeek Chat

Optimized chat model for conversations

chatcode-generationanalysis
Context
33K
Input/1M
$0.2574
IBM
2024-05-06

Granite Code 34B

Code-specialized Granite model

code-generationcode-explanationcode-review
Context
8K
License
Apache 2.0
IBM
2024-05-06

Granite Code 20B

IBM's enterprise-focused code model with strong security awareness

code-generationcode-reviewsecurity-analysis
Context
8K
License
Apache 2.0
IBM
2024-05-06

Granite Code 8B

Efficient IBM code model for resource-constrained deployments

code-generationcode-completion
Context
8K
License
Apache 2.0
Amazon
2024-04-30

Amazon Q Developer

Next-gen AWS coding assistant with broad AWS service integration

code-generationcode-transformationsecurity-scanning+1
Context
33K
License
Open
Snowflake
2024-04-24

Snowflake Arctic

Enterprise-focused MoE model

sql-generationcode-generationanalysis
Context
4K
License
Apache 2.0
Amazon
2024-04-23

Amazon Titan Text Premier

Most capable Titan model for complex tasks

text-generationreasoningcode-generation
Context
32K
License
Open
Microsoft
2024-04-23

Phi-3-mini

Smallest Phi-3 model with strong capabilities

code-generationreasoningchat+1
Context
128K
License
MIT
Meta
2024-04-18

Llama 3 8B

Efficient Llama 3 model for everyday tasks

code-generationchatanalysis
Context
8K
License
Llama 3 Community
Meta
2024-04-18

Llama 3 70B

Large Llama 3 model for complex tasks

code-generationreasoninganalysis
Context
8K
License
Llama 3 Community
WizardLM
2024-04-15

WizardLM 2 8x22B

Large MoE wizard model for complex tasks

code-generationreasoninganalysis
Context
64K
Input/1M
$0.62
Alibaba
2024-04-15

CodeQwen 1.5 7B

Efficient code model based on Qwen 1.5 architecture

code-generationcode-completion
Context
66K
License
Apache 2.0
Snowflake
2024-04-11

Snowflake Arctic Embed L

Enterprise embedding model from Snowflake

embeddingssemantic-search
Context
512
License
Apache 2.0
Mistral AI
2024-04-10

Mixtral 8x22B

Large MoE model for complex tasks

code-generationreasoningmath+1
Context
66K
License
Apache 2.0
Google
2024-04-09

CodeGemma 7B

Code-specialized open model based on Gemma for programming tasks

code-generationcode-completioncode-infilling
Context
8K
License
Gemma Terms of Use
Cohere
2024-04-04

Command R+

Most capable Cohere model for complex tasks

ragcode-generationreasoning+2
Context
128K
License
CC-BY-NC-4.0
xAI
2024-03-28

Grok-1.5

Enhanced Grok with improved reasoning

code-generationreasoningmath+1
Context
128K
License
Open
Databricks
2024-03-27

DBRX

MoE model optimized for enterprise

code-generationreasoninganalysis+1
Context
33K
License
Databricks Open Model
xAI
2024-03-17

Grok-1

Original open-weight Grok model

code-generationreasoningchat
Context
8K
License
Apache 2.0
Anthropic
2024-03-14

Claude 3 Haiku

Fastest Claude 3 model for instant responses

code-generationanalysischat
Context
200K
Input/1M
$0.25
Cohere
2024-03-11

Command R

RAG-optimized model for enterprise search

ragcode-generationanalysis+1
Context
128K
License
CC-BY-NC-4.0
Mixed Bread
2024-03-07

mxbai-embed-large

High-quality embedding model

embeddingssemantic-searchclustering
Context
512
License
Apache 2.0
Anthropic
2024-03-04

Claude 3 Opus

Powerful model for complex tasks requiring deep expertise

code-generationcomplex-reasoninganalysis+1
Context
200K
Input/1M
$15
Anthropic
2024-03-04

Claude 3 Sonnet

Balanced Claude 3 model for enterprise tasks

code-generationanalysisreasoning
Context
200K
Input/1M
$3
TabNine
2024-03-01

TabNine Enterprise

Enterprise AI code completion with custom model training

code-completioncode-generationcustom-training
Context
33K
License
Open
BigCode
2024-02-28

StarCoder2 15B

Code-focused model trained on The Stack v2

code-generationcode-completioncode-explanation
Context
16K
License
BigCode OpenRAIL-M
BigCode
2024-02-28

StarCoder2 7B

Efficient code model for development

code-generationcode-completion
Context
16K
License
BigCode OpenRAIL-M
BigCode
2024-02-28

StarCoder2 3B

Compact code model for edge deployment

code-generationcode-completion
Context
16K
License
BigCode OpenRAIL-M
Mistral AI
2024-02-26

Mistral Small

Cost-effective model for simple tasks

code-generationchatclassification
Context
33K
Input/1M
$1
Google
2024-02-21

Gemma 7B

Original Gemma model for lightweight tasks

code-generationchatanalysis
Context
8K
License
Gemma License
Google
2024-02-08

Gemini Ultra

Most capable Gemini model for complex tasks

code-generationreasoningmultimodal+1
Context
128K
License
Open
Alibaba
2024-02-05

Qwen 1.5 72B

Older Qwen model for compatibility

code-generationchatanalysis
Context
33K
License
Qianwen License
Nomic AI
2024-02-01

Nomic Embed Text

Open-source text embedding model

embeddingslong-context-embeddings
Context
8K
License
Apache 2.0
Supermaven
2024-02-01

Supermaven

Ultra-fast AI code completion with 1M token context

code-completioncode-generationlarge-context
Context
1.0M
License
Open
Alibaba
2024-01-30

Qwen Max

Most capable Qwen via API

code-generationreasoninganalysis+1
Context
33K
License
Open
Alibaba
2024-01-30

Qwen Plus

Balanced Qwen model via API

code-generationchatanalysis
Context
33K
Input/1M
$0.26
Alibaba
2024-01-30

Qwen Turbo

Fast Qwen model for quick tasks

chatcode-generationsummarization
Context
8K
License
Open
BAAI
2024-01-30

BGE-M3

Multi-lingual, multi-functionality embedding model

embeddingsmultilingualdense-sparse-retrieval
Context
8K
License
MIT
Meta
2024-01-29

Code Llama 70B

Specialized code model fine-tuned from Llama 2 for programming tasks

code-generationcode-completioninfilling
Context
100K
License
Llama 2 Community License
Meta
2024-01-29

Code Llama 70B Instruct

Instruction-tuned Code Llama for following complex coding instructions

code-generationcode-reviewdebugging+1
Context
16K
License
Llama 2 Community License
OpenAI
2024-01-25

text-embedding-3-large

OpenAI's latest embedding model

embeddingssemantic-searchflexible-dimensions
Context
8K
License
Open
OpenAI
2024-01-25

text-embedding-3-small

Efficient OpenAI embedding model

embeddingssemantic-search
Context
8K
License
Open
Shanghai AI Laboratory
2024-01-17

InternLM 2 20B

Bilingual model with strong reasoning

code-generationreasoningmath+1
Context
200K
License
Apache 2.0
Sourcegraph
2024-01-15

Sourcegraph Cody

AI coding assistant with deep codebase understanding

code-generationcode-searchcode-explanation+1
Context
100K
License
Open
Stability AI
2024-01-09

Stable Code 3B

Lightweight code model optimized for fast inference and local deployment

code-generationcode-completion
Context
16K
License
Stability AI Non-Commercial Research Community License
Mistral AI
2024-01-08

Mistral Medium

Balanced model for diverse tasks

code-generationreasoninganalysis
Context
33K
Input/1M
$2.7
Mistral AI
2024-01-08

Mistral Embed

Embedding model for semantic search

embeddingssemantic-searchrag
Context
8K
License
Open
Cursor
2024-01-01

Cursor AI

AI-native code editor with advanced code understanding

code-generationcode-reviewcodebase-understanding+1
Context
128K
License
Open
Microsoft
2023-12-19

E5-Mistral-7B-Instruct

Instruction-following embedding model

embeddingsinstruction-followingsemantic-search
Context
4K
License
MIT
Upstage
2023-12-13

SOLAR 10.7B

Depth-upscaled model with strong performance

code-generationreasoningchat
Context
4K
License
Apache 2.0
Microsoft
2023-12-12

Phi-2

Small but capable model rivaling larger ones

code-generationreasoningmath+1
Context
2K
License
MIT
Mistral AI
2023-12-11

Mixtral 8x7B

Mixture-of-experts model with efficient inference

code-generationreasoningmultilingual
Context
33K
License
Apache 2.0
Meta
2023-12-08

SeamlessM4T v2

Multilingual speech and text translation

speech-to-texttext-to-speechtranslation
Context
N/A
License
CC-BY-NC-4.0
Google
2023-12-06

Gemini 1.0 Pro

Original Gemini Pro model for general tasks

code-generationanalysischat
Context
32K
Input/1M
$0.5
ise-uiuc
2023-12-04

Magicoder S-DS 6.7B

Efficient code model trained with OSS-Instruct methodology

code-generationcode-completion
Context
16K
License
Apache 2.0
OpenAI
2023-11-06

GPT-4 Turbo

Enhanced GPT-4 with 128K context and improved performance

code-generationvisionreasoning+1
Context
128K
Input/1M
$10
01.AI
2023-11-06

Yi 34B

Large bilingual model from Yi series

code-generationreasoningbilingual
Context
4K
License
Apache 2.0
01.AI
2023-11-06

Yi 6B

Efficient Yi model for lighter tasks

code-generationchatbilingual
Context
4K
License
Apache 2.0
OpenAI
2023-11-06

Whisper Large v3

Speech recognition model for transcription

speech-to-texttranscriptionmultilingual
Context
N/A
License
MIT
Cohere
2023-11-02

Cohere Embed v3

Enterprise-grade embedding model

embeddingsmultilingualcompression
Context
512
License
Open
OpenChat
2023-11-01

OpenChat 3.5

Open chat model with RLHF training

chatcode-generationreasoning
Context
8K
License
Apache 2.0
DeepSeek
2023-11-01

DeepSeek Coder 33B Instruct

Instruction-tuned DeepSeek coding model for following coding instructions

code-generationinstruction-followingdebugging
Context
16K
License
DeepSeek License
Refact AI
2023-10-12

Refact 1.6B

Ultra-efficient code model for real-time code completion

code-completioncode-infilling
Context
4K
License
BigScience OpenRAIL-M
OpenAI
2023-10-03

DALL-E 3

OpenAI's latest image generation model

image-generationtext-to-imageprompt-understanding
Context
N/A
License
Open
Amazon
2023-09-28

Amazon Titan Text Express

Fast and cost-effective model for general tasks

text-generationsummarizationchat
Context
8K
License
Open
Amazon
2023-09-28

Amazon Titan Text Lite

Lightweight model for cost-sensitive applications

text-generationsummarization
Context
4K
License
Open
Mistral AI
2023-09-27

Mistral 7B

Efficient base model with sliding window attention

code-generationchatanalysis
Context
33K
License
Apache 2.0
Microsoft
2023-09-11

Phi-1.5

Enhanced Phi with improved reasoning

code-generationreasoningchat
Context
2K
License
MIT
Technology Innovation Institute
2023-09-06

Falcon 180B

Largest open Falcon model

text-generationreasoninganalysis
Context
2K
License
Falcon-180B TII
Baichuan
2023-09-06

Baichuan 2 13B

Chinese-focused large language model

text-generationcode-generationchinese
Context
4K
License
Baichuan 2
WizardLM
2023-08-26

WizardCoder 34B

Instruction-following code model with strong complex task performance

code-generationcode-reviewdebugging
Context
8K
License
Llama 2 Community License
Phind
2023-08-25

Phind CodeLlama 34B

Fine-tuned Code Llama optimized for code generation and explanation

code-generationcode-explanationdebugging
Context
16K
License
Llama 2 Community License
Meta
2023-08-24

Code Llama 7B

Code-specialized Llama model for development

code-generationcode-completioncode-infilling
Context
100K
License
Llama 2 Community
Meta
2023-08-24

Code Llama 13B

Mid-size code-specialized Llama model

code-generationcode-completiondebugging
Context
100K
License
Llama 2 Community
Meta
2023-08-24

Code Llama 34B

Large code-specialized Llama model

code-generationcode-reasoningdebugging
Context
100K
License
Llama 2 Community
Meta
2023-08-24

Code Llama Instruct 34B

Instruction-tuned Code Llama for complex tasks

code-generationinstruction-followingdebugging
Context
16K
License
Llama 2 Community
Meta
2023-08-24

Code Llama Python 34B

Python-specialized Code Llama model

python-generationcode-completionpython-debugging
Context
16K
License
Llama 2 Community
Meta
2023-08-24

Code Llama 34B Instruct

Efficient instruction-tuned Code Llama for coding tasks

code-generationcode-reviewdebugging
Context
16K
License
Llama 2 Community License
Replit
2023-07-26

Replit Code V1.5 3B

Efficient code model trained on Replit's diverse codebase

code-generationcode-completion
Context
4K
License
CC BY-SA 4.0
Meta
2023-07-18

Llama 2 7B

Previous generation efficient Llama model

chatcode-generationanalysis
Context
4K
License
Llama 2 Community
Meta
2023-07-18

Llama 2 13B

Mid-size previous generation Llama model

chatcode-generationreasoning
Context
4K
License
Llama 2 Community
Meta
2023-07-18

Llama 2 70B

Largest previous generation Llama model

code-generationreasoninganalysis
Context
4K
License
Llama 2 Community
MosaicML
2023-06-22

MPT-30B

Commercial-friendly open model

text-generationcode-generationanalysis
Context
8K
License
Apache 2.0
Microsoft
2023-06-20

Phi-1

First Phi model focused on coding

code-generationcode-completion
Context
2K
License
MIT
Aider
2023-06-01

Aider

AI pair programming tool for terminal with git integration

code-generationcode-refactoringgit-integration
Context
128K
License
Apache 2.0
Google
2023-06-01

Codey

Google's code-specialized model for enterprise development

code-generationcode-completioncode-explanation
Context
33K
License
Open
Technology Innovation Institute
2023-05-25

Falcon 40B

Mid-size Falcon model

text-generationcode-generationanalysis
Context
2K
License
Apache 2.0
Technology Innovation Institute
2023-05-25

Falcon 7B

Efficient Falcon model

text-generationchat
Context
2K
License
Apache 2.0
Google
2023-05-10

PaLM 2

Google's previous generation foundation model

code-generationreasoningmultilingual
Context
32K
License
Open
Continue
2023-05-01

Continue

Open-source AI code assistant supporting multiple models

code-generationcode-reviewchat+1
Context
128K
License
Apache 2.0
Amazon
2023-04-13

Amazon CodeWhisperer

AWS-native AI coding assistant with security scanning

code-generationcode-completionsecurity-scanning
Context
8K
License
Open
Suno
2023-04-10

Bark

Open-source text-to-audio model

text-to-speechmusicsound-effects
Context
N/A
License
MIT
OpenAI
2023-03-14

GPT-4

Original GPT-4 model with strong reasoning and coding capabilities

code-generationreasoninganalysis+1
Context
8K
Input/1M
$30
Cohere
2023-03-01

Command Light

Lightweight model for simple tasks

text-generationsummarizationclassification
Context
4K
License
Open
Cohere
2023-03-01

Command

General-purpose instruction-following model

text-generationsummarizationclassification+1
Context
4K
License
Open
BigCode
2023-01-01

SantaCoder

Efficient code model trained on Python, Java, and JavaScript

code-generationcode-completion
Context
2K
License
BigCode OpenRAIL-M
OpenAI
2022-11-30

GPT-3.5 Turbo

Fast and cost-effective model for everyday tasks

code-generationchatanalysis
Context
16K
Input/1M
$0.5
Codeium
2022-11-01

Codeium

Free AI code completion with broad IDE support

code-completioncode-generationchat
Context
33K
License
Open
BigScience
2022-07-06

BLOOM

Multilingual open model supporting 46 languages

text-generationmultilingualcode-generation
Context
2K
License
BigScience RAIL
GitHub
2022-06-21

GitHub Copilot

AI pair programmer powered by OpenAI with deep GitHub integration

code-generationcode-completioncode-review+1
Context
8K
Input/1M
$0
Meta
2022-04-01

InCoder 6B

Infilling-capable code model for completion and generation

code-generationcode-infilling
Context
2K
License
Apache 2.0
OpenAI
2021-08-10

OpenAI Codex

OpenAI's code model powering GitHub Copilot

code-generationcode-completioncode-translation
Context
8K
License
Open