Openrouter
60 items tagged with "openrouter"
Models60
DeepSeek Pro Latest
A hosted DeepSeek Pro-class model endpoint added to OpenRouter with a 1,048,576-token context window. It is intended for long-context text generation and general-purpose assistant workloads.
DeepSeek Flash Latest
A hosted DeepSeek Flash-class model endpoint added to OpenRouter with a 1,048,576-token context window. It targets long-context text generation in a faster Flash-tier model line.
Schematron v2 Turbo
A 128k-context hosted language model in the Schematron v2 family, added to OpenRouter on 2026-09-12. The available data verifies its model ID and context length but does not provide benchmark, modality, or pricing details.
Schematron v2 Small
A 128k-context hosted language model in the Schematron v2 family, added to OpenRouter on 2026-09-12. The available data verifies its model ID and context length but does not provide benchmark, modality, or pricing details.
Fugu Ultra v2
A Sakana AI hosted foundation model with a 1,000,000-token context window, added to OpenRouter on 2026-09-11. It appears to be a new v2 update in the Fugu Ultra line, with available data verifying context length but not pricing or benchmark details.
Fugu Max
A Sakana AI hosted foundation model with a 1,000,000-token context window, added to OpenRouter on 2026-09-11. The OpenRouter data verifies it as a new Fugu-family model with very long context support, but does not provide pricing or benchmark details.
DeepSeek V4.1 Flash
DeepSeek V4.1 Flash is a long-context foundation model from DeepSeek, newly listed on OpenRouter and also present in the Ollama library. It is positioned as a fast general-purpose model with a 1,048,576-token context window.
Ling 3.0 Flash VL
Ling 3.0 Flash VL is a vision-language variant of InclusionAI's Ling 3.0 Flash family, newly added to OpenRouter. It supports multimodal use cases with a 262,144-token context window.
Mercury 2.5
Hosted language model added to OpenRouter with a 260k-token context window.
NEX N2.5 Pro
Free hosted NEX AGI language model added to OpenRouter with a 262,144-token context window.
NEX N2.5 Mini
Free hosted compact NEX AGI language model added to OpenRouter with a 262,144-token context window.
Muse Spark 1.3
Muse Spark 1.3 is a Meta hosted foundation model newly added to OpenRouter with a 1,048,576-token context window. It appears to be the next base release in the Muse Spark series, suited for long-context general-purpose generation.
Gemini 3.8 Flash
Gemini 3.8 Flash is a Google Gemini Flash-series model newly added to OpenRouter with a 1,048,576-token context window. It is positioned as a fast, large-context general-purpose model.
Mercury 2.5 Preview
Mercury 2.5 Preview is a hosted preview model added to OpenRouter under the inception namespace. It provides a large 260k-token context window for long-context text tasks.
Granite 4.2 8B
Granite 4.2 8B is an IBM Granite open-weight foundation model added to OpenRouter with a 131k-token context window. It is also present in the Ollama library as granite4.2 for local deployment.
HY4 Preview
Tencent HY4 Preview is a newly added hosted foundation model on OpenRouter with a 1,048,576-token context window. It appears to target general long-context chat and text generation workloads.
Ling 3.0 Flash Fin
A newly added hosted Ling 3.0 Flash financial-domain model on OpenRouter with a 262,144-token context window. It appears intended for long-context finance-oriented text generation and analysis tasks.
Qwen3.8 Flash
Qwen3.8 Flash is an Alibaba Qwen model newly added on OpenRouter with a 1,000,000-token context window. It is positioned as a fast, long-context general-purpose model.
GLM-5.3-Flash
GLM-5.3-Flash is a Zhipu AI / Z.ai GLM model newly added on OpenRouter with a 1,310,720-token context window. It is also present in the Ollama library, indicating local open-weight availability.
DeepSeek V4 Flash Vision Exp
Experimental vision-capable variant of DeepSeek V4 Flash available on OpenRouter, adding multimodal image understanding to the Flash model line with a 1M-token context window.
HY-MT2-1.8B
Tencent HY-MT2-1.8B is a newly listed hosted text model on OpenRouter with an 8K-token context window. It appears to be part of Tencent's HY-MT2 model family.
HY-MT2-30B-A3B
Tencent HY-MT2-30B-A3B is a newly listed hosted text model on OpenRouter with an 8K-token context window. It is the larger HY-MT2 variant surfaced in the discovery data.
HY-MT2 7B
Tencent's 7B-parameter HY-MT2 model made available on OpenRouter with an 8K-token context window. It appears to be part of Tencent's HY-MT2 model family, positioned for multilingual text and translation-style workloads.
GLM-5.3
GLM-5.3 is a Zhipu AI / Z.ai foundation model newly added on OpenRouter. The provided data verifies a very large 1,048,576-token context window, making it suitable for long-context text tasks.
Dots 3 Note Preview
OpenRouter-listed free preview model from Dots Studio with a 512k-token context window. The discovery data identifies it as a recent preview model suitable for long-context text workflows.
Seed 2.1 Turbo
Seed 2.1 Turbo is a newly listed ByteDance Seed hosted model on OpenRouter with a 262,144-token context window. The Turbo designation indicates a speed-oriented variant for general-purpose long-context workloads.
Sakana Namazu
Sakana Namazu is a Sakana AI model newly listed on OpenRouter with a 262k-token context window.
Solar Pro4
Upstage Solar Pro4 is a newly listed long-context hosted foundation model on OpenRouter, supporting a 524k-token context window.
Muse Glimmer 30B
Meta's Muse Glimmer 30B is a 30B-parameter open-weight model newly listed on OpenRouter with a 131k-token context window and also present in the Ollama library.
Ling 3.0 Tiny
A smaller Ling 3.0-series hosted foundation model added to OpenRouter with a 262,144-token context window. The OpenRouter listing identifies it as a free-access model variant.
Muse Spark 1.2
A Meta hosted foundation model added to OpenRouter with a 1,048,576-token context window. It is a newer Muse Spark release than the existing 1.1 model.
Qwen3.8 Max
Alibaba’s Qwen3.8 Max is a hosted Qwen model added to OpenRouter with a 1,000,000-token context window. It is positioned for long-context general-purpose language tasks, reasoning, and generation.
Inkling Small
Inkling Small is a hosted 524K-context model from Thinking Machines, added on OpenRouter as a smaller variant of the Inkling model family. It is suited for long-context text and reasoning workloads where a lighter model is preferred.
Qwen3.7 Flash
Qwen3.7 Flash is a fast hosted Qwen model variant added to OpenRouter with a 1,000,000-token context window. It is suited for long-context text, reasoning, and general assistant workloads where lower latency is important.
Ling 3.0 Flash
Ling 3.0 Flash is a hosted general-purpose language model from InclusionAI, newly listed on OpenRouter with a 262K-token context window. The Flash variant is positioned for fast, long-context chat and reasoning workloads.
Laguna S 2.1
Laguna S 2.1 is a long-context foundation model from Poolside, newly listed on OpenRouter and also present in the Ollama library. The OpenRouter listing reports a 1,048,576-token context window.
Gemini 3.6 Flash
Gemini 3.6 Flash is a Google Gemini model newly added to OpenRouter with a 1,048,576-token context window. It is positioned as a Flash-family hosted model for long-context text generation and reasoning workloads.
Gemini 3.5 Flash-Lite
Gemini 3.5 Flash-Lite is a lightweight Google Gemini hosted model newly added to OpenRouter with a 1,048,576-token context window. It is suited to cost- and latency-sensitive long-context text generation workloads.
LongCat 2.0
LongCat 2.0 is a Meituan long-context foundation model newly added to OpenRouter. The listing reports a 1,048,756-token context window for hosted text-generation workloads.
Inkling
Inkling is a hosted model newly listed on OpenRouter with a 1,048,576-token context window. The discovery data verifies its availability and long-context capability, but does not provide further architectural or benchmark details.
Kimi K3
Kimi K3 is a Moonshot AI long-context hosted language model added to OpenRouter with a 1,048,576-token context window. It is positioned for large-context text, reasoning, and coding workloads.
Muse Spark 1.1
Muse Spark 1.1 is a Meta model added to OpenRouter with a 1,048,576-token context window. The discovery data verifies it as a newly listed long-context hosted model.
KAT Coder Air v2.5
KAT Coder Air v2.5 is a code-focused model from the KwaiPilot namespace added to OpenRouter with a 256,000-token context window. It appears to be the lighter Air variant of the KAT Coder v2.5 family.
KAT Coder Pro v2.5
KAT Coder Pro v2.5 is a code-focused model from the KwaiPilot namespace added to OpenRouter with a 256,000-token context window. It appears to be the higher-capability Pro variant of the KAT Coder v2.5 family.
GPT-5.6 Luna Pro
OpenRouter-listed GPT-5.6 hosted model variant with a 1,050,000-token context window. It is positioned as a long-context frontier text model for demanding reasoning and productivity workloads.
GPT-5.6 Luna
OpenRouter-listed GPT-5.6 hosted model variant with a 1,050,000-token context window. It is intended for broad long-context text, coding, and reasoning use cases.
GPT-5.6 Terra
OpenRouter-listed GPT-5.6 hosted model variant with a 1,050,000-token context window. It targets long-context general-purpose assistance, reasoning, and code-generation tasks.
Grok 4.5
xAI hosted Grok model newly added on OpenRouter with a 500k-token context window for long-context general AI tasks.
Aion 3.0
Aion Labs hosted general-purpose model newly added on OpenRouter with a 131k-token context window.
Aion 3.0 Mini
Smaller Aion Labs hosted model newly added on OpenRouter with a 131k-token context window.
Tencent HY3
Tencent long-context hosted model newly added on OpenRouter with a 262k-token context window.
Laguna XS 2.1
Laguna XS 2.1 is a Poolside model added to OpenRouter with a 262K-token context window and also available in the Ollama library. It is suited to long-context coding and text workflows.
Gemini 3.1 Flash-Lite Image
A Google Gemini 3.1 Flash-Lite image model added to OpenRouter, providing image-focused multimodal capabilities with a 65,536-token context window.
Fugu Ultra
Fugu Ultra is a Sakana model listed on OpenRouter with a 1M-token context window. It is positioned for long-context general-purpose reasoning and text generation.
Gemini 3.1 Flash Image
Google Gemini image-focused model listed on OpenRouter with a 131,072-token context window. It is positioned as a Flash-tier multimodal/image model for lower-latency image-centric workloads.
GLM-5.2
GLM-5.2 is a Z.ai / Zhipu AI foundation model listed on OpenRouter with a 1,048,576-token context window and available in the Ollama library. It targets long-context reasoning, generation, and coding workloads.
Kimi K2.7 Code
Kimi K2.7 Code is a Moonshot AI coding-focused model listed on OpenRouter with a 262K-token context window and available in the Ollama library. It is aimed at software engineering and long-context code understanding tasks.
Claude Fable 5
Claude Fable 5 is a newly listed Anthropic model on OpenRouter with a 1,000,000-token context window. The discovery data verifies it as a new Anthropic model added during the target date range.
NVIDIA Nemotron 3 Ultra 550B A55B
NVIDIA Nemotron 3 Ultra 550B A55B is a large open-weight Nemotron-family model listed on OpenRouter with a 1M-token context window and available in the Ollama library. It is intended for long-context reasoning and general text-generation workloads.
Qwen3.7 Plus
A Qwen-family large language model added on OpenRouter with a 1,000,000-token context window for long-context general-purpose AI workloads.