Skip to main content
Back to Tags

Text Generation

68 items tagged with "text-generation"

Filter by type:

Models68

Model

NVIDIA Nemotron 3 Super

NVIDIA Nemotron 3 Super is a foundation model referenced as the base for Salesforce's Koa CRM reasoning model and is available in the Ollama library for local use.

Model

DeepSeek Pro Latest

A hosted DeepSeek Pro-class model endpoint added to OpenRouter with a 1,048,576-token context window. It is intended for long-context text generation and general-purpose assistant workloads.

Model

DeepSeek Flash Latest

A hosted DeepSeek Flash-class model endpoint added to OpenRouter with a 1,048,576-token context window. It targets long-context text generation in a faster Flash-tier model line.

Model

Schematron v2 Turbo

A 128k-context hosted language model in the Schematron v2 family, added to OpenRouter on 2026-09-12. The available data verifies its model ID and context length but does not provide benchmark, modality, or pricing details.

Model

Schematron v2 Small

A 128k-context hosted language model in the Schematron v2 family, added to OpenRouter on 2026-09-12. The available data verifies its model ID and context length but does not provide benchmark, modality, or pricing details.

Model

Fugu Ultra v2

A Sakana AI hosted foundation model with a 1,000,000-token context window, added to OpenRouter on 2026-09-11. It appears to be a new v2 update in the Fugu Ultra line, with available data verifying context length but not pricing or benchmark details.

Model

Fugu Max

A Sakana AI hosted foundation model with a 1,000,000-token context window, added to OpenRouter on 2026-09-11. The OpenRouter data verifies it as a new Fugu-family model with very long context support, but does not provide pricing or benchmark details.

Model

DeepSeek V4.1 Flash

DeepSeek V4.1 Flash is a long-context foundation model from DeepSeek, newly listed on OpenRouter and also present in the Ollama library. It is positioned as a fast general-purpose model with a 1,048,576-token context window.

Model

Ling 3.0 Flash VL

Ling 3.0 Flash VL is a vision-language variant of InclusionAI's Ling 3.0 Flash family, newly added to OpenRouter. It supports multimodal use cases with a 262,144-token context window.

Model

Mercury 2.5

Hosted language model added to OpenRouter with a 260k-token context window.

Model

NEX N2.5 Pro

Free hosted NEX AGI language model added to OpenRouter with a 262,144-token context window.

Model

NEX N2.5 Mini

Free hosted compact NEX AGI language model added to OpenRouter with a 262,144-token context window.

Model

Muse Spark 1.3

Muse Spark 1.3 is a Meta hosted foundation model newly added to OpenRouter with a 1,048,576-token context window. It appears to be the next base release in the Muse Spark series, suited for long-context general-purpose generation.

Model

Gemini 3.8 Flash

Gemini 3.8 Flash is a Google Gemini Flash-series model newly added to OpenRouter with a 1,048,576-token context window. It is positioned as a fast, large-context general-purpose model.

Model

Mercury 2.5 Preview

Mercury 2.5 Preview is a hosted preview model added to OpenRouter under the inception namespace. It provides a large 260k-token context window for long-context text tasks.

Model

HY4 Preview

Tencent HY4 Preview is a newly added hosted foundation model on OpenRouter with a 1,048,576-token context window. It appears to target general long-context chat and text generation workloads.

Model

Ling 3.0 Flash Fin

A newly added hosted Ling 3.0 Flash financial-domain model on OpenRouter with a 262,144-token context window. It appears intended for long-context finance-oriented text generation and analysis tasks.

Model

Qwen3.8 Flash

Qwen3.8 Flash is an Alibaba Qwen model newly added on OpenRouter with a 1,000,000-token context window. It is positioned as a fast, long-context general-purpose model.

Model

GLM-5.3-Flash

GLM-5.3-Flash is a Zhipu AI / Z.ai GLM model newly added on OpenRouter with a 1,310,720-token context window. It is also present in the Ollama library, indicating local open-weight availability.

Model

HY-MT2-1.8B

Tencent HY-MT2-1.8B is a newly listed hosted text model on OpenRouter with an 8K-token context window. It appears to be part of Tencent's HY-MT2 model family.

Model

HY-MT2-30B-A3B

Tencent HY-MT2-30B-A3B is a newly listed hosted text model on OpenRouter with an 8K-token context window. It is the larger HY-MT2 variant surfaced in the discovery data.

Model

HY-MT2 7B

Tencent's 7B-parameter HY-MT2 model made available on OpenRouter with an 8K-token context window. It appears to be part of Tencent's HY-MT2 model family, positioned for multilingual text and translation-style workloads.

Model

GLM-5.3

GLM-5.3 is a Zhipu AI / Z.ai foundation model newly added on OpenRouter. The provided data verifies a very large 1,048,576-token context window, making it suitable for long-context text tasks.

Model

Qwen3.8-27B

Qwen3.8-27B is a Qwen 3.8 family language model from Alibaba, listed with a 262k-token context window. It appears to be a smaller open-weight member of the Qwen3.8 lineup suitable for general-purpose long-context language tasks.

Model

Dots 3 Note Preview

OpenRouter-listed free preview model from Dots Studio with a 512k-token context window. The discovery data identifies it as a recent preview model suitable for long-context text workflows.

Model

Gemini 3.7 Flash

Google's Gemini 3.7 Flash is a newly listed hosted model with a 1,048,576-token context window. It appears positioned as a fast long-context Gemini model for general-purpose AI workloads.

Model

Seed 2.1 Turbo

Seed 2.1 Turbo is a newly listed ByteDance Seed hosted model on OpenRouter with a 262,144-token context window. The Turbo designation indicates a speed-oriented variant for general-purpose long-context workloads.

Model

Qwen3.8-2.4T-A95B

Qwen3.8-2.4T-A95B is a newly listed Alibaba Qwen model with a 1,010,000-token context window. The name indicates a large mixture-style Qwen model variant intended for high-capacity long-context reasoning and generation.

Model

Grok 4.6

Grok 4.6 is a newly listed xAI hosted model with a 500,000-token context window. It appears as a new Grok-series general-purpose reasoning model.

Model

LFM 2.5 2.6B

Liquid AI's LFM 2.5 2.6B is a compact long-context foundation model listed on OpenRouter with a 128k-token context window.

Model

NVIDIA Nemotron 3.5 Lightning

NVIDIA's Nemotron 3.5 Lightning expands the Nemotron 3 family as an efficient open model for long-running agentic AI workloads.

Model

Sakana Namazu

Sakana Namazu is a Sakana AI model newly listed on OpenRouter with a 262k-token context window.

Model

Solar Pro4

Upstage Solar Pro4 is a newly listed long-context hosted foundation model on OpenRouter, supporting a 524k-token context window.

Model

Muse Glimmer 30B

Meta's Muse Glimmer 30B is a 30B-parameter open-weight model newly listed on OpenRouter with a 131k-token context window and also present in the Ollama library.

Model

Ling 3.0 Tiny

A smaller Ling 3.0-series hosted foundation model added to OpenRouter with a 262,144-token context window. The OpenRouter listing identifies it as a free-access model variant.

Model

Muse Spark 1.2

A Meta hosted foundation model added to OpenRouter with a 1,048,576-token context window. It is a newer Muse Spark release than the existing 1.1 model.

Model

Qwen3.8 Max

Alibaba’s Qwen3.8 Max is a hosted Qwen model added to OpenRouter with a 1,000,000-token context window. It is positioned for long-context general-purpose language tasks, reasoning, and generation.

Model

Inkling Small

Inkling Small is a hosted 524K-context model from Thinking Machines, added on OpenRouter as a smaller variant of the Inkling model family. It is suited for long-context text and reasoning workloads where a lighter model is preferred.

Model

Qwen3.7 Flash

Qwen3.7 Flash is a fast hosted Qwen model variant added to OpenRouter with a 1,000,000-token context window. It is suited for long-context text, reasoning, and general assistant workloads where lower latency is important.

Model

Claude Opus 5

Anthropic's most capable Opus model, positioned for advanced reasoning, agentic systems, and production inference workloads.

Model

Claude Opus 5 Fast

A faster hosted variant of Claude Opus 5 with the same 1M-token context window, aimed at lower-latency use in advanced reasoning and agentic workflows.

Model

Ling 3.0 Flash

Ling 3.0 Flash is a hosted general-purpose language model from InclusionAI, newly listed on OpenRouter with a 262K-token context window. The Flash variant is positioned for fast, long-context chat and reasoning workloads.

Model

Laguna S 2.1

Laguna S 2.1 is a long-context foundation model from Poolside, newly listed on OpenRouter and also present in the Ollama library. The OpenRouter listing reports a 1,048,576-token context window.

Model

Gemini 3.6 Flash

Gemini 3.6 Flash is a Google Gemini model newly added to OpenRouter with a 1,048,576-token context window. It is positioned as a Flash-family hosted model for long-context text generation and reasoning workloads.

Model

Gemini 3.5 Flash-Lite

Gemini 3.5 Flash-Lite is a lightweight Google Gemini hosted model newly added to OpenRouter with a 1,048,576-token context window. It is suited to cost- and latency-sensitive long-context text generation workloads.

Model

LongCat 2.0

LongCat 2.0 is a Meituan long-context foundation model newly added to OpenRouter. The listing reports a 1,048,756-token context window for hosted text-generation workloads.

Model

Inkling

Inkling is a hosted model newly listed on OpenRouter with a 1,048,576-token context window. The discovery data verifies its availability and long-context capability, but does not provide further architectural or benchmark details.

Model

Kimi K3

Kimi K3 is a Moonshot AI long-context hosted language model added to OpenRouter with a 1,048,576-token context window. It is positioned for large-context text, reasoning, and coding workloads.

Model

Muse Spark 1.1

Muse Spark 1.1 is a Meta model added to OpenRouter with a 1,048,576-token context window. The discovery data verifies it as a newly listed long-context hosted model.

Model

KAT Coder Air v2.5

KAT Coder Air v2.5 is a code-focused model from the KwaiPilot namespace added to OpenRouter with a 256,000-token context window. It appears to be the lighter Air variant of the KAT Coder v2.5 family.

Model

KAT Coder Pro v2.5

KAT Coder Pro v2.5 is a code-focused model from the KwaiPilot namespace added to OpenRouter with a 256,000-token context window. It appears to be the higher-capability Pro variant of the KAT Coder v2.5 family.

Model

GPT-5.6 Luna Pro

OpenRouter-listed GPT-5.6 hosted model variant with a 1,050,000-token context window. It is positioned as a long-context frontier text model for demanding reasoning and productivity workloads.

Model

GPT-5.6 Luna

OpenRouter-listed GPT-5.6 hosted model variant with a 1,050,000-token context window. It is intended for broad long-context text, coding, and reasoning use cases.

Model

GPT-5.6 Terra Pro

OpenRouter-listed GPT-5.6 hosted model variant with a 1,050,000-token context window. It is a proprietary long-context model for advanced text generation, reasoning, and coding workflows.

Model

GPT-5.6 Terra

OpenRouter-listed GPT-5.6 hosted model variant with a 1,050,000-token context window. It targets long-context general-purpose assistance, reasoning, and code-generation tasks.

Model

GPT-5.6 Sol Pro

OpenRouter-listed GPT-5.6 hosted model variant with a 1,050,000-token context window. It is a proprietary long-context model for high-end reasoning, coding, and productivity workloads.

Model

Grok 4.5

xAI hosted Grok model newly added on OpenRouter with a 500k-token context window for long-context general AI tasks.

Model

Aion 3.0

Aion Labs hosted general-purpose model newly added on OpenRouter with a 131k-token context window.

Model

Aion 3.0 Mini

Smaller Aion Labs hosted model newly added on OpenRouter with a 131k-token context window.

Model

Tencent HY3

Tencent long-context hosted model newly added on OpenRouter with a 262k-token context window.

Model

Laguna XS 2.1

Laguna XS 2.1 is a Poolside model added to OpenRouter with a 262K-token context window and also available in the Ollama library. It is suited to long-context coding and text workflows.

Model

NVIDIA Nemotron 3 Nano

NVIDIA Nemotron 3 Nano is an open-weight Nemotron model family referenced as newly supported on Amazon Bedrock in AWS GovCloud, including Nano 9B v2, Nano 12B v2, and Nano 30B variants.

Model

Claude Sonnet 5

Anthropic's latest-generation Sonnet model, described as its most capable Sonnet model and made available on Amazon Bedrock, Claude Platform on AWS, and OpenRouter.

Model

Fugu Ultra

Fugu Ultra is a Sakana model listed on OpenRouter with a 1M-token context window. It is positioned for long-context general-purpose reasoning and text generation.

Model

DiffusionGemma

An experimental open model from Google DeepMind built for exceptionally fast text generation, with NVIDIA optimizations for local and accelerated inference across RTX, RTX PRO, and DGX Spark systems.

Model

North Mini Code

Cohere’s first model for developers, focused on coding and developer-assistance workflows.

Model

NVIDIA Nemotron 3 Ultra 550B A55B

NVIDIA Nemotron 3 Ultra 550B A55B is a large open-weight Nemotron-family model listed on OpenRouter with a 1M-token context window and available in the Ollama library. It is intended for long-context reasoning and general text-generation workloads.

Model

Mistral Medium 3.5

A Mistral AI foundation model newly listed on OpenRouter with a 262k token context window, positioned as a balanced medium-tier model for general purpose generation and reasoning tasks.