endueendue

Text

Every model here reads and writes text: chat, writing, analysis, code and tool calls.

models
72
providers
15

Prices and specs checked 2026-09-20

72 models

Filter models

Provider
Reads
Capabilities
Good for
  • xAI

    Grok 4.6

    Grok 4

    xAI’s top Grok for coding, STEM and knowledge work.

    TextText + Images + Files → Text
    • Coding
    • Reasoning
    • Long documents
    Pricing
    $2 / $6 per 1M tokens, in / out
    Context window
    500K context
    Ready on endueUse in agent
  • xAI

    Grok 4.3

    Grok 4

    Reasoning model that follows instructions closely, with a 1M context. endue’s default.

    TextText + Images + Files → Text
    • Agentic work
    • Reasoning
    • Long documents
    Pricing
    $1.25 / $2.5 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • xAI

    Grok 4.20

    Grok 4

    Fast tool calling with a 2M context for very long inputs.

    TextText + Images + Files → Text
    • Agentic work
    • Long documents
    • Fast replies
    Pricing
    $1.25 / $2.5 per 1M tokens, in / out
    Context window
    2M context
    Ready on endueUse in agent
  • xAI

    Grok 4.5

    Grok 4

    Balanced Grok for coding and knowledge work, 500K context.

    TextText + Images + Files → Text
    • Coding
    • Reasoning
    Pricing
    $2 / $6 per 1M tokens, in / out
    Context window
    500K context
    Ready on endueUse in agent
  • xAI

    Grok Build 0.1

    Grok Build

    Fast coding model tuned for interactive software engineering.

    TextText + Images + Files → Text
    • Coding
    • Fast replies
    • Low cost
    Pricing
    $1 / $2 per 1M tokens, in / out
    Context window
    256K context
    Ready on endueUse in agent
  • OpenAI

    GPT-6 Astra

    GPT-6

    OpenAI’s flagship for long, demanding work: analysis, research, large codebases.

    TextFiles + Images + Text → Text
    • Reasoning
    • Coding
    • Long documents
    Pricing
    $10 / $50 per 1M tokens, in / out
    Context window
    1.1M context
    Ready on endueUse in agent
  • OpenAI

    GPT-6 Astra Pro

    GPT-6

    GPT-6 Astra with deeper reasoning per answer. Same price, slower replies.

    TextFiles + Images + Text → Text
    • Reasoning
    • Web research
    Pricing
    $10 / $50 per 1M tokens, in / out
    Context window
    1.1M context
    Ready on endueUse in agent
  • OpenAI

    GPT-5.6 Sol

    GPT-5.6

    Top GPT-5.6 tier, strong at multi-step and command-line coding.

    TextFiles + Images + Text → Text
    • Coding
    • Agentic work
    • Reasoning
    Pricing
    $2 / $10 per 1M tokens, in / out
    Context window
    1.1M context
    Ready on endueUse in agent
  • OpenAI

    GPT-5.6 Terra

    GPT-5.6

    Middle GPT-5.6 tier for everyday coding and agent tasks.

    TextFiles + Images + Text → Text
    • Coding
    • Agentic work
    Pricing
    $2 / $12 per 1M tokens, in / out
    Context window
    1.1M context
    Ready on endueUse in agent
  • OpenAI

    GPT-5.6 Luna

    GPT-5.6

    Cheap, fast GPT-5.6 tier for chat, classification and high volume.

    TextFiles + Images + Text → Text
    • Fast replies
    • Low cost
    • Long documents
    Pricing
    $0.2 / $1.2 per 1M tokens, in / out
    Context window
    1.1M context
    Ready on endueUse in agent
  • OpenAI

    GPT-5.5

    GPT-5

    Reliable reasoning for professional workloads, 1M context.

    TextFiles + Images + Text → Text
    • Reasoning
    • Long documents
    Pricing
    $5 / $30 per 1M tokens, in / out
    Context window
    1.1M context
    Ready on endueUse in agent
  • OpenAI

    GPT-5.4

    GPT-5

    Codex and GPT merged into one model, with a 1M context.

    TextText + Images + Files → Text
    • Coding
    • Reasoning
    • Long documents
    Pricing
    $2.5 / $15 per 1M tokens, in / out
    Context window
    1.1M context
    Ready on endueUse in agent
  • OpenAI

    GPT-5.4 Mini

    GPT-5

    Smaller GPT-5.4 for high-throughput reasoning and coding.

    TextFiles + Images + Text → Text
    • Fast replies
    • Coding
    Pricing
    $0.75 / $4.5 per 1M tokens, in / out
    Context window
    400K context
    Ready on endueUse in agent
  • OpenAI

    GPT-5.4 Nano

    GPT-5

    Lightest GPT-5.4 for latency-sensitive, high-volume tasks.

    TextFiles + Images + Text → Text
    • Fast replies
    • Low cost
    Pricing
    $0.2 / $1.25 per 1M tokens, in / out
    Context window
    400K context
    Ready on endueUse in agent
  • OpenAI

    GPT-5.3 Codex

    GPT-5 Codex

    OpenAI’s agentic coding model for long software tasks.

    TextText + Images + Files → Text
    • Coding
    • Agentic work
    Pricing
    $1.75 / $14 per 1M tokens, in / out
    Context window
    400K context
    Ready on endueUse in agent
  • OpenAI

    GPT-5

    GPT-5

    Step-by-step reasoning with careful instruction following.

    TextText + Images + Files → Text
    • Reasoning
    • Coding
    Pricing
    $1.25 / $10 per 1M tokens, in / out
    Context window
    400K context
    Ready on endueUse in agent
  • OpenAI

    GPT-OSS 120B

    GPT-OSS

    Open-weight MoE with reasoning at a very low price. Text only.

    TextText → Text
    • Open weights
    • Low cost
    • Reasoning
    Pricing
    $0.15 / $0.6 per 1M tokens, in / out
    Context window
    131K context
    Ready on endueUse in agent
  • OpenAI

    GPT-OSS 20B

    GPT-OSS

    Small Apache-2.0 open-weight model for cheap, simple tasks.

    TextText → Text
    • Open weights
    • Low cost
    • Fast replies
    Pricing
    $0.03 / $0.13 per 1M tokens, in / out
    Context window
    131K context
    Ready on endueUse in agent
  • OpenAI

    GPT-4o mini

    GPT-4o

    Small non-reasoning model for quick, inexpensive replies.

    TextText + Images + Files → Text
    • Fast replies
    • Low cost
    Pricing
    $0.15 / $0.6 per 1M tokens, in / out
    Context window
    128K context
    Ready on endueUse in agent
  • Anthropic

    Claude Opus 5

    Claude Opus

    Anthropic’s flagship for code review, bug finding and long agent runs.

    TextText + Images + Files → Text
    • Coding
    • Agentic work
    • Reasoning
    Pricing
    $5 / $25 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Anthropic

    Claude Fable 5.1

    Claude Fable

    Fable 5 improved for long refactors, front-end work and knowledge tasks.

    TextText + Images + Files → Text
    • Coding
    • Agentic work
    • Long documents
    Pricing
    $10 / $50 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Anthropic

    Claude Fable 5

    Claude Fable

    Built for autonomous knowledge work and coding.

    TextText + Images + Files → Text
    • Agentic work
    • Coding
    Pricing
    $10 / $50 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Anthropic

    Claude Opus 4.8

    Claude Opus

    Latest Opus 4 release, 1M context.

    TextText + Images + Files → Text
    • Coding
    • Reasoning
    • Long documents
    Pricing
    $5 / $25 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Anthropic

    Claude Sonnet 5

    Claude Sonnet

    Strong coding and agent work at a lower price than Opus.

    TextText + Images + Files → Text
    • Coding
    • Agentic work
    Pricing
    $2 / $10 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Anthropic

    Claude Sonnet 4.5

    Claude Sonnet

    Sonnet tuned for real-world agents and coding workflows.

    TextText + Images + Files → Text
    • Coding
    • Agentic work
    Pricing
    $3 / $15 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Anthropic

    Claude Haiku 4.5

    Claude Haiku

    Fastest Claude, near Sonnet 4 quality at lower cost.

    TextText + Images + Files → Text
    • Fast replies
    • Low cost
    Pricing
    $1 / $5 per 1M tokens, in / out
    Context window
    200K context
    Ready on endueUse in agent
  • Anthropic

    Claude Opus 4.7

    Claude Opus

    Opus for long-running, asynchronous agents.

    TextText + Images + Files → Text
    • Agentic work
    • Coding
    Pricing
    $5 / $25 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Anthropic

    Claude Opus 4.6

    Claude Opus

    Opus for agents that run whole workflows, not single prompts.

    TextText + Images + Files → Text
    • Agentic work
    • Coding
    Pricing
    $5 / $25 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Anthropic

    Claude Sonnet 4.6

    Claude Sonnet

    Iterative development and navigating large codebases.

    TextText + Images + Files → Text
    • Coding
    • Agentic work
    Pricing
    $3 / $15 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Anthropic

    Claude Opus 4.5

    Claude Opus

    Opus for complex software engineering and computer use, 200K context.

    TextFiles + Images + Text → Text
    • Coding
    • Agentic work
    Pricing
    $5 / $25 per 1M tokens, in / out
    Context window
    200K context
    Ready on endueUse in agent
  • Anthropic

    Claude Opus 4.1

    Claude Opus

    Earlier Opus. The most expensive Claude here; newer Opus models cost less.

    TextImages + Text + Files → Text
    • Coding
    • Reasoning
    Pricing
    $15 / $75 per 1M tokens, in / out
    Context window
    200K context
    Ready on endueUse in agent
  • Google

    Gemini 3.8 Flash

    Gemini Flash

    Latest Gemini Flash. Reads text, images, audio and video; 1M context.

    MultimodalText + Images + Video + Files + Audio → Text
    • Audio & video
    • Agentic work
    • Long documents
    Pricing
    $0.75 / $3.75 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Google

    Gemini 3.5 Flash

    Gemini Flash

    Near-Pro coding and reasoning at Flash cost, with parallel tool use.

    MultimodalText + Images + Video + Files + Audio → Text
    • Coding
    • Audio & video
    • Agentic work
    Pricing
    $1.5 / $9 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Google

    Gemini 3.6 Flash

    Gemini Flash

    Flash model for coding and web/app development with fewer stray edits.

    MultimodalText + Images + Video + Files + Audio → Text
    • Coding
    • Audio & video
    Pricing
    $0.75 / $3.75 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Google

    Gemini 3.5 Flash Lite

    Gemini Flash-Lite

    Cheap Gemini for sub-agents that run focused tasks.

    MultimodalText + Images + Video + Files + Audio → Text
    • Low cost
    • Audio & video
    • Fast replies
    Pricing
    $0.3 / $2.5 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Google

    Gemini 2.5 Pro

    Gemini Pro

    Gemini with step-by-step thinking for math, science and code.

    MultimodalText + Images + Files + Audio + Video → Text
    • Reasoning
    • Audio & video
    • Long documents
    Pricing
    $1.25 / $10 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Google

    Gemini 3.1 Flash Lite

    Gemini Flash-Lite

    Low-latency Gemini for high volume. Reads audio, video and PDF.

    MultimodalText + Images + Video + Files + Audio → Text
    • Low cost
    • Fast replies
    • Audio & video
    Pricing
    $0.25 / $1.5 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Google

    Gemini 2.5 Flash Lite

    Gemini Flash-Lite

    Among the cheapest multimodal models here, 1M context.

    MultimodalText + Images + Files + Audio + Video → Text
    • Low cost
    • Fast replies
    • Audio & video
    Pricing
    $0.1 / $0.4 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Google

    Gemma 4 31B

    Gemma 4

    Open-weight dense model with image and video input and function calling.

    MultimodalImages + Text + Video → Text
    • Open weights
    • Low cost
    • Images
    Pricing
    $0.09 / $0.34 per 1M tokens, in / out
    Context window
    262K context
    Ready on endueUse in agent
  • Google

    Gemma 4 26B

    Gemma 4

    Open-weight MoE close to Gemma 4 31B quality at lower compute.

    MultimodalImages + Text + Video → Text
    • Open weights
    • Low cost
    • Fast replies
    Pricing
    $0.09 / $0.3 per 1M tokens, in / out
    Context window
    262K context
    Ready on endueUse in agent
  • Moonshot AI

    Kimi K3

    Kimi

    Large open-weight multimodal model for coding and long agent runs.

    MultimodalText + Images + Video → Text
    • Coding
    • Agentic work
    • Open weights
    Pricing
    $1.7 / $8.5 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Moonshot AI

    Kimi K2.7 Code

    Kimi

    Coding-focused Kimi for end-to-end programming over long contexts.

    TextText + Images → Text
    • Coding
    • Agentic work
    Pricing
    $0.71 / $3.21 per 1M tokens, in / out
    Context window
    262K context
    Ready on endueUse in agent
  • Moonshot AI

    Kimi K2.6

    Kimi

    Long coding tasks, UI generation from code and multi-agent setups.

    TextText + Images → Text
    • Coding
    • Agentic work
    • Images
    Pricing
    $0.95 / $4 per 1M tokens, in / out
    Context window
    262K context
    Ready on endueUse in agent
  • DeepSeek

    DeepSeek V4.1 Flash

    DeepSeek V4

    Low-cost DeepSeek with image input and a 1M context.

    TextText + Images → Text
    • Low cost
    • Long documents
    • Images
    Pricing
    $0.15 / $0.6 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • DeepSeek

    DeepSeek V4 Pro

    DeepSeek V4

    Large MoE for reasoning and coding, 1M context. Text only.

    TextText → Text
    • Reasoning
    • Coding
    • Long documents
    Pricing
    $0.42 / $0.84 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • DeepSeek

    DeepSeek V4 Flash

    DeepSeek V4

    Very cheap text model with a 1M context for fast inference.

    TextText → Text
    • Low cost
    • Fast replies
    • Long documents
    Pricing
    $0.036 / $0.071 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • DeepSeek

    DeepSeek V3.2

    DeepSeek V3

    Efficient reasoning and tool use at low cost. Text only.

    TextText → Text
    • Low cost
    • Reasoning
    Pricing
    $0.27 / $0.4 per 1M tokens, in / out
    Context window
    164K context
    Ready on endueUse in agent
  • Alibaba Qwen

    Qwen3.8 Max

    Qwen3.8

    Largest Qwen, with text, image and video input.

    MultimodalText + Images + Video → Text
    • Reasoning
    • Coding
    • Audio & video
    Pricing
    $2 / $6 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Alibaba Qwen

    Qwen3.8 Flash

    Qwen3.8

    Cheap multimodal Qwen for charts, documents and long videos.

    MultimodalText + Images + Video → Text
    • Low cost
    • Audio & video
    • Images
    Pricing
    $0.15 / $0.47 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Alibaba Qwen

    Qwen3.7 Max

    Qwen3.7

    Qwen3.7 flagship for agent work, coding and office tasks. Text only.

    TextText → Text
    • Agentic work
    • Coding
    Pricing
    $1.48 / $4.42 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Alibaba Qwen

    Qwen3.7 Plus

    Qwen3.7

    Cost-effective Qwen3.7 with image input.

    TextText + Images → Text
    • Low cost
    • Images
    Pricing
    $0.32 / $1.28 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Alibaba Qwen

    Qwen3.6 Flash

    Qwen3.6

    Fast Qwen with image and video input and a 1M context.

    MultimodalText + Images + Video → Text
    • Fast replies
    • Low cost
    • Audio & video
    Pricing
    $0.19 / $1.13 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Alibaba Qwen

    Qwen3 Coder Next

    Qwen3 Coder

    Open-weight coding model for coding agents. No reasoning mode.

    TextText → Text
    • Coding
    • Open weights
    • Low cost
    Pricing
    $0.12 / $0.8 per 1M tokens, in / out
    Context window
    262K context
    Ready on endueUse in agent
  • Z.ai

    GLM-5.3

    GLM-5

    Reasoning model for complex engineering and long agent tasks. Text only.

    TextText → Text
    • Coding
    • Agentic work
    • Long documents
    Pricing
    $0.91 / $2.86 per 1M tokens, in / out
    Context window
    1.3M context
    Ready on endueUse in agent
  • Z.ai

    GLM-5.2

    GLM-5

    Project-level software work with a 1M context at a lower price.

    TextText → Text
    • Coding
    • Long documents
    • Low cost
    Pricing
    $0.65 / $2.04 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Z.ai

    GLM-4.7 Flash

    GLM-4

    Small, very cheap GLM tuned for agentic coding.

    TextText → Text
    • Low cost
    • Coding
    • Fast replies
    Pricing
    $0.061 / $0.4 per 1M tokens, in / out
    Context window
    200K context
    Ready on endueUse in agent
  • Z.ai

    GLM-5V Turbo

    GLM-5V

    GLM that reads images and video, for vision-based coding.

    MultimodalImages + Text + Video → Text
    • Images
    • Coding
    • Audio & video
    Pricing
    $1.2 / $4 per 1M tokens, in / out
    Context window
    203K context
    Ready on endueUse in agent
  • MiniMax

    MiniMax M3

    MiniMax

    Multimodal model with a 1M context for long agent work.

    MultimodalText + Images + Video → Text
    • Audio & video
    • Agentic work
    • Low cost
    Pricing
    $0.3 / $1.2 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • MiniMax

    MiniMax M2.7

    MiniMax

    Low-cost text model built for multi-agent productivity.

    TextText → Text
    • Agentic work
    • Low cost
    Pricing
    $0.3 / $1.2 per 1M tokens, in / out
    Context window
    205K context
    Ready on endueUse in agent
  • Mistral AI

    Mistral Medium 3.5

    Mistral Medium

    Dense 128B model for agent workflows and coding.

    TextText + Images + Files → Text
    • Agentic work
    • Coding
    Pricing
    $1.5 / $7.5 per 1M tokens, in / out
    Context window
    262K context
    Ready on endueUse in agent
  • Mistral AI

    Mistral Small 2603

    Mistral Small

    Mistral Small 4: reasoning and image input in one cheap model.

    TextText + Images → Text
    • Low cost
    • Reasoning
    • Images
    Pricing
    $0.15 / $0.6 per 1M tokens, in / out
    Context window
    262K context
    Ready on endueUse in agent
  • Mistral AI

    Codestral 2508

    Codestral

    Low-latency code completion, fixes and test generation.

    TextText + Files → Text
    • Coding
    • Fast replies
    • Low cost
    Pricing
    $0.3 / $0.9 per 1M tokens, in / out
    Context window
    256K context
    Ready on endueUse in agent
  • Mistral AI

    Devstral 2

    Devstral

    Open-source 123B model for agentic coding across a repository.

    TextText + Files → Text
    • Coding
    • Agentic work
    • Open weights
    Pricing
    $0.4 / $2 per 1M tokens, in / out
    Context window
    262K context
    Ready on endueUse in agent
  • Mistral AI

    Ministral 14B

    Ministral

    Efficient 14B with image input at a flat low price.

    TextText + Images → Text
    • Low cost
    • Images
    • Fast replies
    Pricing
    $0.2 / $0.2 per 1M tokens, in / out
    Context window
    262K context
    Ready on endueUse in agent
  • Meta

    Muse Spark 1.3

    Muse Spark

    Keeps track of long tasks. Reads text, images, audio, video and PDF.

    MultimodalText + Images + Video + Files + Audio → Text
    • Agentic work
    • Audio & video
    • Long documents
    Pricing
    $1.25 / $4.25 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • Meta

    Muse Spark 1.1

    Muse Spark

    Multimodal reasoning for agent tasks, 1M context.

    MultimodalText + Images + Video + Files + Audio → Text
    • Agentic work
    • Audio & video
    Pricing
    $1.25 / $4.25 per 1M tokens, in / out
    Context window
    1M context
    Ready on endueUse in agent
  • NVIDIA

    Nemotron 3 Ultra

    Nemotron 3

    Open reasoning model for orchestrating other agents. Text only.

    TextText → Text
    • Reasoning
    • Agentic work
    • Open weights
    Pricing
    $0.6 / $2.4 per 1M tokens, in / out
    Context window
    262K context
    Ready on endueUse in agent
  • NVIDIA

    Nemotron 3 Super

    Nemotron 3

    Open hybrid MoE for multi-agent apps at low compute.

    TextText → Text
    • Open weights
    • Low cost
    • Agentic work
    Pricing
    $0.08 / $0.45 per 1M tokens, in / out
    Context window
    262K context
    Ready on endueUse in agent
  • NVIDIA

    Nemotron 3 Nano

    Nemotron 3

    Small open model for narrow, specialized agents.

    TextText → Text
    • Open weights
    • Low cost
    • Fast replies
    Pricing
    $0.06 / $0.24 per 1M tokens, in / out
    Context window
    262K context
    Ready on endueUse in agent
  • inclusionAI

    Ling 3.0 Flash

    Ling

    Token-efficient MoE, the cheapest model here. Text only.

    TextText → Text
    • Low cost
    • Fast replies
    Pricing
    $0.021 / $0.063 per 1M tokens, in / out
    Context window
    262K context
    Ready on endueUse in agent
  • StepFun

    Step 3.7 Flash

    Step

    Efficient MoE with native image and video understanding.

    MultimodalText + Images + Video → Text
    • Audio & video
    • Low cost
    • Images
    Pricing
    $0.2 / $1.15 per 1M tokens, in / out
    Context window
    262K context
    Ready on endueUse in agent
  • Perplexity

    Sonar Pro Search

    Sonar

    Searches the web and reasons over results. Cannot call tools.

    TextText + Images → Text
    • Web research
    Pricing
    $3 / $15 per 1M tokens, in / out
    Context window
    200K context
    Ready on endueUse in agent