Support

AI-powered help

Welcome!

Please introduce yourself before we start.

    • Docs
    • Pricing
    • Pricing
    • Docs
    • Models
    1.6k
    Log InGet Started

    Text Generation Models

    Models designed for text generation, chat, code, and content creation

    Compare

    Use Case

    Capabilities

    Provider

    Input Price ($/M tokens)

    Output Price ($/M tokens)

    Context Size (tokens)

    230
    Models
    36
    Providers
    112
    Vision Models
    162
    Tool-enabled
    3
    Free Models
    Features
    Nebius AI
    qwen3-235b-a22b-thinking-2507
    $0.20$0.60—
    NovitaAI
    qwen3-235b-a22b-thinking-2507
    $0.30$3.00—
    CanopyWave
    qwen3-coder
    $0.22$0.15
    -30% off
    $0.95$0.66
    -30% off
    —
    Alibaba Cloud(cn-beijing)
    qwen3-coder-flash
    $0.14$0.12
    -20% off
    $0.57$0.46
    -20% off
    $0.03$0.02
    -20% off
    Alibaba Cloud
    qwen3-coder-flash
    $0.30$0.24
    -20% off
    $1.50$1.20
    -20% off
    $0.06$0.05
    -20% off
    Alibaba Cloud(us-virginia)
    qwen3-coder-flash
    $0.14$0.12
    -20% off
    $0.57$0.46
    -20% off
    $0.03$0.02
    -20% off
    Alibaba Cloud(singapore)
    qwen3-coder-flash
    $0.30$0.24
    -20% off
    $1.50$1.20
    -20% off
    $0.06$0.05
    -20% off
    Google Vertex AI
    gemini-2.5-flash-lite
    $0.10$0.40$0.01
    Google AI Studio
    gemini-2.5-flash-lite
    $0.10$0.40$0.01
    Nebius AI
    qwen3-235b-a22b-instruct-2507
    $0.20$0.60—
    Cerebras
    qwen3-235b-a22b-instruct-2507
    $0.60$1.20—
    NovitaAI
    qwen3-235b-a22b-instruct-2507
    $0.09$0.58—
    Mistral AI
    devstral-small-2507
    $0.10$0.30—
    ByteDance
    kimi-k2
    $0.60$2.50$0.12
    Moonshot AI
    kimi-k2
    $0.60$2.50$0.15
    Groq
    kimi-k2
    $1.00$3.00$0.50
    Alibaba Cloud
    kimi-k2
    $0.57$2.29—
    NovitaAI
    kimi-k2
    $0.57$2.30—
    Nebius AI
    kimi-k2
    $0.50$2.40—
    Alibaba Cloud(cn-beijing)
    kimi-k2
    $0.57$2.29—
    xAI
    grok-4-fast
    $0.20$0.50$0.05
    xAI
    grok-4-fast-reasoning
    $0.20$0.50$0.05
    xAI
    grok-4
    $3.00$15.00$0.75
    xAI
    grok-4-0709
    $3.00$15.00$0.75
    Google AI Studio
    gemma-3n-e4b-it
    $0.07$0.30—
    Google AI Studio
    gemma-3n-e2b-it
    $0.07$0.30—
    ByteDance
    seed-1-6-250615
    $0.25$2.00$0.05
    Mistral AI
    mistral-small-2506
    $0.10$0.30—
    Google AI Studio
    gemini-2.5-pro-preview-06-05
    $1.25$10.00—
    Google Vertex AI
    gemini-2.5-pro-preview-06-05
    $1.25$10.00—
    OpenAI
    o3-mini
    $1.10$4.40$0.55
    Azure
    o3-mini
    $1.10$4.40$0.55
    Azure
    o3
    $2.00$8.00$0.50
    OpenAI
    o3
    $2.00$8.00$0.50
    Nebius AI
    deepseek-r1-0528
    $0.80$2.40—
    Anthropic
    claude-opus-4-20250514
    $15.00$75.00$1.50
    AWS Bedrock
    claude-opus-4-20250514
    $15.00$10.50
    -30% off
    $75.00$52.50
    -30% off
    $1.50$1.05
    -30% off
    Google AI Studio
    gemini-2.5-flash-preview-05-20
    $0.15$0.60—
    Google Vertex AI
    gemini-2.5-flash-preview-05-20
    $0.15$0.60—
    Anthropic
    claude-sonnet-4-20250514
    $3.00$15.00$0.30
    AWS Bedrock
    claude-sonnet-4-20250514
    $3.00$2.10
    -30% off
    $15.00$10.50
    -30% off
    $0.30$0.21
    -30% off
    Google Vertex AI
    gemini-2.5-pro-preview-05-06
    $1.25$10.00—
    Google AI Studio
    gemini-2.5-pro-preview-05-06
    $1.25$10.00—
    Groq
    llama-guard-4-12b
    $0.20$0.20—
    NovitaAI
    qwen3-4b-fp8
    $0.03$0.03—
    NovitaAI
    qwen3-30b-a3b-fp8
    $0.09$0.45—
    NovitaAI
    qwen3-32b-fp8
    $0.10$0.45—
    Nebius AI
    qwen3-30b-a3b
    $0.10$0.30—
    Nebius AI
    qwen3-32b
    $0.10$0.30—
    Cerebras
    qwen3-32b
    $0.40$0.80—
    Page 6 of 9

    Newsletter

    Stay ahead of the curve

    Join developers who get weekly insights on LLM routing, new model launches, and cost optimization — straight to their inbox.

    • New models & providers as they drop
    • Tips to cut latency & costs
    • Early access to beta features

    No spam. Unsubscribe anytime.

    Product

    • Features
    • Models
    • Providers
    • Chat Playground
    • Changelog
    • DevPass
    • Compare Models
    • Enterprise

    Resources

    • Templates
    • Agents
    • MCP Server
    • Blog
    • Documentation
    • Integrations
    • Guides
    • Brand Assets
    • Token Cost Calculator
    • Referral Program
    • GitHub
    • Contact Us

    Community

    • Twitter
    • Discord

    Compare

    • OpenRouter
    • LiteLLM

    Models

    • Text Generation
    • Text to Image
    • Image to Image
    • Vision
    • Reasoning
    • Tool Calling
    • Web Search
    • Discounted

    Providers

    • OpenAI
    • Anthropic
    • Google AI Studio
    • Glacier
    • Google Vertex AI
    • Quartz
    • Avalanche
    • Groq
    • Cerebras
    • xAI
    • DeepSeek
    • Bluestone
    • Alibaba Cloud
    • NovitaAI
    • AWS Bedrock
    • Azure
    • Z AI
    • Moonshot AI
    • Perplexity
    • Nebius AI
    • Mistral AI
    • CanopyWave
    • Inference.net
    • Together AI
    • Custom
    • NanoGPT
    • ByteDance
    • MiniMax
    • EmberCloud
    • Fireworks AI
    • Parasail
    • DeepInfra
    • W&B Inference (CoreWeave)
    • GMI Cloud
    • Bitdeer AI

    © 2026 LLM Gateway. All rights reserved.

    All systems operationalPrivacy PolicyTerms of Use