Premium Models

    Models priced at $5+ per million input tokens or $15+ per million output tokens — the exact list used for DevPass fair-use limits

    Compare
    Use Case
    Capabilities
    Provider
    Status
    Input Price ($/M tokens)
    Output Price ($/M tokens)
    Context Size (tokens)
    46
    Models
    59
    Providers
    38
    Vision Models
    40
    Tool-enabled
    0
    Free Models
    Features
    AWS Bedrock(us)
    claude-sonnet-4-6
    $3.30$16.50$0.33
    Anthropic
    claude-sonnet-4-6
    $3.00$15.00$0.30

    S

    AvalAI
    gpt-image-2
    $5.00$10.00$1.25

    S

    GapGPT
    gpt-image-2
    $8.00$30.00—

    S

    Metis
    gpt-image-2
    $8.00$15.00$2.00
    Azure
    gpt-image-2
    $5.00$0.00$1.25
    OpenAI
    gpt-image-2
    $5.00$0.00$1.25

    S

    Hormouz
    gpt-image-2
    $8.00$15.00$2.00

    S

    OpenRouter
    gpt-5.5-pro
    $30.00$180.00—
    OpenAI
    gpt-5.5-pro
    $30.00$180.00—

    S

    Metis
    gpt-5.5
    $5.00$30.00$0.50
    OpenAI
    gpt-5.5
    $5.00$30.00$0.50

    S

    AvalAI
    gpt-5.5
    $5.00$30.00$0.50

    S

    Hormouz
    gpt-5.5
    $5.00$30.00$0.50

    S

    GapGPT
    gpt-5.5
    $5.00$30.00$0.50

    S

    OpenRouter
    gpt-5.5
    $5.00$30.00$0.50
    Azure
    gpt-5.5
    $5.00$30.00$0.50

    S

    GapGPT
    gpt-5.4-nano
    $75.00$600.00—

    S

    Hormouz
    gpt-5.4-nano
    $0.20$1.25$0.02

    S

    AvalAI
    gpt-5.4-nano
    $0.20$1.25$0.02
    Azure
    gpt-5.4-nano
    $0.20$1.25$0.02

    S

    Metis
    gpt-5.4-nano
    $0.20$1.25$0.02

    S

    OpenRouter
    gpt-5.4-nano
    $0.20$1.25$0.02
    OpenAI
    gpt-5.4-nano
    $0.20$1.25$0.02
    Azure
    gpt-5.4-pro
    $30.00$180.00—

    S

    OpenRouter
    gpt-5.4-pro
    $30.00$180.00—
    OpenAI
    gpt-5.4-pro
    $30.00$180.00—

    S

    Hormouz
    gpt-5.4
    $2.50$15.00$0.25
    OpenAI
    gpt-5.4
    $2.50$15.00$0.25
    Azure
    gpt-5.4
    $2.50$15.00$0.25

    S

    GapGPT
    gpt-5.4
    $2.50$15.00$0.25

    S

    AvalAI
    gpt-5.4
    $2.50$15.00$0.25

    S

    OpenRouter
    gpt-5.4
    $2.50$15.00$0.25

    S

    Metis
    gpt-5.4
    $2.50$15.00$0.25
    AWS Bedrock
    claude-opus-4-6
    $5.00$25.00$0.50

    S

    AvalAI
    claude-opus-4-6
    $5.00$25.00$1.50
    AWS Bedrock(eu-west-2)
    claude-opus-4-6
    $5.50$27.50$0.55

    S

    GapGPT
    claude-opus-4-6
    $5.00$25.00$0.50

    S

    OpenRouter
    claude-opus-4-6
    $5.00$25.00$0.50
    AWS Bedrock(us)
    claude-opus-4-6
    $5.50$27.50$0.55
    AWS Bedrock(eu)
    claude-opus-4-6
    $5.50$27.50$0.55
    AWS Bedrock(au)
    claude-opus-4-6
    $5.50$27.50$0.55
    AWS Bedrock(global)
    claude-opus-4-6
    $5.00$25.00$0.50
    Vertex AI (Anthropic)
    claude-opus-4-6
    $5.00$25.00$0.50
    Anthropic
    claude-opus-4-6
    $5.00$25.00$0.50
    Alibaba Cloud
    qwen3-max
    $3.00$15.00$0.60
    NovitaAI
    qwen3-max
    $0.84$3.38—

    S

    OpenRouter
    qwen3-max
    $0.78$3.90$0.16

    S

    AvalAI
    qwen3-max
    $1.20$6.00$0.10
    Alibaba Cloud
    qwen3-coder-plus
    $6.00$60.00$1.20
    Page 3 of 5

    Newsletter

    Stay ahead of the curve

    Join developers who get weekly insights on LLM routing, new model launches, and cost optimization — straight to their inbox.

    • New models & providers as they drop
    • Tips to cut latency & costs
    • Early access to beta features

    No spam. Unsubscribe anytime.

    S

    SOLOP
    All systems operational
    AICPA SOC for Service Organizations badgeSOC 2 Type II
    compliant

    Product

    • Features
    • AI Gateway
    • Observability
    • Models
    • Providers
    • Rankings
    • Add Provider
    • Lounge
    • Changelog
    • DevPass
    • Compare Models
    • Enterprise

    Resources

    • Apps
    • Templates
    • Agents
    • MCP Server
    • Use Cases
    • Blog
    • Documentation
    • Integrations
    • Guides
    • Brand Assets
    • Token Cost Calculator
    • Copilot Cost Calculator
    • Referral Program
    • GitHub
    • Discord
    • Twitter
    • Contact Us

    Compliance

    • Trust Center
    • Security Portal
    • Terms
    • Privacy Policy
    • GDPR
    • SOC 2 Type II
    • Status

    Compare

    • All Comparisons
    • GitHub Copilot
    • OpenRouter
    • LiteLLM
    • Portkey
    • AWS Bedrock
    • Azure AI Foundry
    • Vercel AI Gateway
    • Migration Guides

    Models

    • Text Generation
    • Text to Image
    • Image to Image
    • Video Generation
    • Embeddings
    • Vision
    • Reasoning
    • Tool Calling
    • Web Search
    • Discounted
    • Best for Roleplay
    • Best for Coding
    • Best for Creative Writing
    • Best for Translation
    • Best for Math
    • Long Context
    • Cheapest
    • Open Source

    Providers

    • OpenAI
    • Anthropic
    • Google AI Studio
    • Glacier
    • Iceberg
    • Granite
    • Google Vertex AI
    • Vertex AI (OpenAI-compatible)
    • Vertex AI (Anthropic)
    • Quartz
    • Avalanche
    • Obsidian
    • Groq
    • Cerebras
    • xAI
    • DeepSeek
    • Alibaba Cloud
    • NovitaAI
    • AtlasCloud
    • AWS Bedrock
    • AWS Mantle
    • Azure
    • Azure AI Foundry
    • Z AI
    • Moonshot AI
    • Perplexity
    • Nebius AI
    • Mistral AI
    • CanopyWave
    • Inference.net
    • Together AI
    • SCX.ai (Turbo)
    • SCX.ai
    • Custom
    • RouteWay
    • CloudRift
    • NanoGPT
    • ByteDance
    • MiniMax
    • OpenRouter
    • EmberCloud
    • Meta
    • Sakana AI
    • Tundra
    • Bluestone
    • Xiaomi
    • AvalAI
    • Hormouz
    • Metis
    • GapGPT
    • DeepInfra
    • Reve
    • ElevenLabs
    • Runware
    • Gonka24
    • Fireworks AI
    • RanoAI

    © 2026 SOLOP. All rights reserved.

    Premium is a pricing classification, not a curated list: a model lands on this page when any of its providers charges at least $5 per million input tokens or $15 per million output tokens. The classification is computed directly from catalogue prices, so this page always shows exactly which models are premium right now — every model not listed here is standard.

    The distinction matters for DevPass coding plans, where premium models are subject to a weekly fair-use allowance (a percentage of the plan's monthly credits) on top of the normal credit balance. When you call the LLM Gateway API directly with pay-as-you-go credits, premium models have no extra cap and no markup — you pay the same per-token provider prices shown here.

    Frequently asked questions

    What makes a model premium?

    Pricing alone. A model is premium when at least one of its providers charges $5 or more per million input tokens, or $15 or more per million output tokens. There is no hand-picked list — the classification is recomputed from the live model catalogue, so it updates automatically when prices change.

    Do premium models cost extra on LLM Gateway?

    No. LLM Gateway charges the same per-token provider prices for premium models as for any other model, with no markup. The premium classification only affects DevPass fair-use limits — it never changes what a request costs.

    How do premium models work on DevPass plans?

    DevPass plans include a weekly fair-use allowance for premium models: 12% of monthly credits on Lite, 15% on Pro, and 18% on Max. The allowance works on a fixed 7-day window that opens with your first premium request and fully resets when it ends. Standard models are never affected — they're limited only by the plan's credit balance.

    How do I check whether a specific model is premium?

    If it appears on this page, it's premium; otherwise it's standard. In the full models directory, premium models are marked with a gem icon, and the DevPass models directory offers a pricing-tier filter, since the distinction only affects DevPass fair-use limits. The classification is also documented in the model categories guide in our docs.

    S

    LLM Gateway
    • DevPass
    • Lounge
    • Models
    • Docs
    • Pricing
    • DevPass
    • Lounge
    • Pricing
    • Docs
    • Models
      • AI Gateway
      • DevPass
      • Lounge
      • Observability
      • Enterprise
      • Blog
      • Changelog
      • Integrations
      • Reliability
      • Guardrails
      • Providers
      • Rankings
      • Apps
      • Models
      • Model Timeline
      • Compare
      • Token Cost Calculator
      • Referral Program
      • MCP Server
      • Agents
      • AI SDK Provider
      • Agent Skills
      • Templates
      • Guides
    Log InGet Started
    1.6k