Support

AI-powered help

Welcome!

Please introduce yourself before we start.

    AI Models Directory

    Browse and compare 200+ AI models from OpenAI, Anthropic, Google, and 40+ providers — filter by capabilities, pricing, and context size.

    Compare
    Use Case
    Capabilities
    Provider
    Status
    Input Price ($/M tokens)
    Output Price ($/M tokens)
    Context Size (tokens)
    351
    Models
    51
    Providers
    174
    Vision Models
    225
    Tool-enabled
    4
    Free Models
    Features
    Alibaba Cloud(us-virginia)
    glm-5.2
    $1.40$4.40$0.28
    Alibaba Cloud
    glm-5.2
    $1.40$4.40$0.28
    NovitaAI
    glm-5.2
    $1.40$4.40$0.26
    Alibaba Cloud(eu-frankfurt)
    glm-5.2
    $1.40$4.40$0.28
    Z AI
    glm-5.2
    $1.40$4.40$0.26
    CanopyWave
    glm-5.2
    $1.40$4.40$0.26
    Granite
    glm-5.2
    $1.40$4.40$0.26
    NovitaAI
    glm-5.2
    $1.40$4.40$0.26
    Z AI
    glm-5.2
    $1.40$4.40$0.26
    Runware
    glm-5.2
    $0.80$2.55$0.16
    SCX.ai
    glm-5.2
    $0.55$1.78$0.11
    EmberCloud
    glm-5.2
    $1.26$3.96$0.23
    Nebius AI
    glm-5.2
    $1.40$4.40—
    Baidu
    glm-5.2
    $1.40$4.40$0.26
    Baidu
    glm-5.2
    $1.40$4.40$0.26
    Nebius AI
    glm-5.2
    $1.40$4.40—
    ByteDance
    glm-5.2
    $1.40$4.40$0.26
    Runware
    glm-5.2
    $0.80$2.55$0.16
    CanopyWave
    glm-5.2
    $1.40$4.40$0.26
    SCX.ai(au)
    glm-5.2
    $0.55$1.78$0.11
    Z AI
    glm-5.3
    $1.40$4.40$0.26
    Z AI
    glm-5.3
    $1.40$4.40$0.26
    Nebius AI
    minicpm-v-4.5
    $0.66$1.11—
    Nebius AI
    minicpm-v-4.5
    $0.66$1.11—
    Nebius AI
    cosmos3-super-reasoner
    $0.10$0.30—
    Nebius AI
    cosmos3-super-reasoner
    $0.10$0.30—
    Nebius AI
    nemotron-3-nano-omni
    $0.06$0.24—
    Nebius AI
    nemotron-3-nano-omni
    $0.06$0.24—
    Nebius AI
    nemotron-3-nano-30b
    $0.06$0.24—
    Nebius AI
    nemotron-3-nano-30b
    $0.06$0.24—
    Nebius AI
    nemotron-3-super-120b
    $0.30$0.90—
    Nebius AI
    nemotron-3-super-120b
    $0.30$0.90—
    DeepInfra
    nemotron-3-ultra-550b
    $0.50$2.20$0.10
    DeepInfra
    nemotron-3-ultra-550b
    $0.50$2.20$0.10
    Nebius AI
    nemotron-3-ultra-550b
    $1.00$3.00—
    Nebius AI
    nemotron-3-ultra-550b
    $1.00$3.00—
    DeepInfra
    hy3
    $0.14$0.58$0.04
    NovitaAI
    hy3
    $0.14$0.58$0.04
    NovitaAI
    hy3
    $0.14$0.58$0.04
    DeepInfra
    hy3
    $0.14$0.58$0.04
    Sakana AI
    fugu-ultra
    $5.00$30.00$0.50
    Sakana AI
    fugu-ultra
    $5.00$30.00$0.50
    Reve
    reve-create
    $0.024/req——
    NovitaAI
    hermes-2-pro-llama-3-8b
    $0.14$0.14—
    Nebius AI
    hermes-3-llama-405b
    $1.00$3.00—
    Nebius AI
    hermes-4-70b
    $0.13$0.40—
    Nebius AI
    hermes-4-70b
    $0.13$0.40—
    Nebius AI
    hermes-4-405b
    $1.00$3.00—
    Nebius AI
    hermes-4-405b
    $1.00$3.00—
    ByteDance
    seedream-5-0-pro
    $0.090/req——
    Page 3 of 25

    Newsletter

    Stay ahead of the curve

    Join developers who get weekly insights on LLM routing, new model launches, and cost optimization — straight to their inbox.

    • New models & providers as they drop
    • Tips to cut latency & costs
    • Early access to beta features

    No spam. Unsubscribe anytime.

    All systems operational
    AICPA SOC for Service Organizations badgeSOC 2 Type II
    compliant

    Product

    • Features
    • AI Gateway
    • Observability
    • Models
    • Providers
    • Rankings
    • Add Provider
    • Partners
    • Lounge
    • Changelog
    • Compare Models
    • Enterprise

    Resources

    • Legal Overview
    • Apps
    • MCP Server
    • Use Cases
    • Blog
    • Documentation
    • Integrations
    • Guides
    • Brand Assets
    • Token Cost Calculator
    • Copilot Cost Calculator
    • Referral Program
    • GitHub
    • Discord
    • Twitter
    • Contact Us

    Compliance

    • Trust Center
    • Security Portal
    • Terms
    • Privacy Policy
    • Provider Information
    • Sub-processors
    • SOC 2 Type II
    • Status

    Compare

    • All Comparisons
    • GitHub Copilot
    • OpenRouter
    • LiteLLM
    • Portkey
    • AWS Bedrock
    • Azure AI Foundry
    • Vercel AI Gateway
    • Migration Guides

    Models

    • Text Generation
    • Text to Image
    • Image to Image
    • Video Generation
    • Embeddings
    • Vision
    • Reasoning
    • Tool Calling
    • Web Search
    • Discounted
    • Best for Roleplay
    • Best for Coding
    • Best for Creative Writing
    • Best for Translation
    • Best for Math
    • Long Context
    • Cheapest
    • Open Source

    Providers

    • OpenAI
    • Anthropic
    • Google AI Studio
    • Google Vertex AI
    • Vertex AI (OpenAI-compatible)
    • Vertex AI (Anthropic)
    • Groq
    • Cerebras
    • xAI
    • DeepSeek
    • Alibaba Cloud
    • NovitaAI
    • AtlasCloud
    • AWS Bedrock
    • AWS Mantle
    • Azure
    • Azure AI Foundry
    • Z AI
    • Moonshot AI
    • Baidu
    • Perplexity
    • Nebius AI
    • Mistral AI
    • CanopyWave
    • Inference.net
    • Together AI
    • SCX.ai (Turbo)
    • SCX.ai
    • ByteDance
    • MiniMax
    • EmberCloud
    • Meta
    • Sakana AI
    • Xiaomi
    • DeepInfra
    • ElevenLabs
    • Runware
    • Gonka24
    • Fireworks AI

    © 2026 OffRail. All rights reserved.

    Browse models by use case

    • Best models for coding
    • Reasoning models
    • Best models for roleplay
    • Creative writing models
    • Translation models
    • Best models for math
    • Long context models
    • Cheapest models
    • Premium models
    • Open source models
    • Vision models
    • Tool-calling models
    • Web search models
    • Embedding models
    • Text generation models
    • Text-to-image models
    • Image editing models
    • Video generation models
    • Discounted models

    How to choose an AI model

    Start from the capability you need — reasoning, vision, tool calling, or long context — then compare price per million tokens and context window. The filters above narrow the directory, and each model's page lists provider availability, live pricing, and uptime. Not sure where to start? See which models developers actually run in production in the live rankings.

    Compare AI model pricing

    Prices are shown per million input and output tokens, exactly as providers publish them. Sort by price to find the cheapest models, or estimate a monthly bill for your traffic with the token cost calculator.

    Try a model before you integrate

    Every model here is callable through one OpenAI-compatible API — switch models by changing a single string. Chat with any of them first in the Lounge to compare quality, speed, and cost side by side.

    OffRail
    • Lounge
    • Models
    • Docs
    • Pricing
    • Lounge
    • Pricing
    • Docs
    • Models
      • AI Gateway
      • Lounge
      • Observability
      • Enterprise
      • Blog
      • Changelog
      • Integrations
      • Reliability
      • Guardrails
      • Providers
      • Partners
      • Rankings
      • Apps
      • Models
      • Model Timeline
      • Compare
      • Token Cost Calculator
      • Referral Program
      • MCP Server
      • AI SDK Provider
      • Guides
    Log InGet Started
    ★