Support

AI-powered help

Welcome!

Please introduce yourself before we start.

    PassingRight
    • Models
    • Pricing
    • Docs
    • PassingRight Chat
    • Models
    • Pricing
    • Docs
    • PassingRight Chat
    Log inGet an API key

    Open Source Models

    Open-weight models — Llama, DeepSeek, Qwen, GLM, Kimi, GPT-OSS, Gemma, and more — served through one API

    Compare
    Use Case
    Capabilities
    Provider
    Status
    Input Price ($/M tokens)
    Output Price ($/M tokens)
    Context Size (tokens)
    74/283
    Models
    28/55
    Providers
    26
    Vision Models (filtered)
    64
    Tool-enabled (filtered)
    0
    Free Models (filtered)
    Features
    Xiaomi
    mimo-v2.6-flash
    $0.14$0.28$0.00
    Xiaomi
    mimo-v2.6-pro
    $0.43$0.87$0.00
    Alibaba Cloud(cn-beijing)
    deepseek-v4.1-flash
    $0.28$1.13$0.03
    Alibaba Cloud(us-virginia)
    deepseek-v4.1-flash
    $0.28$1.13$0.03
    Alibaba Cloud(eu-frankfurt)
    deepseek-v4.1-flash
    $0.28$1.13$0.03
    Alibaba Cloud(singapore)
    deepseek-v4.1-flash
    $0.30$1.20$0.03
    Alibaba Cloud
    deepseek-v4.1-flash
    $0.30$1.20$0.03
    DeepInfra
    deepseek-v4.1-flash
    $0.20$0.60$0.01
    Fireworks AI
    deepseek-v4.1-flash
    $0.22$0.66$0.01
    Together AI
    deepseek-v4.1-flash
    $0.30$1.20$0.01
    NovitaAI
    deepseek-v4.1-flash
    $0.30$1.20$0.01
    DeepSeek
    deepseek-v4.1-flash
    $0.15$0.13
    -15% off
    $0.60$0.51
    -15% off
    $0.00$0.00
    -15% off
    Baidu
    deepseek-v4.1-flash
    $0.30$1.20$0.01
    Z AI
    glm-5.3-flash
    $0.15$0.50$0.03
    Together AI
    glm-5.3-flash
    $0.15$0.50$0.03
    SCX.ai
    glm-5.3-flash
    $0.13$0.40$0.02
    NovitaAI
    glm-5.3-flash
    $0.15$0.50$0.03
    Runware
    glm-5.3-flash
    $0.15$0.50$0.03
    SCX.ai
    glm-5.2-fast
    $1.99$6.16$0.40
    Alibaba Cloud(eu-frankfurt)
    glm-5.3
    $1.13$3.96$0.23
    Alibaba Cloud(cn-beijing)
    glm-5.3
    $1.13$3.96$0.23
    Alibaba Cloud(singapore)
    glm-5.3
    $1.40$4.40$0.28
    Z AI
    glm-5.3
    $1.40$4.40$0.26
    Runware
    glm-5.3
    $1.20$4.00$0.20
    DeepInfra
    glm-5.3
    $1.20$4.00$0.20
    Together AI
    glm-5.3
    $1.40$4.40$0.26
    Baidu
    glm-5.3
    $1.40$4.40$0.26
    SCX.ai
    glm-5.3
    $1.30$4.00$0.25
    NovitaAI
    glm-5.3
    $1.40$4.40$0.26
    Alibaba Cloud
    glm-5.3
    $1.40$4.40$0.28
    DeepInfra
    nemotron-3.5-lightning
    $0.08$0.20—
    DeepInfra
    muse-glimmer-30b
    $0.30$1.20$0.04
    Together AI
    muse-glimmer-30b
    $0.35$1.50$0.04
    Fireworks AI
    kimi-k3-fast
    $4.50$22.50$0.45
    Alibaba Cloud(cn-beijing)
    glm-5.2
    $1.40$4.40$0.28
    Alibaba Cloud(us-virginia)
    glm-5.2
    $1.40$4.40$0.28
    Alibaba Cloud(eu-frankfurt)
    glm-5.2
    $1.40$4.40$0.28
    Alibaba Cloud(singapore)
    glm-5.2
    $1.40$4.40$0.28
    NovitaAI
    glm-5.2
    $1.40$4.40$0.26
    SCX.ai
    glm-5.2
    $0.55$1.78$0.11
    Runware
    glm-5.2
    $0.80$2.55$0.16
    Baidu
    glm-5.2
    $1.40$4.40$0.26
    Tencent Cloud
    glm-5.2
    $1.40$4.40$0.26
    Z AI
    glm-5.2
    $1.40$4.40$0.26
    ByteDance
    glm-5.2
    $1.40$4.40$0.26
    EmberCloud
    glm-5.2
    $1.26$3.96$0.23
    CanopyWave
    glm-5.2
    $1.40$4.40$0.26
    Alibaba Cloud
    glm-5.2
    $1.40$4.40$0.28
    NovitaAI
    kimi-k2.7-code
    $0.95$4.00$0.19
    Moonshot AI
    kimi-k2.7-code
    $0.95$4.00$0.19
    Page 1 of 5

    Open-weight models have closed most of the gap with proprietary frontiers: DeepSeek V4, Qwen3.7, GLM-5, Kimi K2, and MiniMax M3 sit near the top of real-world leaderboards, joined by OpenAI's GPT-OSS and Google's Gemma releases. Their weights are public — but running a 200B+ parameter model yourself means serious GPU infrastructure.

    This page lists open-weight models served by hosted providers, so you get the openness — inspectable weights, no lock-in, the option to self-host later — with API convenience. PassingRight itself is open source (AGPLv3) and self-hostable, so the whole stack can run on your terms.

    Frequently asked questions

    What is the best open source LLM?

    DeepSeek V4, Qwen3.7, GLM-5.2, Kimi K2.6, and MiniMax M3 are the current leaders, each within striking distance of proprietary frontier models. For smaller, hardware-friendly options, GPT-OSS 20B, Gemma 4, and Qwen3.5 9B are the standouts.

    What does 'open source' mean for LLMs?

    Usually 'open weight': the trained weights are downloadable, but licenses vary — some are Apache 2.0 or MIT, others (like the Llama license) carry usage restrictions, and training data is rarely published. Check the license of a specific model before building on it.

    Should I self-host or use an API?

    Self-hosting pays off with steady high volume, strict data-residency needs, or fine-tuned weights. For everything else, per-token APIs are cheaper than idle GPUs. A middle path: develop against hosted open models and keep self-hosting as an exit option, since the weights are public.

    Are open models cheaper than proprietary ones?

    Dramatically, per token. Competition among hosts drives prices down — DeepSeek V4 Flash and Qwen3 Coder 30B cost 10–50x less than frontier proprietary models. The list above shows every provider's price for each model.

    Newsletter

    Stay ahead of the curve

    Join developers who get weekly insights on LLM routing, new model launches, and cost optimization — straight to their inbox.

    • New models & providers as they drop
    • Tips to cut latency & costs
    • Early access to beta features

    No spam. Unsubscribe anytime.

    All systems operational
    AICPA SOC for Service Organizations badgeSOC 2 Type II
    compliant

    Product

    • Features
    • AI Gateway
    • Observability
    • Models
    • Providers
    • Rankings
    • Add Provider
    • Partners
    • PassingRight Chat
    • DevPass
    • Compare Models
    • Enterprise

    Resources

    • Legal Overview
    • Apps
    • Templates
    • Agents
    • MCP Server
    • Use Cases
    • Documentation
    • Developer resources
    • Integrations
    • Guides
    • Brand Assets
    • Token Cost Calculator
    • Copilot Cost Calculator
    • Referral Program
    • About
    • Contact Us

    Compliance

    • Trust Center
    • Security Portal
    • Terms
    • Privacy Policy
    • Provider Information
    • Sub-processors
    • SOC 2 Type II
    • Status

    Compare

    • All Comparisons
    • GitHub Copilot
    • OpenRouter
    • LiteLLM
    • Portkey
    • AWS Bedrock
    • Microsoft Foundry
    • Vercel AI Gateway
    • Migration Guides

    Models

    • Text Generation
    • Text to Image
    • Image to Image
    • Video Generation
    • Embeddings
    • Vision
    • Reasoning
    • Tool Calling
    • Web Search
    • Discounted
    • Best for Roleplay
    • Best for Coding
    • Best for Creative Writing
    • Best for Translation
    • Best for Math
    • Long Context
    • Cheapest
    • Open Source

    Providers

    • OpenAI
    • Anthropic
    • Google AI Studio
    • Google Vertex AI
    • Vertex AI (OpenAI-compatible)
    • Vertex AI (Anthropic)
    • Groq
    • Cerebras
    • xAI
    • DeepSeek
    • Alibaba Cloud
    • Runpod
    • NovitaAI
    • AtlasCloud
    • AWS Bedrock
    • AWS Mantle
    • Azure
    • Azure AI Foundry
    • Azure Anthropic
    • Z AI
    • Moonshot AI
    • Baidu
    • Perplexity
    • Mistral AI
    • CanopyWave
    • Inference.net
    • Together AI
    • SCX.ai (Turbo)
    • SCX.ai
    • ByteDance
    • MiniMax
    • EmberCloud
    • Meta
    • Meta Contributor
    • Sakana AI
    • Xiaomi
    • DeepInfra
    • ElevenLabs
    • Runware
    • Gonka24
    • Fireworks AI
    • RanoAI
    • Consensus Protocol
    • Tencent Cloud
    • Atria
    • TypeSafe AI

    © 2026 PassingRight. All rights reserved.