Blog

Latest news and updates from PassingRight

A glowing decision switch on a central chip with light traces branching into three weighted paths, surrounded by probability dials and a routing signpost

An LLM Decision API That Returns Values, Not Text

Most classification and routing calls use a chat model, then parse prose and hope the JSON holds. The new /v1/systemone LLM decision API on LLM Gateway returns typed yes/no, choice, and score answers with calibrated probabilities, billed on input tokens only.

September 20, 2026
A glowing compass rose on a central chip with light traces rerouting from an old endpoint plug to a new one, beside dated source cards on a circuit board

The Perplexity Sonar API Retirement: What Changes

Perplexity retires its Sonar chat completions API on September 27, 2026. On LLM Gateway, perplexity/sonar moves to Perplexity's Agent API on September 25 with no changes to your requests and lower prices, while sonar-pro and sonar-reasoning-pro stop being routable on September 27. Here is exactly what changes, why, and what to do about it.

September 20, 2026
A glowing discount tag on a circuit board connected to seven AI model chips and a calendar

SCX Model Discount: 20% Off Through October 9

The SCX model discount brings 20% off seven selected models on LLM Gateway, starting when the Runware promotion ends on September 9 and now extended through October 9. See the eligible models and how to route requests through SCX.

September 9, 2026
A glowing airport control tower with a lit runway leading to it on a circuit board, surrounded by paper planes, luggage tags, coins, and a boarding-gate arch

Introducing Airside: List Your Models on LLM Gateway

Airside is the self-serve carrier console where LLM providers claim their listing on LLM Gateway, register models, file prices for review, and tune the discount and margin that win routed traffic. Listing costs a one-time $2,500 fee per provider company, and every model is verified live before it goes on the departure board.

September 5, 2026
A circuit board with a glowing doorway on the central chip, representing a drop-in Vercel AI Gateway alternative for the AI SDK

A Vercel AI Gateway Alternative Without the Rewrite

Switching off the Vercel AI Gateway used to mean rewriting how your AI SDK app resolves models — and losing provider-native web search on the way out. LLM Gateway now implements the AI SDK's own gateway protocol, so the migration is one line and your bare model strings keep working.

August 12, 2026
A glossy 3D circuit board with a glowing folder holding documents at its center, representing a project knowledge base feeding AI chats

Projects: A Knowledge Base for Your AI Chats

LLM Gateway Chat now has Projects: group related chats, upload files — PDFs and spreadsheets included — as a knowledge base, and get answers grounded in your own documents via RAG, with source citations, on any of 280+ models. Projects also remember durable facts across chats. Available to every Chat user.

July 5, 2026
Q2 2026 Feature Roundup

Q2 2026: Speech, Embeddings & Coding Plans

Three months of updates: speech generation and Audio Studio, OpenAI-compatible embeddings, OCR, DevPass coding plans, chat subscription plans, enterprise IAM and master keys, SOC 2 Type II, 40+ new models, and much more.

June 30, 2026
Enterprise LLM analytics on LLM Gateway: cost, requests, and tokens broken down by model, project, API key, and team member

Enterprise LLM Analytics: See Where Every Dollar Goes

Most dashboards show what you spent, not where it went. LLM Gateway's enterprise LLM analytics break cost, requests, and tokens down by model, project, API key, and team member — no data warehouse to build. Member and organization-wide analytics are available on the Enterprise plan.

June 26, 2026