OpenRouter
Unified API marketplace for 200+ LLMs from OpenAI, Anthropic, Google, Mistral, Meta and more.
OpenRouter launches Batch API for half-price async inference while building out its decision model catalog.
◆Recent moves
- 2d ago
Is Kimi K3 Open Source? Weights, License, and How to Call It
- 3d ago
Best Embedding Models in 2026
- 3d ago
How to Use Jev: Moderation with the Jev API in TypeScript
End-to-end tutorial for using Jev on marketplace listing moderation, covering prompt structure, probability thresholds, and publish/hold/reject routing logic. Instructional content; no new platform capability.
- 4d ago
Is Jev as Accurate as Frontier Models at Classification?
Benchmark putting Jev 1.13 against Claude Opus 5 on Banking77 classification: Jev scores 81.0% vs 84.4% at 175ms and $0.11/thousand vs 2.3s and $2.42. Useful comparative data but a blog piece, not a platform change.
- 4d ago
What Is Nemotron 3.5 Lightning
Explainer on Nemotron 3.5 Lightning — NVIDIA's 30B MoE model with ~3B active parameters per token — covering architecture, endpoint features, and structured output usage via OpenRouter. Catalog documentation, not a platform change.
- 4d ago
Batch API launches: half-price inference for async workloads
⚡ SPARKThe Batch API ships out of beta: send a full workload in one POST, pay roughly half the per-token price, collect results within 24 hours. 230k+ batches completed during the beta with a median of 7 minutes — far shorter than the 24-hour SLA. This is the first time OpenRouter has offered a fundamentally different pricing tier based on latency tolerance, making bulk inference economics competitive with self-hosted options.