2026
- 1月5日 - vLLM Semantic Router v0.1 Iris: The First Major Release
- 1月23日 - Building Mixture-of-Models on AMD GPUs with vLLM-SR
- 3月10日 - vLLM Semantic Router v0.2 Athena: ClawOS, Model Refresh, and the System Brain
- 3月12日 - v0.3 Themis Roadmap: Stability at Scale
- 3月25日 - Deploying vLLM Semantic Router on AMD Developer Cloud
- 5月28日 - From Text to Multimodal Routing: Hardening Vision Signals in vLLM Semantic Router
- 6月2日 - Session-Aware Agentic Routing: Continuity-Aware Model Selection for Long-Horizon LLM Agents
- 6月5日 - vLLM Semantic Router v0.3 Themis: From Signals to Stateful Production Routing
- 6月16日 - Beyond One Model: Fusion in vLLM Semantic Router
- 6月18日 - Agentic Routing on AMD ROCm
- 6月28日 - Giving AgentGateway a Semantic Brain with vLLM Semantic Router
- 6月29日 - Micro-Agent: Beat Frontier Models with Collaboration inside Model API
- 7月9日 - Adding Cursor-Style Auto Model Selection to OpenCode with vLLM Semantic Router
- 7月21日 - Beyond a Single Model: Building Mixture-of-Models Systems with vLLM Semantic Router
- 8月5日 - LettuceDetect v2 in Semantic Router: Generative Hallucination Detection as a vLLM Endpoint
- 8月24日 - Find Your Focus: How to Join and Work Together
- 9月18日 - Introducing Vela 1.0
- 9月22日 - Introducing Decision 1.0: Open Decision Foundation Models
- 9月24日 - vLLM Semantic Router v0.4 Hermes: Many Models, One Improving System
- 9月25日 - Per-Call Model Selection: Measuring Cost and Quality Across a Multi-Agent Run
- 10月2日 - Beyond Prompt Routing: Model Selection, Conversation State, and KV Cache
- 10月6日 - Vela 2.0: Open Foundation Routing Models
- 10月10日 - Decision Models Need a Router, Too
2025
- 9月11日 - vLLM Semantic Router: Next Phase in LLM inference
- 10月20日 - Semantic Router Q4 2025 Roadmap: Journey to Iris
- 10月27日 - From Monolithic to Modular: Scaling Semantic Routing with Extensible LoRA
- 11月7日 - Semantic Tool Selection: Building Smarter AI Agents with Context-Aware Routing
- 11月19日 - Signal-Decision Driven Architecture: Reshaping Semantic Routing at Scale
- 12月14日 - Token-Level Truth: Real-Time Hallucination Detection for Production LLMs
- 12月16日 - AMD × vLLM Semantic Router: Building the System Intelligence Together