New Models
Gemini 3.8 Live & Live Extended Thinking
Gemini 3.8 Live claims #1 on the speech-to-speech index with 97 languages and async tool calls
Google launched Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, real-time speech-to-speech models scoring 82.6 to take #1 on the speech-to-speech quality index across 97 languages with asynchronous tool calls. Extended Thinking brings longer reasoning into a live model without breaking real-time interaction, and the release reaches everyone on Android and Google Search rather than just API users.
82.6 #1 on the speech-to-speech quality index97 languages
New Models
Union Alpha
Union Alpha: anonymous stealth model free on OpenRouter with 262K context
A new anonymous stealth model called Union Alpha appeared for free on OpenRouter with a 262K context window, processing over 100B tokens within hours of listing. The community's leading guess for the lab behind it is Z.ai, but the provider remains unconfirmed.
262K context window100B+ tokens processed within hours
New Models
StepAudio 3
StepFun's StepAudio 3 family takes #1 on Artificial Analysis real-time voice
StepFun launched StepAudio 3, a five-model audio family spanning Real-Time Preview, ASR Max, and TTS — a full suite for building end-to-end voice assistants in code. Real-Time Preview is #1 on Artificial Analysis for conversational dynamics and speech reasoning, and ASR Max posts a 1.7% word error rate, significantly below Whisper. API only, no open weights.
#1 Artificial Analysis real-time voice1.7% ASR Max word error rate5 models in the family
New Models
Jev
TypeSafe AI launches Jev, the first public non-LLM System 1 decision model
TypeSafe AI, the stealth lab led by RLHF and ChatGPT co-creator Diogo Almeida, launched Jev — the first public System 1 model, a new class of AI that returns calibrated probabilities instead of generating text. Developers define questions with three primitives (choice, score, null) in natural language and get typed, machine-readable probability outputs back in 70-500ms with a 32K context window, trained with what TypeSafe calls RLCD (reinforcement learning for calibrated decisions). Pricing is $42 per billion input tokens with free output tokens, and TypeSafe cites 133x faster and 444x cheaper than competitive-level LLMs on tested decision tasks. Access is via waitlist.
$42 / 1B input token pricing; output tokens free70-500ms decision latency133x / 444x faster / cheaper vs competitive-level models (TypeSafe)