Google introduced real-time simultaneous interpretation via Gemini 3.5 'Live Translate.' The signal centers on official DeepMind sources with community (simon_willison) and arXiv involvement—an official announcement drawing developer-community reaction. The focus is a cross-conversation interpreting experience that translates speech back with low latency; the through-line is a shift from a translation-accuracy race toward the experience value of real-time responsiveness. It reads as multimodal models extending from 'read and write' to 'listen and translate on the spot.' But this is mainly a feature announcement—real-use latency and accuracy, language coverage, and the gap versus existing interpreting await validation.
Google ships Gemini Live Translate
Google ships Gemini Live Translate
Fluid, natural voice translation with Gemini 3.5 Live Translate
Google debuts Gemini 3.5 Live Translate for real-time speech
Google releases DiffusionGemma as an open-weight diffusion LLM
Willison weighs WWDC 2026 Siri: Gemini-derived model on Apple's PCC
Academic (arxiv etc.) 5 ▾
MSUE: Multi-Modal Soccer Understanding Expert
MSUE multi-expert architecture takes third in 2026 SoccerNet VQA
The Shibboleth Effect: Auditing the Cross-Lingual Distributional Skew of Large Language Models
The Shibboleth Effect: language shifts LLM behavior in a wargame
A History-Aware Visually Grounded Critic for Computer Use Agents
HiViG: a history-aware, visually grounded critic for computer-use agents
Auditing cultural translation of math word problems across LLMs
Beyond APIs: Probing the Limits of MLLMs in Physical Tool Use
PhysTool-Bench probes the limits of MLLMs in physical tool use