#real-time
- Thinking Machines Lab Unveils TML-Interaction-Small: 276B MoE Real-Time Multimodal Model Thinking Machines Lab models-llm
- Gemini 3.5 Live Translate: Real-Time Speech-to-Speech in 70+ Languages Google DeepMind audio
- Vidu S1: Real-Time Interactive Video Generation at 42 FPS on Consumer GPUs Shengshu Technology / Tsinghua University research
- ShengShu Technology Unveils Vidu S1: Real-Time Interactive Video on Consumer GPUs ShengShu Technology video
- OpenAI Releases GPT-Realtime-2.1 and GPT-Realtime-2.1-mini for Voice Agents OpenAI models-llm
- Alibaba Amap Launches ABot-World Studio: Infinite Interactive Video and 3D World Generation on a Single GPU Alibaba video
- Runway introduces Solaris, its first Interface World Model Runway video
- Runway previews real-time video generation: streaming video steered as it is described Runway video
- JoyAI-VL-Interaction: Open-Source 8B Real-Time VLM with Autonomous Turn-Taking JD.com research
- ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU Alibaba AMAP CV Lab research
- Vidu S2: real-time interactive, editable, and spatial video generation ShengShu Intelligence research
- Causal Forcing++: 2-Step Distillation Enables Real-Time Interactive Video Generation Tsinghua University research
- Runway introduces GWM Worlds 2, a real-time interactive world model Runway video
- Gander: end-to-end omni interaction agent with Cerebellum-Brain architecture Tencent Hunyuan research
- Wan-Streamer v0.1: End-to-End Real-Time Interactive Foundation Model Under 550ms Latency Wan-AI research
- Meta Superintelligence Labs releases Muse Voice Transcribe, a real-time speech model Meta audio
- Yandex Smart Camera Gains Visual Q&A via Alice AI VLM Yandex tools
- StepAudio 3 Realtime: think-while-speaking audio-language model with asynchronous tool calls StepFun research
- Yandex launches Yandex Sim mobile operator with call-transcribing AI assistant Yandex tools
- Yandex Maps AI Chat Expands to Cultural and Outdoor Leisure Recommendations Yandex tools