#streaming
- Thinking Machines Lab Unveils TML-Interaction-Small: 276B MoE Real-Time Multimodal Model Thinking Machines Lab models-llm
- OpenAI GPT-Live: Full-Duplex Voice Models Replace Advanced Voice Mode OpenAI audio
- MiniCPM-o 4.5: Real-Time Full-Duplex Omni-Modal AI on Edge Devices OpenBMB / Tsinghua University research
- ElevenLabs Launches ElevenMusic: AI Music Creation, Remixing, and Streaming in One Platform ElevenLabs audio
- OpenAI brings ChatGPT Voice to the desktop app with Codex control OpenAI tools
- Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model Microsoft research
- JoyAI-VL-Interaction: Open-Source 8B Real-Time VLM with Autonomous Turn-Taking JD.com research
- Audio Interaction Model: Unified Streaming Framework Combining Offline and Real-Time Audio Instruction Following research
- Wan-Streamer v0.1: End-to-End Real-Time Interactive Foundation Model Under 550ms Latency Wan-AI research
- ElevenLabs Launches ElevenMusic: AI Music Creation, Remixing, and Streaming Platform ElevenLabs audio
- LangChain Stack: Provider-Agnostic Content Block Token Callbacks for Anthropic, Groq, Mistral LangChain tools
- llama.cpp b9754: Real-Time Model Load Progress via SSE and PEG Grammar Parser tools
- Claude Code v2.1.199: Stacked Slash-Skill Invocations and Streaming Reliability Anthropic tools
- GitHub Copilot CLI Drops PAT Requirement in Actions; Agent Session Streaming in Preview GitHub tools
- LightMem-Ego: Lightweight Streaming Multimodal Memory for Smartphones and AI Glasses ZJUNLP research
- Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model Microsoft research