-
ElevenLabs Surpasses $500M ARR, Adds BlackRock and Nvidia to Series D
ElevenLabs
audio
-
OpenAI Launches GPT-Realtime-2, GPT-Realtime-Translate, and GPT-Realtime-Whisper
OpenAI
audio
-
ElevenLabs Surpasses $500M ARR and Closes Series D with BlackRock and NVIDIA
ElevenLabs
audio
-
ByteDance Launches Seed-Audio 1.0: Unified Speech, Music, and Ambient Sound Generation
ByteDance
audio
-
NVIDIA Releases Nemotron-Labs-Audex-30B-A3B: Unified Audio-Text MoE Model
NVIDIA
audio
-
ElevenLabs Deploys Google DeepMind SynthID Audio Watermarking to Free-Tier Users
ElevenLabs
audio
-
Runway launches Media Router, an automatic model-selection API for generative media
Runway
tools
-
Pika launches Pika Audio: four frontier foundation sound models at up to 20x lower cost
Pika
audio
-
Pika Audio Models: Soundtrack, Music, SFX, and Speech
Pika Labs
audio
-
Cartesia launches Sonic-3.6 TTS, claims top Elo over Eleven v3
Cartesia
audio
-
Google launches Gemini 3.8 Flash TTS and Flash-Lite TTS with prompt-based voice design
Google DeepMind
audio
-
Tencent open-sources AuK, a 1.5B speech foundation model for generation and editing
Tencent Hunyuan
audio
-
ElevenLabs Launches Avatars in ElevenCreative: TTS-Native AI Talking-Head Video
ElevenLabs
video
-
Gemini 3.8 TTS: custom voice creation and 100+ language synthesis
Google DeepMind
audio
-
ElevenLabs Licenses Stan Lee's Voice and Likeness for AI Commercial Use
ElevenLabs
audio
-
xAI Expands Grok Voice with 21 New Multilingual Flagship Voices
xAI
audio
-
xAI Grok Voice Becomes Default Engine for Vapi's 2.5M+ Voice Agents
xAI
audio
-
Navana.ai launches Bodhi TTS voice AI model for 10 Indian languages
Navana.ai
audio