StepFun

@StepFun_ai

Introducing StepAudio 3, our new family of 5 audio models for real-time voice, speech recognition, speech generation, audio generation and music. Realtime ranks #1 on Artificial Analysis for both Conversational Dynamics (98.9%) and Speech Reasoning (99.7%). ASR reaches 1.7% WER, matching the best result on the leaderboard. Build voice agents that handle interruptions, reason while speaking, and call tools. Transcribe speech, generate expressive voices, and create full audio scenes and music. Available now: Voice AI Lab: https://t.co/9lQaKPc5AF Blog: https://t.co/cU0cQvI7zM
打开原帖#511482
  1. Frontier

    Logan Kilpatrick: Say hello (literally) to Gemini 3.8 Live and 3.8 Live Extended Thinki…
  2. Frontier

    Guillermo Rauch: Excited to formally introduce Vercel Labs
  3. Industry

    Rohan Paul: Google released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking…