Join us for an evening bringing together founders, engineers, researchers, and product teams shaping the next generation of Voice AI.
Introducing StepAudio 3
Xuerui Yang, Research Lead at StepAudio, will introduce StepAudio 3, StepFun’s next-generation audio model family, and demo how it goes beyond words to better understand voice, sound, context, and the natural rhythm of conversation. We’ll also open the floor for a live AMA with the StepFun audio model and product team.
Panel: The Next Interface for AI, From Voice to Real-Time Interaction
♦️ Yijia Zhang, Partner & Head of AI Algorithm & Platform at PLAUD.AI♦️ Kobi Hudson, Head of Engineering at Coval♦️ Jingwen Gu, Maintainer of SGLang-Omni
Another panelist from Cresta will be announced shortly.
What does it actually take to turn better voice models into AI products that feel natural, responsive, and useful? This panel will bring together perspectives across models, real-time infrastructure, evaluation, product experience, and enterprise deployment to unpack what’s working, what still breaks in production, what users actually care about, and where Voice AI and real-time agents are heading next.
What to Expect
🚀 StepAudio 3 launch + Live AMA 🗣️ A cross-stack Voice AI panel : from models to real-world products🎁 $100 StepAudio API credits for in-person attendees🥂 Drinks, light dinner & networking
Who Should Join
Founders, engineers, researchers, product leaders, enterprise teams, and investors building or exploring voice agents, audio models, multimodal products, intelligent devices, and human–AI interaction.
Agenda
6:00 PM, Check-in, light dinner, drinks & mingle6:40 PM, Introducing StepAudio 3, Live Demos & AMA7:25 PM, The Next Interface for AI: From Voice to Real-Time Interaction, Panel & Audience Q&A8:10–9:00 PM, Networking
About StepFun
StepFun is a leading foundation model company committed to the long-term pursuit of AGI. Its Step model family spans language, audio, multimodal, and reasoning capabilities.
StepFun also works closely with intelligent devices, with its models deployed across smartphones, automotive, and in-car agent experiences.