pyannoteAI, Gladia, and Modal are getting together for an evening on the engineering behind production voice AI.
Voice applications stack multiple layers: diarization, transcription, and the compute infrastructure underneath. We'll get into the trade-offs that show up when these layers meet real production workloads:
-
Designing the voice stack across speech intelligence and infrastructure
-
Managing latency, reliability, and cost
-
Getting from a working demo to something that holds up in production
This is for engineers and builders working on real voice products. Come with the problem you've been stuck on - there'll be plenty of people who've hit the same wall.
Agenda
-
7:00PM: Arrival: drinks, and mingle with other engineers
-
7:15PM: Panel & demo
-
7:50PM: Open Q&A
-
8:00PM: Food, drinks & more chat