Nvidia has open-sourced PersonaPlex, a voice AI model that listens and speaks at the same time. The 7-billion-parameter model switches speakers in 0.07 seconds, against 1.3 seconds for Google’s Gemini Live in the same tests, and its licence allows commercial use without paying Nvidia a cent.
