pyannoteAI
pyannoteAI is a speaker intelligence platform that decodes who speaks, when, and how in real-world audio, powering voice AI products built for developers and regulated industries.
Publisher details

Pyannote is the speaker intelligence layer for human conversations. Where STT and existing Voice AI models tell you what was said, Pyannote tells you who said it, when, and how. That is the foundational understanding every pipeline needs.
Built on more than a decade of dedicated research and the most accurate diarization (speaker recognition) models in the world, and available as open source, a simple API, or on your own infrastructure, Pyannote turns every conversation into structured, reliable, private metadata, wherever it's captured. So any machine can understand what is happening in a human conversation.
We are the infrastructure layer. Built diarization-first, Pyannote is the base layer of the pipeline. It pairs with any STT, TTS, or Voice AI models and makes everything downstream work.