Voxa.Audio.Diarization
Speaker diarization for the Voxa pipeline (VLS-005) — "who spoke when", filling TranscriptionFrame.SpeakerId. Defines the ISpeakerSegmentation / ISpeakerEmbedding / IDiarizer seams and a pure-C# DiarizationPipeline (constrained agglomerative clustering by cosine distance — no ML runtime in the orchestration). Ships the seams + clustering; the reference ONNX impls (Pyannote segmentation + WeSpeaker embedding) are a separate opt-in package.
Activity
- Latest release
- 2mo ago
- Total releases
- 4
- Cadence
- ~daily
- Last 12 months
- 4
Details
- License
- MIT
- First release
- Jun 22, 2026
Releases
| Version | Released | |
|---|---|---|
0.7.2-alpha
pre
|
0.7.2-alpha
pre
|
|
0.7.1-alpha
pre
|
0.7.1-alpha
pre
|
|
0.7.0-alpha
pre
|
0.7.0-alpha
pre
|
|
0.6.0-alpha
pre
|
0.6.0-alpha
pre
|