Whisper Diarization Post-Processor
Enhances OpenAI Whisper transcription output with speaker diarization using pyannote.audio pipeline and speechbrain embeddings. Aligns word-level timestamps from whisper-timestamped with speaker segments for multi-speaker meeting transcript generation.
Whisper Diarization Post-Processor
Enhances OpenAI Whisper transcription output with speaker diarization using pyannote.audio pipeline and speechbrain embeddings. Aligns word-level timestamps from whisper-timestamped with speaker segments for multi-speaker meeting transcript generation.
Installation
Method 1, Agent Skill Exchange
- Install from the marketplace listing: https://agentskillexchange.com/skills/whisper-diarization-post-processor/
Method 2, Git clone
git clone https://github.com/agentskillexchange/skills.git && cd skills/skills/whisper-diarization-post-processor
Method 3, Download ZIP
- Download the repository ZIP and extract
skills/whisper-diarization-post-processor.
Method 4, Manual copy
- Copy this skill folder into your local skills directory, then reload your agent tooling.
Method 5, Fork and sync
- Fork the repository if you want to maintain local edits while syncing upstream changes.