Blog

Meeting Intelligence

Uzbek Meeting Intelligence: Speaker Diarization and AI Summaries

Transform Uzbek business meetings, calls, and interviews into structured transcripts with millisecond-accurate speaker diarization and automated executive summaries.

·Biruniy Research

Business meetings, call center discussions, and executive briefings in Uzbekistan involve multiple speakers, rapid turn-taking, overlapping voices, and natural switching between Uzbek and Russian.

Standard transcription models treat audio as a single continuous stream of text, losing the critical answer to: Who said what, and when?

Today, Biruniy delivers end-to-end Uzbek Meeting Intelligence combining state-of-the-art speaker diarization, Nava ASR transcription, and structured meeting summarization.

The multi-speaker challenge in Uzbek meetings

Transcribing a meeting accurately requires solving three distinct challenges simultaneously:

  1. Speaker Diarization: Identifying distinct voice embeddings to partition audio into speaker-specific segments (Speaker 1, Speaker 2, etc.).
  2. Turn-Taking and Overlap Handling: Detecting when one speaker interrupts or interjects without losing words from either participant.
  3. Dialect and Code-Switching Continuity: Preserving speaker context even when individual participants switch between regional dialects or blend Russian and Uzbek technical terms.

Fine-tuned Pyannote 3.1 on Uzbek conversational speech

Biruniy's meeting intelligence pipeline utilizes a customized Pyannote 3.1 neural diarization architecture fine-tuned on our 200+ hour Gold dataset with over 100+ unique annotated speakers.

Key capabilities include:

  • Diarization Error Rate (DER) < 8%: Accurately tracks speaker boundaries and speaker re-identification throughout long multi-hour recordings.
  • Millisecond-level word alignment: Every transcribed word is directly mapped to its speaker ID and time boundary.
  • Voice Activity Detection (VAD): Suppresses room background noise, keyboard typing, and HVAC hum before audio reaches the recognition engine.

Actionable meeting intelligence in Biruniy Studio

When processed through Biruniy Studio, multi-speaker Uzbek recordings are automatically converted into executive-ready assets:

  • Speaker-Attributed Transcripts: Full readable transcript with speaker avatars, timestamps, and search capability.
  • Talk Time & Participation Analytics: Visual breakdown of speaker contribution percentage, turn frequency, and pace.
  • AI Executive Summaries: Key decisions, discussion topics, and next steps extracted in Uzbek, Russian, or English.
  • Export Formats: Download as JSON, SRT subtitles, PDF meeting minutes, or integrate via REST API into internal enterprise tools.

Enterprise deployment & air-gapped security

For financial institutions, government bodies, and legal teams handling sensitive conversations, Biruniy Meeting Intelligence can be deployed completely on-premise in your private cloud or air-gapped data centers. Audio never leaves your infrastructure.

Get started

Try uploading your meeting recordings in Biruniy Studio or contact our enterprise team for on-premise deployment and custom integrations.