Diarization — Speaker Labels
Automatically identify and label different speakers in your meetings.
Diarization requires a Pro subscription. Upgrade →
What Is Diarization?
Diarization is the process of identifying who said what in a meeting. Instead of a flat transcript, you get:
Rather than:
How Clearminutes Labels Speakers
Clearminutes uses machine learning to identify speakers based on:
- Voice characteristics: Pitch, tone, speech patterns (entirely local, no external analysis)
- Audio segmentation: Detecting speaker changes in the waveform
- Consistency: Assuming the same voice throughout = same speaker
All processing happens on your device. No audio is sent to external services.
Enable Diarization
Diarization is enabled by default for Pro users. To verify:
- Go to Settings → Transcription
- Check Diarization is toggled On
- Your next recording will have speaker labels
If diarization is off, toggle it on and re-record or regenerate transcripts.
[SCREENSHOT: Diarization toggle in Settings]
How to Use Speaker Labels
View Labels in Transcript
Open any meeting with diarization enabled. Each segment shows the speaker:
Rename Speakers
Replace "Speaker 1", "Speaker 2" with actual names:
- Click on a speaker label in the transcript
- Type the person's name
- Press Enter
[SCREENSHOT: Inline speaker name editing]
Clearminutes remembers the name and applies it to other segments of the same speaker throughout the meeting.
Auto-Detect Speaker Names (Pro)
Auto-detect speaker names requires a Pro subscription. Upgrade →
After diarization labels speakers, Clearminutes can automatically suggest real names:
- After recording stops, Clearminutes analyses the transcript for name mentions
- Suggested names appear next to each "Speaker X" label
- Review and confirm or correct each suggestion
- Confirmed names replace "Speaker 1", "Speaker 2" throughout the transcript
This works best when participants introduce themselves or are mentioned by name during the meeting.
Rename All Instances
Change a speaker name globally:
- Click Edit → Rename Speaker
- Select the speaker (or type their current label)
- Enter new name
- Click Apply to All
All segments from that speaker update instantly.
Merge Speakers
If the same person was labeled as two different speakers:
- Click Edit → Merge Speakers
- Select speaker 1 and speaker 2
- Choose which name to keep
- All segments merge under one speaker
Split Speakers
If different people were labeled as one speaker:
- Click in the transcript where the speaker changes
- Click Edit → Change Speaker
- Select the new speaker name
- Confirm
Accuracy and Tips
Diarization accuracy depends on several factors:
Best Conditions for Accurate Diarization
- Distinct voices: Obvious differences between speakers (tone, pitch)
- Clear audio: No background noise, wind, or distortion
- Consistent speaking patterns: Speakers talk in turns (not overlapping)
- Few speakers: 2-3 speakers is most accurate; 5+ becomes harder
- Good microphone: Professional audio input (USB headset is better than laptop mic)
Challenging Conditions
- Similar voices: People with similar accents or vocal characteristics — diarization may confuse them
- Noisy environment: Loud background noise makes it harder to distinguish speakers
- Lots of overlap: People interrupting each other — labels may swap or merge
- Background speakers: People off-screen contributing (may be missed entirely)
When to Disable Diarization
Consider disabling if:
- You're recording a presentation (only one speaker, unnecessary processing)
- Audio quality is very poor (may create false labels)
- You have many similar voices (confusing labels)
- You need maximum processing speed (diarization adds 10-20% processing time)
To disable:
- Go to Settings → Transcription
- Toggle Diarization to Off
- Re-generate transcript (or record new meeting)
Speaker Labels in Summaries
When diarization is enabled, summaries include speaker context:
Without diarization, speaker context is lost:
Speaker Labels in Exports
When you export a meeting with speaker labels:
Markdown:
Plain Text:
Both formats include speaker names.
Talk-Time Breakdown
View how much each speaker talked:
[SCREENSHOT: Pie chart showing speaker percentages]
See this in:
- Analytics dashboard (Settings → Analytics)
- Meeting details (click a meeting to see speaker summary)
- Summary (includes speaker statistics)
This helps identify:
- Dominant speakers (talking 70%+ of time)
- Quiet participants (less than 20%)
- Balanced discussions (roughly equal talk time)
Accuracy Issues & Fixes
Speaker labels are wrong (mixing two people)
Cause: Similar voices or overlapping speech confused the model.
Fix:
- Manually correct labels in transcript
- Use Rename Speaker to fix all instances
- Lower background noise in your setup for next recording
- Use microphones that capture distinct voices better
Missed some speakers
Cause: Background speakers or low-volume participants not captured.
Fix:
- Manually add missing speaker labels
- Ensure all participants are close to microphone
- Reduce background noise
- Try recording with better audio equipment
Too many false labels
Cause: Background noise, music, or echoes interpreted as speakers.
Fix:
- Enable Noise Suppression in Settings → Recording
- Record in a quieter environment
- Use a directional microphone pointed at speakers
- Disable diarization if quality is poor
Diarization and Privacy
Diarization:
- Runs entirely on your device
- Does not identify people by name
- Does not use facial recognition
- Does not connect to external services
- Voice characteristics are analyzed locally only
You maintain full control over speaker labels — Clearminutes doesn't know or store identity information.
Performance Impact
Diarization adds processing time:
- Recording: No impact (happens during playback)
- Transcription: +10-20% time (depends on meeting length and audio quality)
- CPU usage: Moderate increase during processing
- GPU acceleration: Speeds up diarization significantly
Example timings (30-min meeting, large-v3 model):
- Without diarization: 4-5 minutes
- With diarization: 5-6 minutes
Related
- Advanced Analytics — View speaker talk-time breakdown
- Edit Transcripts — Manually fix speaker labels
- Generate Summaries — Summaries include speaker attribution