Transcription Problems
Troubleshoot issues with Whisper transcript accuracy and processing.
Transcript Is Empty
Transcription completed but no text appears.
Check Whisper Model Downloaded
- Go to Settings → Transcription → Model
- Look for a green checkmark or "Downloaded" label next to your model
- If not downloaded:
- Click Download next to the model
- Wait for download to complete (can take 5-30 minutes depending on model size)
- Retry transcription
VAD Sensitivity Too Aggressive
Voice Activity Detection (VAD) may have filtered out all audio as silence:
- Go to Settings → Recording → VAD Sensitivity
- Lower the sensitivity (move slider left toward "Low")
- Re-record test segment or regenerate transcript
- Check if speech is now captured
Note: Very low voices or soft speakers are more likely to be filtered. Test with normal speaking volume.
Audio Recording Is Actually Silent
Verify the original audio wasn't captured:
- Open the meeting
- Click any transcript segment
- Press Play to hear original audio
- If silent, see Recording Issues →
If audio plays but transcript is empty, Whisper model isn't working. See below.
Whisper Model Needs Reinstall
The model file may be corrupted:
- Go to Settings → Transcription → Model
- Click the gear icon next to your model
- Click Delete
- Click Download to re-download
- Retry transcription
Transcript Is Garbled or Gibberish
Transcription contains nonsensical words, wrong language, or corrupted text.
Check Audio Quality
- Play back the recording: Open meeting → click segment → Play
- Listen for clarity
- If audio is distorted, crackling, or very quiet:
- See Recording Issues →
- Adjust microphone gain
- Retry in quieter environment
Try a Larger Whisper Model
Larger models are more accurate and better at noisy audio:
- Go to Settings → Transcription → Model
- Choose a larger model:
- Upgrade:
tiny→base→small→medium→large-v3
- Upgrade:
- Download the new model
- Regenerate transcript
Accuracy vs Speed tradeoff:
tiny/base: Faster but less accurate (better for near-perfect audio)medium/large-v3: More accurate (can handle noisy audio) but slower
Set Language Manually
If Whisper auto-detected wrong language:
- Go to Settings → Transcription → Language
- Choose the correct language (e.g. English, Spanish, French)
- Regenerate transcript
Auto-detection fails sometimes with mixed languages or heavy accents.
Lower VAD Sensitivity
VAD filtering might be cutting out speech:
- Go to Settings → Recording → VAD Sensitivity
- Lower sensitivity to Low
- Regenerate transcript
- Check if previously missing words now appear
Transcript Is Very Slow
Transcription is taking longer than expected (over 2-3 minutes for 30-min meeting).
Use Smaller Model
Larger models take exponentially longer. Switch to a smaller one:
- Go to Settings → Transcription → Model
- Choose smaller model:
large-v3(slowest, best quality) →medium(good balance)medium→small(faster)small→baseortiny(fastest)
Typical times for 30-minute meeting:
tiny: 5 seconds (CPU), 2 seconds (GPU)base: 15 seconds (CPU), 5 seconds (GPU)small: 45 seconds (CPU), 12 seconds (GPU)medium: 2 minutes (CPU), 20 seconds (GPU)large-v3: 5 minutes (CPU), 30-45 seconds (GPU)
Enable GPU Acceleration
GPU is 5-10x faster than CPU. Check if it's enabled:
- Go to Settings → Advanced → System Info
- Look for GPU Status
- Should show "Metal enabled" (macOS) or "CUDA/Vulkan enabled" (Windows)
If showing "CPU only":
- See GPU Acceleration → for setup
Enable VAD
VAD filters silence and reduces audio sent to Whisper (~70% reduction):
- Go to Settings → Recording → VAD
- Toggle On
- Set sensitivity to Medium or High
- Regenerate transcript
With VAD enabled, transcription is 2-3x faster.
Close Other Applications
Free up system resources:
- Close browser tabs (Chrome, Firefox use lots of memory)
- Close large apps (photo/video editors, IDE)
- Close other music apps
- Restart computer if system is sluggish
GPU acceleration speeds up if system is less busy.
Check System Memory
If Clearminutes is running out of memory:
- Open Task Manager (Windows) or Activity Monitor (macOS)
- Look for "Clearminutes" or "app_lib" process
- If "Memory" shows 4+ GB, system is under strain
- Close other apps or increase available RAM
Transcript Misses Words
Transcript is incomplete — whole words or phrases are missing.
Lower VAD Sensitivity
VAD may be filtering speech as silence:
- Go to Settings → Recording → VAD Sensitivity
- Lower to Low (more permissive)
- Regenerate transcript
- Check if previously missing words return
Trade-off: Lower VAD may include more background noise.
Try a Larger Whisper Model
Smaller models sometimes miss quiet words. Upgrade model:
- Go to Settings → Transcription → Model
- Choose larger model (e.g.
mediumorlarge-v3) - Download and regenerate transcript
Check Microphone Levels
Words may be missing because they were recorded too quietly:
- Go to Settings → Recording → Microphone Gain
- Increase to 70-80%
- Re-record test segment
- Check transcription quality
Manually Add Missing Words
You can edit the transcript directly:
- Open meeting
- Click segment where word is missing
- Click to edit and add the word
- Save
Use Cmd+Z (macOS) or Ctrl+Z (Windows) to undo if needed.
Wrong Language Detected
Transcript is in wrong language or garbled because of language auto-detection.
Set Language Manually
- Go to Settings → Transcription → Language
- Select correct language
- Regenerate transcript
For Mixed Language Meetings
If meeting has multiple languages:
- Set the primary language in settings
- Manually edit any segments in other languages
- Consider recording multiple meetings if heavily multilingual
Whisper works best with one primary language.
Diacritics and Special Characters Missing
Accents, umlauts, or special characters not appearing in transcript.
This Is a Whisper Limitation
Some non-English characters may be:
- Converted to their base letter (é → e)
- Omitted entirely
- Substituted with similar characters
Workaround:
- Edit transcript manually to add correct characters
- Use Find & Replace to fix recurring issues
Try Large-v3 Model
The latest large-v3 model handles non-English characters better:
- Go to Settings → Transcription → Model
- Download and switch to
large-v3 - Regenerate transcript
Personal Names Not Recognized
People's names are misspelled or not recognized.
Whisper Limitation
Whisper sometimes misspells proper names and technical terms:
- "TensorFlow" → "tensorflow"
- "API" → "Api"
- Acronyms → expanded or capitalized incorrectly
Fix Names in Transcript
- Open meeting
- Use Find & Replace (Cmd+H / Ctrl+H)
- Find: "incorrect name", Replace: "correct name"
- Click Replace All
Rename Speakers (Pro)
If using diarization:
- Click speaker name in transcript
- Type correct name
- Apply to all segments of that speaker
Transcription Cuts Off Mid-Meeting
Transcript ends abruptly in the middle of recording.
Check Recording File
- Open meeting
- Click a segment near the end
- Press Play to verify audio continues
- If audio cuts off, see Recording Issues →
Regenerate Transcript
Sometimes transcription fails due to temporary issue:
- Open the meeting
- Click Regenerate Transcript
- Wait for processing to complete
- Check if full transcript appears
Check Disk Space
If disk was full during transcription:
- Go to Settings → Storage
- Free up space
- Regenerate transcript
Too Much Background Noise in Transcript
Transcript picks up background sounds as words (air conditioning, keyboard, etc.).
Enable Noise Suppression
- Go to Settings → Recording → Noise Suppression
- Toggle On
- Set to Moderate or Strong
- Re-record test segment
Raise VAD Sensitivity
More aggressive VAD filtering removes quiet background noise:
- Go to Settings → Recording → VAD Sensitivity
- Increase to High
- Regenerate transcript
Improve Recording Environment
- Close windows and doors
- Turn off HVAC/fans if possible
- Move away from air vents
- Use a directional microphone
Use Large-v3 Model
Larger models are better at distinguishing speech from noise:
- Go to Settings → Transcription → Model
- Download
large-v3 - Regenerate transcript
Regenerate a Transcript
To re-transcribe a meeting with different settings:
- Open the meeting
- Click Regenerate Transcript
- Choose new settings (different model, language, etc.)
- Wait for processing
- New transcript replaces old one
Note: Old transcript is saved in edit history (if you regenerated immediately). Undo if needed.
Check Processing Logs
If you're still having issues, check the logs:
- Go to Settings → Advanced → View Logs
- Look for error messages related to Whisper
- Screenshot errors and email to https://support.clearminutes.app
Include:
- Meeting duration
- Whisper model used
- Language setting
- Error message from logs
Related
- Recording Quality — VAD and audio preprocessing
- Model Selection — Choosing the right Whisper model
- GPU Acceleration — Speed up transcription
- Recording Issues — Microphone and audio capture problems