Pardon · MOSS noise playground

Record up to one minute, add noise, and compare every MOSS-Transcribe-Diarize 0.9B transcription with speaker labels and segment timestamps.

The full recording is used. Maximum length: 60 seconds.

2 16
1 8
10 120
64 2048
Phonetic comparison language
Mandarin tones
Listen to full recording / noise variant

Ready. Default: 4 noise samples around 2 dB + original, maximum batch 4.

Every transcription

At ~2 dB, noise is a strong stress test. Stability is not correctness: all samples can share the same wrong answer. Four noisy samples give only six pairwise comparisons. Phonetic scores exclude the original and are withheld on incomplete runs. Timestamps are model-predicted segment boundaries, not exact uncertain-word boundaries. Speaker IDs are local to each rollout and can change between variants. Mandarin uses dictionary pinyin (tones ignored by default); English uses eSpeak IPA. Disagreement scores ignore punctuation; probability curves retain it. Phonetic scores use speech text only; curves include generated timestamps and speaker labels. Audio and results are temporary server files, not published as a dataset; download runs you want to keep.