TIMIT
PulseAugur coverage of TIMIT — every cluster mentioning TIMIT across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
CrisperWhisper 2.0 released with verbatim and intended transcription modes
Nyra Labs has released CrisperWhisper 2.0, an advanced speech-to-text model designed for production use. This model offers distinct modes for verbatim transcription, capturing every spoken detail, and an "intended" mode…
-
New gradient-based method aligns speech-to-text across all ASR models
Researchers have developed a novel gradient-based method for aligning speech-to-text, applicable to any differentiable automatic speech recognition (ASR) model. This technique derives word timings from the gradient of t…
-
New method improves multilingual word-level speech alignment
Researchers have developed a novel method for multilingual word-level forced alignment, integrating representations from the Massively Multilingual Speech (MMS) model and a self-supervised phoneme boundary detector. Thi…
-
New acoustic models achieve SOTA on TIMIT phonetic recognition
Researchers have analyzed the error patterns of raw waveform acoustic models used for phonetic recognition on the TIMIT dataset. They decomposed the phone error rate (PER) across phonetic categories and constructed conf…