PulseAugur
EN
LIVE 15:54:24
ENTITY Whisper-small

Whisper-small

PulseAugur coverage of Whisper-small — every cluster mentioning Whisper-small across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
10
10 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
9
9 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

3 day(s) with sentiment data

RECENT · PAGE 1/1 · 10 TOTAL
  1. TOOL · CL_273429 ·

    New dataset MGhana-ST targets low-resource Ghanaian languages

    Researchers have introduced MGhana-ST, a new speech translation dataset designed for four low-resource Ghanaian languages: Ga, Twi, Ewe, and Fante. The dataset, which includes paired audio and English translations with …

  2. TOOL · CL_254602 ·

    Pruning Whisper-small improves ASR accuracy by acting as a regularizer

    Researchers have demonstrated that neural network pruning can act as a regularization technique for Automatic Speech Recognition (ASR) systems, rather than just a method for compression. By analyzing the sensitivity of …

  3. TOOL · CL_244925 ·

    New research probes acoustic information loss in audio-conditioned LLMs

    Researchers have investigated why audio-conditioned language models often fail to utilize crucial acoustic cues like prosody and emotion. Their study, detailed on arXiv, tested various audio encoders including Whisper-T…

  4. TOOL · CL_221030 ·

    Whisper model adapted for indigenous Baniwa language speech recognition

    Researchers have successfully adapted OpenAI's Whisper model to perform automatic speech recognition for the Baniwa language, an indigenous Arawakan language spoken across Brazil, Colombia, and Venezuela. Using a small …

  5. RESEARCH · CL_210255 ·

    Whisper ASR Models Adapted for Multilingual Medical Use

    Researchers have analyzed how multilingual medical adaptation affects the internal representations of Whisper models. The study compared various fine-tuning strategies across different Whisper model sizes, finding that …

  6. TOOL · CL_140740 ·

    Apple's SpeechAnalyzer API enhances on-device transcription speed

    Apple has introduced its SpeechAnalyzer API, designed for on-device English speech transcription. This new API utilizes the Apple Neural Engine to achieve significant improvements in speed and accuracy, reportedly being…

  7. TOOL · CL_104723 ·

    ASR systems evaluated for low-resource African language text corpora

    Researchers have evaluated the effectiveness of Automatic Speech Recognition (ASR) systems for creating text corpora for low-resource African languages, specifically Fongbe and Hausa. By fine-tuning the MMS-300M model o…

  8. TOOL · CL_74410 ·

    New attack targets ASR by manipulating speech features, not waveforms

    Researchers have developed a new adversarial attack method for automatic speech recognition (ASR) systems that operates in the feature space rather than directly on audio waveforms. This approach, termed the Clean-Refer…

  9. RESEARCH · CL_76814 ·

    New method boosts speech recognition for code-switching

    Researchers have developed a new contrastive training method to improve automatic speech recognition for code-switching, which involves alternating between languages within a single utterance. This approach identifies c…

  10. TOOL · CL_44843 ·

    Quantization study enables smaller, more accurate Whisper-small ASR

    A new study published on arXiv evaluates various post-training quantization (PTQ) techniques for the Whisper-small automatic speech recognition model. The research, which tested libraries like PyTorch, Optimum-Quanto, H…