PulseAugur
EN
LIVE 09:05:14
ENTITY LRS3

LRS3

PulseAugur coverage of LRS3 — every cluster mentioning LRS3 across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
4
6 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
4
6 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

3 day(s) with sentiment data

RECENT · PAGE 1/1 · 7 TOTAL
  1. TOOL · CL_256989 ·

    SyncVoice framework enhances video dubbing with vision-augmented TTS

    Researchers have developed SyncVoice, a novel framework for automatic video dubbing that enhances speech naturalness and temporal synchronization with visual content. By integrating a Text-Visual Fusion Module into a pr…

  2. TOOL · CL_245697 ·

    New Candor-LR dataset pushes audio-visual speech recognition toward natural conversation

    Researchers have introduced Candor-LR, a new dataset designed to advance audio-visual speech recognition (AVSR) by simulating natural conversations. Unlike existing benchmarks like LRS3, which use scripted speech, Cando…

  3. TOOL · CL_245695 ·

    New AVSRBench benchmark reveals generalization gap in speech recognition

    Researchers have developed AVSRBench, a new benchmark designed to evaluate Audio-Visual Speech Recognition (AVSR) systems across a variety of challenging conditions beyond standard broadcast speech. The study found that…

  4. TOOL · CL_223391 ·

    New method boosts LLM-based audio-visual speech recognition

    Researchers have developed a new method called Attention-Guided Reliability Scaling (AGRS) to improve audio-visual speech recognition (AVSR) systems that use large language models. This technique adapts contrastive deco…

  5. TOOL · CL_178315 ·

    New DoubleHelix framework improves audio-visual speech recognition

    Researchers have introduced DoubleHelix, a novel framework for audio-visual speech recognition (AVSR) that enhances the fusion of audio and visual data. Unlike previous methods that treat cross-modal interaction as a si…

  6. TOOL · CL_174035 ·

    New Framework Generates Speech from Facial Images

    Researchers have developed a novel Face-to-Speech (F2S) framework capable of generating plausible voices from static facial images, addressing the limitation of text-to-speech (TTS) systems that require reference audio.…

  7. TOOL · CL_66164 ·

    New VSR method uses head pose to improve accuracy

    Researchers have developed a new framework called HP-VSR-ResFiLM to improve visual speech recognition (VSR) by explicitly incorporating head-pose information. This method uses a pose-conditioned residual Feature-wise Li…