PulseAugur
EN
LIVE 11:47:00

New Arti-JEPA model adapts video models for vocal tract MRI analysis

Researchers have developed Arti-JEPA, a new joint embedding predictive architecture designed to model real-time MRI data of the vocal tract for speech analysis. This model was trained on approximately 62 hours of unlabeled vocal tract videos and evaluated on tasks including phoneme prediction, classification of fluent versus disfluent speech, and characterizing speech changes after surgery. The findings indicate that a temporal video prior is more effective than per-frame encoders, and domain adaptation is crucial for certain tasks like phoneme prediction, though it did not improve stuttering classification. Arti-JEPA demonstrated an ability to decode phoneme signals from post-operative speech, suggesting its potential as a reusable measurement tool for speech science. AI

IMPACT This research offers a new method for analyzing vocal tract dynamics using MRI, potentially advancing speech science and clinical applications.

RANK_REASON The cluster contains an academic paper detailing a new model and its evaluation. [lever_c_demoted from research: ic=1 ai=1.0]

Read on arXiv cs.CV →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

New Arti-JEPA model adapts video models for vocal tract MRI analysis

How we ranked this

Signal score
9 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
The cluster contains an academic paper detailing a new model and its evaluation. [lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
paper, model release
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Same-day
Cluster formed today. Ranking reflects the current source set at time of score.

Full methodology in our editorial standards.

COVERAGE [1]

  1. arXiv cs.CV TIER_1 English(EN) · Hong Nguyen, Sean Foley, Christina Hagedorn, Yijing Lu, Sudarsana Reddy Kadiri, Dani Byrd, Shrikanth Narayanan ·

    Arti-JEPA: Adapting Video World Model to Real-Time MRI of the Vocal Tract for Speech-Production Analysis

    arXiv:2609.09757v1 Announce Type: cross Abstract: Real-time MRI (rtMRI) captures the dynamics of the entire vocal tract during speech, but labeled data are scarce and the modality - single-slice, grayscale, low-resolution - differs substantially from the natural videos that video…