Researchers have developed SCAPES, a new generative model for environmental sounds that is lightweight and resource-efficient. This model synthesizes high-fidelity environmental textures using high-level semantic control by operating on the continuous latent manifold of a neural audio codec. SCAPES utilizes a Continuous Normalizing Flow with Flow Matching to model latent trajectories, allowing a 36-million parameter instance to be trained on limited datasets using a single consumer-grade GPU. AI
IMPACT This model offers a more accessible and flexible tool for creative sound design and open research by reducing computational and ecological costs.
RANK_REASON The item is a research paper detailing a new generative model for environmental sounds. [lever_c_demoted from research: ic=1 ai=1.0]
- alphaXiv
- arXiv
- CatalyzeX
- Continuous Normalizing Flow
- DagsHub
- Flow Matching for Generative Modeling
- Gotit.pub
- Hugging Face
- ScienceCast
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →