MIDI
PulseAugur coverage of MIDI — every cluster mentioning MIDI across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
New LLM techniques advance text-to-music generation
Researchers have developed new methods for text-to-music generation using large language models (LLMs). The first approach, "Agogic," focuses on performance-timed music tokens and demonstrates that the choice of music r…
-
New self-supervised model enhances AI's understanding of symbolic music
Researchers have developed a hierarchical self-supervised world model for symbolic music, utilizing a 2.55M-parameter Swin V2 encoder trained on MIDI data. This model, which does not require labels or music theory vocab…
-
New AI models evaluate expressive MIDI piano performances, releasing open-source library
Researchers have developed a new method for evaluating expressive MIDI piano performances by integrating contextual embeddings from self-supervised symbolic music models like Aria and CLaMP3. This approach addresses lim…
-
Kyutai releases open-weight MuScriptor for multi-instrument music transcription
Kyutai, a French AI lab, has introduced MuScriptor, an open-weight transformer model designed for transcribing multi-instrument music into MIDI format. The model is available in three sizes, with the largest having 1.4 …
-
PipeWire 1.6.8 released; Podcast covers Maine politics, Meta's Threads, and Trump phone
This cluster covers two distinct technology news items: the release of PipeWire 1.6.8 with enhanced audio and video server capabilities for Linux, and a podcast episode discussing a political scandal in Maine, Meta's Th…
-
MuScriptor: Open-weight model transcribes multi-instrument music to MIDI
Researchers have introduced MuScriptor, an open-weight model designed for multi-instrument music transcription. Unlike previous models that struggled with complex mixes or single instruments, MuScriptor can transcribe n…
-
New framework decompiles symbolic music to executable code
Researchers have developed Decomposer, a novel framework designed to translate symbolic music, such as MIDI files, into executable programs within the Strudel music programming language. This system addresses the scarci…
-
Pianist Transformer advances expressive music generation with self-supervised learning
Researchers have developed Pianist Transformer, a novel approach to generating expressive piano performances from symbolic music scores. This method utilizes large-scale self-supervised learning on over 10 billion token…
-
AI researcher explores using LLMs to generate music from MIDI data
An individual is exploring the idea of fine-tuning a small large language model (LLM) to understand and generate music patterns. The approach involves using extended MIDI file protocols instead of actual sound samples, …
-
New algorithm aids pitch spelling and key estimation in musical scores
Researchers have developed an algorithm for pitch spelling and key estimation in musical scores. The algorithm processes MIDI-like input to determine note names, key signatures, and local scales for each bar, optimizing…
-
Beethoven's Moonlight Sonata mirrors ML architectures, study finds
A new research paper explores the structural parallels between Beethoven's "Moonlight Sonata" (Op. 27 No. 2) and machine learning mechanisms. Through computational analysis of the musical score, the study identifies dis…
-
AI Models Compared for Bach-Style Music Generation
A new research paper compares different AI models for generating Bach-style piano music. The study found that autoregressive LSTMs with attention produced the most musically coherent results, while vector quantization i…
-
IntelliJ session to explore JavaFX and MIDI file manipulation
A session focused on JavaFX and MIDI file manipulation within IntelliJ was announced, featuring participants Anton Arhipov and another individual. The event encourages audience interaction with questions and offers a li…
-
Melinda French Gates commits $215M to menopause and midlife women's health
Melinda French Gates has committed $215 million to advance research and care for menopause and midlife women's health. This significant philanthropic investment aims to address the underfunding and lack of attention in …
-
New models unify speech and singing voice generation
Researchers have developed new unified models for generating human vocal audio, capable of producing both speech and singing. UniVoice uses a conditional flow matching approach, separating content, melody, and timbre to…
-
Google releases Magenta RealTime 2 for live AI music collaboration
Google has released Magenta RealTime 2, an AI model capable of improvising music in real-time alongside human musicians. The model is available in two versions: a high-quality 2.4 billion parameter model and a faster 23…
-
New dataset tests AI's grasp of multilingual idioms
Researchers have introduced MIDI, a new dataset designed to evaluate how well multilingual NLP models understand idiomatic expressions. This dataset includes idioms in sentence and conversational contexts across high-, …
-
Data centers prioritize on-site power amid grid interconnection delays
Data center developers are increasingly turning to on-site power generation, primarily natural gas-based microgrids, as a solution to significant interconnection delays with the main power grid. This shift is driven by …
-
PianoCoRe dataset combines and refines MIDI corpora for expressive performance research
Researchers have introduced PianoCoRe, a comprehensive and refined dataset of piano MIDI performances and scores. This new dataset combines and improves upon existing open-source corpora, offering over 250,000 performan…
-
EchoSight offers open-source visual-audio sensory substitution app
EchoSight is an open-source mobile application designed to provide real-time visual-to-audio sensory substitution. This framework aims to assist individuals, particularly those with blindness, by converting visual infor…