Luxembourgish
PulseAugur coverage of Luxembourgish — every cluster mentioning Luxembourgish across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
New datasets and studies explore multilingual instruction tuning for LLMs
Researchers have developed new methods for instruction tuning large language models in low-resource languages. One study introduces LuxInstruct, a cross-lingual dataset for Luxembourgish that avoids machine translation …
-
New dataset boosts Luxembourgish LLM performance
Researchers have developed LuxIT, a new dataset designed to improve the capabilities of Large Language Models (LLMs) for the Luxembourgish language. By synthesizing data from native texts and using DeepSeek-R1-0528 for …
-
New Luxembourgish SQA system uses TTS, new expressive speech corpus released
Researchers have developed LuxSQA, a system for spoken question answering in Luxembourgish, a low-resource language. The system utilizes text-to-speech (TTS) technology to generate synthetic spoken questions, augmenting…
-
New expressive speech corpus released for Luxembourgish language
Researchers have introduced LuxEmo, a new 21-hour corpus of expressive speech for the Luxembourgish language, addressing the underrepresentation of low-resource languages in speech technology. The corpus, derived from R…
-
LLMs improve Luxembourgish borrowing detection with knowledge graph prompts
Researchers have developed a new benchmark, LexNeo-Bench, to evaluate how well large language models understand lexical borrowing in low-resource languages like Luxembourgish. The benchmark, derived from a Luxembourgish…
-
Low-resource NLP needs both cross-lingual transfer and specific data
A new paper argues that low-resource natural language processing (NLP) requires a combination of cross-lingual transfer and language-specific development. While cross-lingual transfer can boost performance using data fr…
-
LLMs analyze language ideologies in Luxembourgish news comments
Researchers have developed a new method using sparse crosscoders to track the emergence and consolidation of linguistic features within large language models during pretraining. This technique, which includes a novel me…