Researchers have introduced ConlangBench, a novel benchmark designed to evaluate and train large language models (LLMs) on constructed languages (conlangs). This benchmark comprises over 21 million conlang-English parallel sentence pairs and 321,000 vocabulary entries across 21 conlangs. Experiments indicate that LLMs perform better on a posteriori conlangs, which derive their vocabulary from natural languages, and that models can successfully learn conlangs when sufficient parallel corpora are available, though learning curves vary based on the language's creation method. ConlangBench offers a unique platform for studying LLM acquisition of low-resource languages. AI
IMPACT Provides a new evaluation framework for LLM understanding of linguistic diversity and low-resource language acquisition.
RANK_REASON The item describes a new academic paper introducing a benchmark for evaluating LLMs on constructed languages. [lever_c_demoted from research: ic=1 ai=1.0]
- alphaXiv
- arXiv
- CatalyzeX
- ConlangBench
- DagsHub
- English
- Esperanto
- Gotit.pub
- Hugging Face
- LLMs
- ScienceCast
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →