Unbounded Labs has developed Bart, a 2.82 billion parameter large language model trained on 20.1 billion tokens of English text predating 1931. The project, which cost approximately $800 and took three months, aimed to explore whether LLMs could replicate the scientific reasoning of historical figures. Bart achieved top performance on the newly created Vintage CORE benchmark for its scale and has open-sourced its datasets, methodology, and training code. AI
IMPACT This release explores the potential for LLMs to engage with historical knowledge and reasoning, potentially opening new avenues for AI research.
RANK_REASON The cluster describes the release of a new LLM with novel training data and benchmarks, fitting the research category.
AI-generated summary · Google Gemini · from 2 sources. How we write summaries →