Researchers have introduced a new dataset and harness called Wikidata Search Traces, designed to train agents for querying large knowledge bases like Wikidata. The dataset addresses limitations in current language models, which often rely on internal memory rather than active graph exploration. The proposed method involves creating multi-hop questions and managing retrieved evidence effectively, showing improved performance for both commercial and open-weight models. AI
IMPACT This research could lead to more capable AI agents for complex information retrieval from large knowledge bases.
RANK_REASON The cluster contains a research paper detailing a new dataset and methodology for training AI agents on knowledge graph search. [lever_c_demoted from research: ic=1 ai=1.0]
- arXiv
- Carlos Rosas-Hinostroza
- GPT-6 Luna
- Hugging Face
- Language Models
- Python
- Qwen3.8-27B
- recursive language model (RLM)
- SPARQL
- Wikidata
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →