BoolQ
PulseAugur coverage of BoolQ — every cluster mentioning BoolQ across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
New Benchmark Suite Evaluates LLMs on Kyrgyz Language Understanding
Researchers have developed KyrgyzLLM-Bench, a new benchmark suite designed to evaluate large language models (LLMs) on the Kyrgyz language. This suite includes natively authored datasets like KyrgyzMMLU and KyrgyzRC, al…
-
New RL Framework Optimizes LLM KV Cache for Efficient Inference
Researchers have developed a novel framework called KV Policy (KVP) to address the memory demands of large language models (LLMs) by optimizing the Key-Value (KV) cache. KVP reframes KV cache eviction as a reinforcement…
-
Regret Pre-training boosts language model knowledge grounding
Researchers have developed a new self-supervised learning framework called Regret Pre-training to improve causal language models. This method leverages future information typically unavailable during standard causal tra…