PulseAugur
EN
LIVE 03:53:09
ENTITY CruxEval

CruxEval

PulseAugur coverage of CruxEval — every cluster mentioning CruxEval across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
0
4 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
4 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 4 TOTAL
  1. TOOL · CL_141553 ·

    LLMs struggle with semantic recall in long code contexts, new paper finds

    A new paper published on arXiv explores the limitations of large language models (LLMs) in understanding long code contexts. Researchers found that while LLMs excel at lexical recall (verbatim code retrieval), their sem…

  2. TOOL · CL_40817 ·

    Quantization impacts LLM performance, with larger models showing more resilience

    A new research paper explores the impact of quantization on large language model performance, examining models from 2-bit to 6-bit precision. The study found that while higher precision generally leads to better perform…

  3. TOOL · CL_29426 ·

    New framework StepCodeReasoner boosts code reasoning with execution traces

    Researchers have developed StepCodeReasoner, a new framework designed to improve code reasoning by focusing on intermediate execution states rather than just final outputs. This approach uses structured print statements…

  4. RESEARCH · CL_07050 ·

    Researchers generate verifiable code reasoning data to boost LLM performance

    Researchers have developed a new method to generate verifiable Chain-of-Thought (CoT) rationales for code reasoning by instrumenting code to capture execution traces. This pipeline narrates these traces into natural lan…