PulseAugur
EN
LIVE 00:50:04
ENTITY PutnamBench

PutnamBench

PulseAugur coverage of PutnamBench — every cluster mentioning PutnamBench across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
1
6 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
4 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 6 TOTAL
  1. SIGNIFICANT · CL_124629 ·

    Mistral AI releases Leanstral 1.5 for advanced formal verification

    Mistral AI has released Leanstral 1.5, an open-source model designed for formal verification tasks. This model, which has 6 billion active parameters and is available under the Apache 2.0 license, demonstrates significa…

  2. TOOL · CL_92404 ·

    Pythagoras-Prover achieves state-of-the-art in efficient formal proving

    Researchers have introduced Pythagoras-Prover, a new family of theorem provers designed for efficiency in formal reasoning tasks. These models utilize curriculum training and augmented formalization techniques to overco…

  3. TOOL · CL_74785 ·

    DeepSeek V4 powers Goedel-Architect to math competition win at low cost

    A new framework named Goedel-Architect, powered by DeepSeek V4, has achieved a 75.6% pass rate on the PutnamBench mathematics competition. This framework offers a significant cost advantage, costing only $294 compared t…

  4. TOOL · CL_65645 ·

    New framework ECP formally solves math answer-construction problems

    Researchers have developed a new neuro-symbolic framework called Enumerate-Conjecture-Prove (ECP) designed to tackle answer-construction problems in formal mathematics. This framework combines general large language mod…

  5. RESEARCH · CL_62835 ·

    AI frameworks boost formal theorem proving with new techniques

    Researchers have developed new frameworks to enhance formal theorem proving capabilities using large language models. Goedel-Architect utilizes a blueprint generation and refinement strategy, achieving state-of-the-art …

  6. RESEARCH · CL_62715 ·

    LLMs optimized for efficient formal theorem proving in Lean

    Two new research papers explore methods to improve the efficiency and effectiveness of large language models (LLMs) in formal theorem proving within the Lean environment. The first paper introduces an action routing age…