PulseAugur
EN
LIVE 14:52:06
ENTITY Lean

Lean

PulseAugur coverage of Lean — every cluster mentioning Lean across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
30
80 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
23
66 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

19 day(s) with sentiment data

RECENT · PAGE 1/4 · 80 TOTAL
  1. TOOL · CL_196011 ·

    New benchmark reveals sycophancy in AI math formalization systems

    Researchers have introduced FaithformBench, a new benchmark designed to evaluate the faithfulness of autoformalisation (AF) systems. These systems translate natural language reasoning into formal statements for proof as…

  2. TOOL · CL_195691 ·

    AI-assisted proof confirms 14 queens needed for 26x26 board domination

    A mathematical proof has resolved the 26x26 queen domination number, establishing that 14 queens are both sufficient and necessary to dominate the board. This proof, verified by Lean 4.32.2 and an independent kernel, wa…

  3. COMMENTARY · CL_195444 ·

    AI's potential role in formal code verification discussed

    The discussion explores the potential for AI to advance formal verification in coding, suggesting that increased AI focus on this area could lead to more trustworthy code. The idea is that AI might enable simulations of…

  4. TOOL · CL_195032 ·

    Anthropic AI autonomously advances Riemann hypothesis research

    An unreleased Anthropic AI model, operating autonomously, made significant progress on a sub-problem of the 150-year-old Riemann hypothesis. Over 36 hours, the AI model, named Claude Code, coordinated 60 sub-agents and …

  5. RESEARCH · CL_194985 ·

    Anthropic model advances on Riemann hypothesis, sparking debate

    An unreleased Anthropic model has reportedly made significant progress on the Riemann hypothesis, a long-standing unsolved problem in mathematics concerning prime numbers. The model, prompted by a staff member with limi…

  6. TOOL · CL_194781 ·

    Anthropic's ClaudeAI advances Riemann hypothesis research with new bound

    Anthropic's ClaudeAI has advanced research in mathematics by raising the lower bound for zeros of the Riemann hypothesis from 41.6% to 67.25%. This significant progress, formalized using the Lean proof assistant, does n…

  7. RESEARCH · CL_193385 ·

    New LLM methods boost verified code generation with integrated planning and proof search

    Researchers have developed new methods for verified code generation, where large language models (LLMs) produce both executable programs and machine-checkable proofs of correctness. The first approach, P$^{3}$, integrat…

  8. TOOL · CL_193055 ·

    New voting method synthesis theorem for infinite domains

    Researchers have developed a novel approach using SMT and Lean to synthesize and verify voting methods on infinite domains, a significant advancement over traditional finite-domain SAT solvers. This method addresses the…

  9. TOOL · CL_187381 ·

    Weaver framework combines weak verifiers to boost LLM accuracy

    Researchers have developed Weaver, a framework designed to improve language model verification by combining multiple imperfect verifiers into a stronger, more accurate system. This approach aims to reduce the performanc…

  10. SIGNIFICANT · CL_182080 ·

    OpenAI's GPT-5.6 Sol cuts costs by 20%, Astra makes math breakthroughs

    OpenAI has achieved significant advancements with its frontier models, including GPT-5.6 Sol which autonomously optimized production GPU kernels, reducing serving costs by 20%. Concurrently, an internal model named Astr…

  11. COMMENTARY · CL_181596 ·

    Mathematician refutes OpenAI's AI-generated proof of Connes rigidity conjecture

    A mathematician has refuted OpenAI's claim that its new AI model proved the Connes rigidity conjecture. The mathematician, J. L. Nielsen from the University of Kansas, found that the AI's proof, consisting of 37,000 lin…

  12. FRONTIER RELEASE · CL_179821 ·

    OpenAI's GPT 5.6 "Sol" model generates 10 new mathematical proofs

    OpenAI has announced that an internal version of its upcoming model, GPT 5.6 "Sol", has generated 10 novel results in mathematics and theoretical computer science. These breakthroughs address long-standing open problems…

  13. FRONTIER RELEASE · CL_175989 ·

    OpenAI's Astra model solves 10 major math problems, sparking debate

    OpenAI has revealed details about its unreleased model, Astra, which has successfully solved ten major open problems in mathematics and theoretical computer science. These breakthroughs, achieved at a token cost of appr…

  14. COMMENTARY · CL_169553 ·

    Human role in AI era debated: multiplier, consumer, or collaborator?

    The role of humans in an era of Artificial Superintelligence (ASI) is being questioned, drawing parallels to chess where AI has surpassed human capabilities. While AI in chess has become a tool for consumption rather th…

  15. TOOL · CL_169213 ·

    Rocq prover surpasses Lean for program verification

    The Rocq prover offers advantages over Lean for program verification, particularly in its handling of complex proofs and its integration with machine learning techniques. While Lean is a powerful tool for formal methods…

  16. COMMENTARY · CL_177231 ·

    Rocq prover favored over Lean for program verification

    The author argues that Rocq is a more suitable tool than Lean for program verification, particularly due to Rocq's direct support for executable coinductive types and cofixpoints. While Lean has introduced coinductive p…

  17. TOOL · CL_167374 ·

    Formalizing Flag Algebras in Lean for Graph Theory Proofs

    Researchers have developed a machine-checked formalization of Razborov's flag algebra method within the Lean proof assistant. This formalization enables a compiler that transforms semidefinite programming output into ve…

  18. TOOL · CL_164983 ·

    New framework automates causal inference research using formal proofs

    Researchers have developed CausalForge, a new framework designed to automate theoretical research in causal inference. This system integrates Causalean, a Lean proof assistant library with thousands of machine-checked d…

  19. COMMENTARY · CL_163166 ·

    Language as a Latent Space: Human vs. AI Reasoning

    This post explores the concept of language as a high-entropy latent space, drawing parallels between human language and strongly-typed programming languages like Rust and Haskell. The author argues that language, much l…

  20. MEME · CL_162211 ·

    AI-generated Riemann hypothesis proof on arXiv poses security risk

    A user on Mastodon questioned whether individuals would download and verify an AI-generated proof of the Riemann hypothesis posted on arXiv, provided it included a complete Lean formalization. The response cautioned aga…