PulseAugur
EN
LIVE 18:46:21
ENTITY DeepSeek-Prover-V2-7B

DeepSeek-Prover-V2-7B

PulseAugur coverage of DeepSeek-Prover-V2-7B — every cluster mentioning DeepSeek-Prover-V2-7B across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
0
3 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
3 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 3 TOTAL
  1. RESEARCH · CL_117485 ·

    Flaws Found in Lean Theorem Proving Benchmarks and RL Model Inference

    Researchers have identified significant flaws in the formal benchmarking of Lean theorem-proving datasets, uncovering thousands of issues including counterexamples and vacuous theorems. A separate study on RL-trained Le…

  2. RESEARCH · CL_91340 ·

    New LLM Frameworks and Benchmarks Advance Formal Mathematical Reasoning

    Researchers are developing new methods and benchmarks to improve the formal mathematical reasoning capabilities of large language models (LLMs). One approach, Diffusion-Proof, utilizes diffusion LLMs (dLLMs) for theorem…

  3. TOOL · CL_27514 ·

    FormalRewardBench benchmark evaluates LLM reward models for theorem proving

    Researchers have introduced FormalRewardBench, a new benchmark designed to evaluate reward models used in formal theorem proving. This benchmark addresses the challenge of sparse credit assignment in reinforcement learn…