PulseAugur
EN
LIVE 12:51:21
ENTITY SWE-rebench

SWE-rebench

PulseAugur coverage of SWE-rebench — every cluster mentioning SWE-rebench across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
1
4 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
1 over 90d
TIER MIX · 90D
TOPICS
TIMELINE
  1. 2026-07-01 research_milestone The SWE-rebench leaderboard was updated with new models and an improved UI for comparing AI performance on coding tasks. source
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 4 TOTAL
  1. TOOL · CL_169015 ·

    SWE-rebench adds multilingual coding tasks, GLM-5.2 leads leaderboard

    The SWE-rebench leaderboard has been updated with a new multilingual slice that evaluates software engineering tasks across five programming languages: Go, Java, Python, Rust, and TypeScript. The update includes perform…

  2. TOOL · CL_120598 ·

    SWE-rebench leaderboard adds Claude Opus 4.8, GLM-5.2, Gemini 3.5 Flash

    The SWE-rebench leaderboard has been updated with new models and improved UI, making it easier to compare AI performance on coding tasks. Notable additions include Claude Opus 4.8 xhigh, GLM-5.2, and Gemini 3.5 Flash, a…

  3. TOOL · CL_99352 ·

    2-bit GGUF models achieve 63% SWE-rebench pass rate with calibration

    A new method has been developed to calibrate 2-bit quantized language models, specifically GGUF formats under 10GB, for agentic coding tasks. These calibrated models, such as Qwopus3.6-27B-Coder, achieve over 60% pass r…

  4. TOOL · CL_55071 ·

    SWE-rebench leaderboard adds 110 new Python tasks for AI models

    The SWE-rebench leaderboard has been updated with 110 new Python tasks from GitHub PRs spanning March, April, and May. This update focuses on evaluating models' ability to read real issues, edit code, and pass test suit…