PulseAugur
EN
LIVE 16:05:18
ENTITY ARC-AGI-1

ARC-AGI-1

PulseAugur coverage of ARC-AGI-1 — every cluster mentioning ARC-AGI-1 across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
3
8 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
5 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

3 day(s) with sentiment data

RECENT · PAGE 1/1 · 9 TOTAL
  1. TOOL · CL_247782 ·

    New 'Looped Flows' method boosts AI reasoning on complex tasks

    Researchers have introduced "looped flows," a novel approach to training recurrent neural networks that enhances their ability to solve complex problems by allowing for more computational updates during inference. This …

  2. TOOL · CL_235965 ·

    Astra language model achieves high scores on ARC-AGI benchmark without CoT

    The Astra language model has achieved impressive scores on the ARC-AGI benchmark, reaching 97% on ARC-AGI-3 and 86% on ARC-AGI-1. Notably, these high scores were attained without the use of Chain-of-Thought (CoT) prompt…

  3. RESEARCH · CL_230007 ·

    Small Transformer Achieves 44% on ARC-AGI-1 Benchmark for 67 Cents

    A researcher has developed a small transformer model that achieves 44% accuracy on the ARC-AGI-1 benchmark, a significant feat accomplished in just 1.5 hours and costing only 67 cents. This model surpasses many existing…

  4. RESEARCH · CL_201277 ·

    150M param recurrent model achieves strong ARC-AGI-1 score at low cost · 2 sources tracked

    A new recurrent latent reasoning model, significantly smaller than typical transformer models, has achieved a notable score of 29.5% on the ARC-AGI-1 benchmark. This model operates at a very low cost of $0.0007 per task…

  5. TOOL · CL_194794 ·

    Pathway's BDH-CQ model sets new ARC-AGI-1 cost-efficiency frontier

    Pathway has announced BDH-CQ, a 150 million parameter post-Transformer model that achieves a score of 29.5% on the ARC-AGI-1 benchmark. This model reportedly sets a new cost-efficiency frontier, with a computed cost of …

  6. SIGNIFICANT · CL_191626 ·

    DeepSeek V4 Flash 0731 achieves high scores on ARC-AGI benchmark

    DeepSeek has released its V4 Flash 0731 model, featuring three reasoning settings: Low, High, and Max. The model achieved impressive scores on the ARC Prize benchmark, a test for abstract reasoning in AI systems. Notabl…

  7. RESEARCH · CL_193219 ·

    BDH-CQ model achieves new state-of-the-art in AI reasoning cost efficiency

    Researchers have developed BDH-CQ, a novel reasoning model that integrates in-context learning with recurrent latent reasoning. This model updates its memory with inference-time inputs and iteratively computes solutions…

  8. TOOL · CL_178347 ·

    TraceViT model enhances AI visual abstract reasoning with step-by-step supervision

    Researchers have introduced TraceViT, a novel looped visual reasoner designed to improve performance on abstract reasoning tasks. Unlike previous methods that only constrain the final output, TraceViT is trained to foll…

  9. RESEARCH · CL_107855 ·

    AI benchmark scores predictable from just two factors, study finds

    A new research paper proposes a method called BenchPress that can predict a frontier model's performance across numerous benchmarks using only two key scores. The study analyzed 84 models and 133 benchmarks, finding tha…