PulseAugur
EN
LIVE 08:43:46
ENTITY Open Pre Trained Transformer

Open Pre Trained Transformer

PulseAugur coverage of Open Pre Trained Transformer — every cluster mentioning Open Pre Trained Transformer across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
0
10 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
9 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 10 TOTAL
  1. TOOL · CL_129256 ·

    New algorithms convert neural network heatmaps to TSP tours with provable guarantees

    Researchers have developed new algorithms to convert heatmaps, generated by neural networks, into tours for the Traveling Salesperson Problem (TSP). These algorithms provide theoretical guarantees that link the quality …

  2. TOOL · CL_106820 ·

    New SVD-Surgeon method optimizes LLM compression without retraining

    Researchers have developed SVD-Surgeon, a novel training-free method for compressing large language models (LLMs) using singular value decomposition (SVD). This technique optimizes the singular values directly, offering…

  3. TOOL · CL_94290 ·

    Shenzhen Big Data Institute's 4 AI research papers accepted by ICML 2026

    The Shenzhen Institute for Big Data Research has had four of its research papers accepted by ICML 2026, a top-tier international conference in machine learning. Two of the papers introduce novel optimization techniques …

  4. RESEARCH · CL_41829 ·

    Self-training restructures language models, research finds

    A new research paper challenges the common understanding of self-training in language models, suggesting it restructures rather than flattens language. The study found that while surface-level linguistic features like d…

  5. TOOL · CL_40799 ·

    New AR1-ZO method boosts LoRA fine-tuning with Zeroth-Order optimization

    Researchers have developed AR1-ZO, a novel method for fine-tuning large language models using Zeroth-Order optimization and Low-Rank Adaptation (LoRA). This technique addresses the challenge of effectively increasing Lo…

  6. TOOL · CL_38117 ·

    Opt adopts Anthropic's Claude Enterprise for ad operations

    Opt, a Japanese advertising company, has fully adopted Anthropic's Claude Enterprise Plan across its organization. This strategic move aims to revolutionize the operational structure of AI agent-based advertising. The c…

  7. RESEARCH · CL_15913 ·

    Researchers explore weight decay, in-context learning, and acceleration for Transformer models

    Researchers have developed several new methods to improve the efficiency and theoretical understanding of Transformer models. One paper provides a functional-analytic characterization of weight decay, demonstrating its …

  8. RESEARCH · CL_14113 ·

    Researchers explore efficient transformers via attention control and algorithmic capture

    Researchers are exploring methods to enhance transformer efficiency and understanding. One paper introduces Budgeted Attention Allocation, a head-gating mechanism that allows for cost-quality trade-offs. Another study d…

  9. RESEARCH · CL_09277 ·

    AI model evaluations are becoming a costly bottleneck, surpassing training expenses

    AI model evaluations are becoming prohibitively expensive, with recent benchmarks costing tens of thousands of dollars and consuming thousands of GPU hours. This high cost is particularly pronounced for agent-based eval…

  10. RESEARCH · CL_05407 ·

    AdaLeZO speeds up LLM fine-tuning with adaptive layer sampling

    Researchers have developed AdaLeZO, a new framework designed to make Zeroth-Order (ZO) optimization more efficient for fine-tuning Large Language Models. This method addresses the slow convergence and high variance typi…