PulseAugur
EN
LIVE 14:22:01
ENTITY Full Attention

Full Attention

PulseAugur coverage of Full Attention — every cluster mentioning Full Attention across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
2
6 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
2
5 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

2 day(s) with sentiment data

RECENT · PAGE 1/1 · 6 TOTAL
  1. TOOL · CL_191193 ·

    New Autonomy-of-Heads method boosts LLM efficiency without data

    Researchers have developed a novel data-free method called Autonomy-of-Heads (AoH) to improve the efficiency of long-context Large Language Models. AoH identifies retrieval and streaming heads by analyzing the spectral …

  2. TOOL · CL_180556 ·

    Bole system accelerates hybrid-attention LLM inference with tree speculation

    Researchers have developed Bole, a new system designed to accelerate inference for hybrid-attention large language models. These models combine full attention with recurrent linear attention to manage long contexts more…

  3. TOOL · CL_137807 ·

    Xiaomi details MiMo-V2.5 AI model efficiency optimizations

    Xiaomi has detailed the engineering optimizations behind its MiMo-V2.5 series of AI models, focusing on achieving efficiency for long-context reasoning and multimodal tasks. The models employ Hybrid Sliding Window Atten…

  4. RESEARCH · CL_115713 ·

    New attention mechanisms boost LLM efficiency and reduce hallucination · 10 sources tracked

    Researchers are developing novel attention mechanisms to improve the efficiency and capabilities of large language models (LLMs) and multimodal large language models (MLLMs). These advancements focus on optimizing spars…

  5. RESEARCH · CL_103889 ·

    HydraHead architecture fuses attention types for improved long-context LLMs

    Researchers have introduced HydraHead, a novel architecture that hybridizes Full Attention and Linear Attention at the head level within transformer models. This approach leverages interpretability to identify critical …

  6. RESEARCH · CL_93108 ·

    New research explores hybrid and sparse attention mechanisms for LLMs

    Researchers are exploring novel methods to optimize attention mechanisms in large language models, particularly for handling long contexts. The HydraHead architecture, for instance, hybridizes Full Attention (FA) and Li…