Kanishk Awadhiya
PulseAugur coverage of Kanishk Awadhiya — every cluster mentioning Kanishk Awadhiya across labs, papers, and developer communities, ranked by signal.
-
New Bifocal Attention method aims to improve LLM algorithmic generalization
A research paper introduced Bifocal Attention, a new architectural paradigm designed to improve algorithmic generalization in large language models. This approach combines standard Rotary Positional Embeddings (RoPE) fo…
-
Withdrawn paper links Vision Transformer sparsity to data complexity
A recently withdrawn arXiv paper explored the phenomenon of "representational sparsity" in Vision Transformers (ViTs). The research, led by Kanishk Awadhiya, proposed that the observed "U-shaped" entropy profile in ViTs…
-
New H-Res method efficiently adapts Transformer models without altering weights
Researchers have introduced H-Res (Hierarchical Residual Steering), a novel method for adapting large Transformer models, which function as Dense Associative Memories (DAMs). This technique addresses the "Plasticity-Sta…