Visual Tokens
PulseAugur coverage of Visual Tokens — every cluster mentioning Visual Tokens across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
New CoVeR method prunes visual tokens for 3D reasoning in VLMs
Researchers have developed CoVeR, a novel method for pruning visual tokens in Vision-Language Models (VLMs) when processing 3D scenes represented by multi-view images. This technique addresses the issue of redundant tok…
-
New research optimizes visual token processing for long-video MLLMs
Researchers are exploring methods to optimize how multimodal large language models (MLLMs) process visual information, particularly for long videos. Several papers introduce techniques for selecting, compressing, and pr…
-
New MM-ShiftKV method optimizes KV caching for multimodal LLMs
Researchers have developed MM-ShiftKV, a novel method for optimizing Key-Value (KV) caching in multimodal large language models (MLLMs). This technique addresses the issue where prefill-stage KV selection methods, which…
-
New AI framework enhances task-oriented communication for 6G networks
Researchers have introduced a new framework called ET-TokenCom designed to improve task-oriented communication in AI-native 6G networks. This framework addresses challenges in representing task-oriented tokens, integrat…