InternVL3.5-8B
PulseAugur coverage of InternVL3.5-8B — every cluster mentioning InternVL3.5-8B across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
VLMs' chart reading capabilities analyzed in new research
Researchers have investigated how vision-language models (VLMs) extract specific values from vertical bar charts, focusing on models like Qwen2.5VL-7B-Instruct and InternVL3.5-8B. Their analysis reveals that the top reg…
-
New Auditing Method Assesses Visual Token Provenance in MLLMs
A new research paper introduces a method for auditing the spatial provenance of visual tokens in multimodal large language models (MLLMs). This approach goes beyond traditional accuracy metrics to assess whether a model…
-
ET-Prune framework optimizes MLLM inference by dynamically pruning visual tokens
Researchers have developed ET-Prune, a novel framework designed to optimize the inference costs of multimodal large language models (MLLMs) by dynamically pruning visual tokens. Unlike fixed pruning ratios, ET-Prune ada…