InternVL3 8B
PulseAugur coverage of InternVL3 8B — every cluster mentioning InternVL3 8B across labs, papers, and developer communities, ranked by signal.
3 day(s) with sentiment data
-
New SLAPBench benchmark tests MLLMs on fingerprint verification
Researchers have introduced SLAPBench, the first benchmark designed to evaluate multimodal large language models (MLLMs) on four-finger SLAP fingerprint verification. The benchmark, built using NIST SD302b data, tests M…
-
New SD-MAR framework boosts VLM analytical reasoning across multiple images
Researchers have introduced SD-MAR, a new framework designed to enhance the analytical reasoning capabilities of vision-language models (VLMs) across multiple images. This framework utilizes synthetic data generated thr…
-
VLMs' reasoning chains offer better uncertainty signals than answer entropy, study finds
A new research paper explores the effectiveness of "thinking chains" in visual language models (VLMs) for quantifying uncertainty. The study found that while some models like Qwen3-VL-8B-Thinking exhibit a complete coll…
-
New VLM-Judge Protocol Evaluates 3D Mesh Quality Reliably
Researchers have developed a de-biased protocol using vision-language models (VLMs) to evaluate the quality of 3D meshes generated from single images. This protocol, which involves using distinct VLM judges for training…
-
HiDe framework boosts MLLM performance on high-res images
Researchers have developed a new training-free framework called HiDe to improve the performance of Multimodal Large Language Models (MLLMs) on high-resolution images. HiDe addresses background interference rather than o…
-
LiteFrame boosts Video LLM frame scaling and cuts latency
Researchers have developed LiteFrame, an efficient vision encoder designed to improve the performance of Video Large Language Models (Video LLMs) when processing extended video content. This new framework uses Compresse…