PulseAugur
EN
LIVE 15:22:53
ENTITY CVPR 2026

CVPR 2026

PulseAugur coverage of CVPR 2026 — every cluster mentioning CVPR 2026 across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
0
10 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
8 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
TIMELINE
  1. 2026-06-08 research_milestone The CVPR 2026 conference concluded, featuring awards for D4RT and Oxford VGG, the release of the PhysInOne dataset, and notable contributions from Chinese researchers. source
RECENT · PAGE 1/3 · 58 TOTAL
  1. TOOL · CL_218291 ·

    1st Asynchronous CASTLE Challenge Results Published for CVPR 2026

    The 1st Asynchronous CASTLE Challenge, held in conjunction with the Egocentric Vision Workshop at CVPR 2026, has concluded. This report details the contributions and outcomes from the challenge participants. Several pla…

  2. SIGNIFICANT · CL_215144 ·

    Humanoid robot plays full tennis match, showcasing integrated AI decision-making and motor control

    Galaxy General has demonstrated a significant advancement in embodied AI with their humanoid robot, AstraTennis, which successfully played a full game of tennis against a professional player. This achievement, dubbed th…

  3. TOOL · CL_204212 ·

    3D Gaussian Splatting advances towards embodied AI with efficiency and adaptability

    Researchers are advancing 3D Gaussian Splatting (3DGS) beyond high-quality rendering to practical applications in embodied AI and robotics. Recent work presented at CVPR 2026 focuses on making 3DGS more efficient and ad…

  4. TOOL · CL_194146 ·

    Multimodal AI boosts animal identification accuracy by 11%

    Researchers have developed a multimodal framework to improve animal identification by combining visual data with semantic information from text descriptions. This approach was tested on a large dataset of nearly 700,000…

  5. RESEARCH · CL_198050 ·

    New Promptable Gaze Estimation Model Integrates Subject Localization

    Researchers have introduced Promptable Gaze Target Estimation (PGE), a novel end-to-end approach for analyzing human gaze in images. Unlike previous methods that rely on multi-stage pipelines and explicit inputs, PGE us…

  6. TOOL · CL_185284 ·

    EgoCross Challenge debuts at CVPR 2026 for egocentric video QA

    The first EgoCross Challenge, held at EgoVis 2026 during CVPR 2026, introduced a benchmark for evaluating multimodal large language models on cross-domain egocentric video question answering. The challenge focused on mo…

  7. TOOL · CL_148244 ·

    Robots' spatial misrepresentation, not physics, hinders generalization: new research

    Researchers from Sun Yat-sen University and X-Era AI Lab have identified that the primary cause of failure in robots operating in new environments is not a lack of physical understanding, but rather misaligned spatial r…

  8. TOOL · CL_138440 ·

    New GCA method uses formal task constraints for VLM spatial reasoning

    A new paper proposes the Geometrically-Constrained Agent for Spatial Reasoning (GCA) to improve how vision-language models (VLMs) handle spatial queries. GCA introduces a two-stage process: first, the VLM formalizes the…

  9. RESEARCH · CL_135124 ·

    New AUTOPILOT VQA benchmark tests AI for dashcam incident understanding

    Researchers have introduced AUTOPILOT-VQA, a new benchmark designed to evaluate the capabilities of vision-language models in understanding safety-critical incidents from dashcam footage. This benchmark utilizes structu…

  10. TOOL · CL_124937 ·

    MACT framework uses specialized agents for better visual document understanding

    Researchers have introduced MACT, a novel multi-agent framework designed to improve visual document understanding. Unlike traditional large vision-language models that attempt a single forward pass, MACT divides the com…

  11. TOOL · CL_103645 ·

    Humanoid robot 'cerebellum' gets GPT-style model with 2B frames of motion data

    Researchers have introduced AstraBrain-WBC 0.5, a novel GPT-style foundational model designed for humanoid robot general cerebellum control. This model leverages a massive dataset of 2 billion frames of human motion dat…

  12. TOOL · CL_103176 ·

    ByteDance Seed's SpatialTree framework accepted for CVPR 2026

    ByteDance Seed, in collaboration with academic partners, has introduced SpatialTree, a novel hierarchical framework designed to enhance the spatial intelligence of multimodal large language models (MLLMs). This new fram…

  13. TOOL · CL_93874 ·

    New method boosts video QA accuracy using cross-model disagreement

    Researchers have developed a novel inference-time procedure called disagreement-based cross-model routing to improve video question answering accuracy. This method leverages the variance in outputs from a primary video …

  14. TOOL · CL_87282 ·

    PS-SR: New AI Method Enhances Video Resolution with Speed and Detail

    Researchers from the University of Science and Technology of China and Zhixiang Future have developed PS-SR, a novel video super-resolution technique that balances speed and detail. The method employs a powerful base mo…

  15. TOOL · CL_87284 ·

    AI Models Shift Focus to Stability and Adaptability in Real-World Deployments

    Recent research presented at CVPR 2026 highlights a shift in AI model development from pure capability expansion to "capability management." This involves ensuring models retain old knowledge while adapting to new data …

  16. RESEARCH · CL_84428 ·

    AutoMine method uses LLMs and VLMs for autonomous driving scenario mining

    Researchers have developed AutoMine, a novel method for extracting critical scenarios from autonomous driving data using Large Language Models (LLMs) and Vision-Language Models (VLMs). This approach enhances prompt sens…

  17. COMMENTARY · CL_82327 ·

    AI agents move from theory to messy reality, demanding governance

    The AI agent revolution is rapidly moving from theoretical concept to operational reality for enterprises, with companies like Workday developing governance tools and security standards. While multimodal AI capabilities…

  18. RESEARCH · CL_82192 ·

    PortraitCraft Challenge advances AI portrait generation and understanding

    Researchers have introduced the PortraitCraft Challenge, a new competition focused on AI's ability to understand and generate portraits. This challenge, held at CVPR 2026, includes two tracks: one for analyzing portrait…

  19. TOOL · CL_80473 ·

    Latent Labs founder: AI enables programmable biology and autonomous drug design

    Latent Labs founder Simon Kohl, a key figure in the AlphaFold project, presented at CVPR 2026 on using generative AI for molecular design. He highlighted the inefficiencies in traditional drug discovery, which takes ove…

  20. RESEARCH · CL_79703 ·

    Claude Code agent aids scenario mining for autonomous driving challenge

    Researchers have developed a novel four-stage pipeline for the CVPR 2026 Argoverse 2 Scenario Mining Challenge. This system leverages a Claude Code agent, powered by GLM 5.1, for autonomous code generation. It then refi…