Calvin
PulseAugur coverage of Calvin — every cluster mentioning Calvin across labs, papers, and developer communities, ranked by signal.
3 day(s) with sentiment data
-
DFM-VLA introduces iterative action refinement for robot manipulation
Researchers have introduced DFM-VLA, a novel approach for robot manipulation that utilizes discrete flow matching to iteratively refine action tokens. Unlike previous methods that fix tokens once generated, DFM-VLA mode…
-
New framework uses exocentric data to improve egocentric 3D hand pose forecasting
Researchers have developed Exo2EgoPose, a novel framework designed to improve the forecasting of egocentric 3D hand poses. This method leverages exocentric demonstrations to guide and compensate for the limited and dyna…
-
New Anchor-Align method boosts VLA policy generalization
Researchers have introduced Anchor-Align, a novel method to improve vision-language-action (VLA) policies by addressing issues with standard behavior cloning (BC) finetuning. BC finetuning can degrade the generalizabili…
-
New methods enhance VLM to VLA adaptation for robotics control · 2 sources tracked
Two new research papers propose methods to improve the adaptation of vision-language models (VLMs) into vision-language-action (VLA) models for robotics. The first paper introduces CLAP (Causal Language-Action Predictio…
-
ThinkProprio integrates robot state to improve VLA model attention and speed
Researchers have developed a novel approach called ThinkProprio for vision-language-action (VLA) models, which integrates proprioceptive data more effectively into the decision-making process. Unlike traditional methods…
-
New benchmarks and frameworks advance video world modeling
Researchers have introduced "ImageTime," a new benchmark designed to evaluate how well image generation models can understand and represent temporal changes. This benchmark assesses spatiotemporal consistency by requiri…
-
New MoLA method bridges robot video imagination and action execution
Researchers have developed a new method called MoLA (Mixture of Latent Actions) to improve robot manipulation by better utilizing predicted future video frames. MoLA transforms these imagined futures into executable act…