ENTITY
HalfCheetah-v4
HalfCheetah-v4
PulseAugur coverage of HalfCheetah-v4 — every cluster mentioning HalfCheetah-v4 across labs, papers, and developer communities, ranked by signal.
Total · 30d
0
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
2 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 2 TOTAL
-
New research explores advanced RL for agent survival, navigation, and explainability · 7 sources tracked
Researchers are exploring advanced techniques in reinforcement learning (RL) to enhance agent performance and interpretability. One study introduces programmatic policies (PERL) as an alternative to neural policies (NER…
-
New Pair-GRPO algorithms enhance LLM alignment stability and generalization
Researchers have introduced the Pair-GRPO family, a novel theoretical framework designed to enhance the stability and generality of reinforcement learning for aligning large language models. This family includes two var…