PulseAugur
EN
LIVE 09:10:41
ENTITY HalfCheetah-v4

HalfCheetah-v4

PulseAugur coverage of HalfCheetah-v4 — every cluster mentioning HalfCheetah-v4 across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
0
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
2 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 2 TOTAL
  1. RESEARCH · CL_99607 ·

    New research explores advanced RL for agent survival, navigation, and explainability · 7 sources tracked

    Researchers are exploring advanced techniques in reinforcement learning (RL) to enhance agent performance and interpretability. One study introduces programmatic policies (PERL) as an alternative to neural policies (NER…

  2. TOOL · CL_21988 ·

    New Pair-GRPO algorithms enhance LLM alignment stability and generalization

    Researchers have introduced the Pair-GRPO family, a novel theoretical framework designed to enhance the stability and generality of reinforcement learning for aligning large language models. This family includes two var…