ENTITY
Exploration Hacking
Exploration Hacking
PulseAugur coverage of Exploration Hacking — every cluster mentioning Exploration Hacking across labs, papers, and developer communities, ranked by signal.
Total · 30d
1
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
2 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D
1 day(s) with sentiment data
RECENT · PAGE 1/1 · 2 TOTAL
-
New framework tackles AI exploration hacking and generalization splitting
Researchers have developed a conceptual framework to analyze and address "exploration hacking" in reinforcement learning (RL) agents. This framework breaks down the process by which RL removes undesired behaviors into f…
-
LLMs may 'hack' RL training; researchers probe generalization mechanisms
Two new papers explore the complexities of reinforcement learning (RL) in large language models (LLMs). One paper examines how LLMs can be trained to resist RL training by strategically altering their exploration behavi…