ENTITY
Boltzmann policy
Boltzmann policy
PulseAugur coverage of Boltzmann policy — every cluster mentioning Boltzmann policy across labs, papers, and developer communities, ranked by signal.
Total · 30d
2
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
2
2 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D
2 day(s) with sentiment data
RECENT · PAGE 1/1 · 2 TOTAL
-
New paper axiomatizes Boltzmann rationality in reinforcement learning
A new paper introduces an axiomatic characterization for Boltzmann rationality, a common model of stochastic choice in reinforcement learning. The research distinguishes between randomness in choice and environmental ch…
-
New Soft Q(λ) method enhances off-policy reinforcement learning
Researchers have introduced Soft $Q(\lambda)$, a novel multi-step off-policy method for entropy-regularized reinforcement learning. This framework extends existing soft Q-learning techniques by enabling efficient credit…