Hao Liang
PulseAugur coverage of Hao Liang — every cluster mentioning Hao Liang across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
New research explains GRPO normalization's role in adaptive gradients
A new research paper explores the necessity and effectiveness of normalization in Group Relative Policy Optimization (GRPO), a standard algorithm for reinforcement learning in language models. The study, published on ar…
-
New benchmark evaluates LLMs' ability to find errors in synthetic math data
Researchers have introduced MathDebugger, a new benchmark designed to evaluate the ability of large language models to detect and diagnose errors within synthetic mathematical data. The benchmark includes a dataset of c…
-
Withdrawn paper analyzes differential privacy in wireless federated learning
This paper, titled "When Differential Privacy Meets Wireless Federated Learning: An Improved Analysis for Privacy and Convergence," was withdrawn by its author, Hao Liang. The research aimed to address limitations in ex…