Announcement_25
[ICML 2026 Spotlight] We are delighted that our paper “Critique-GRPO: Advancing LLM Reasoning with Natural Language and Numerical Feedback” has been accepted to ICML 2026 as a Spotlight. Critique-GRPO integrates natural-language critiques with numerical rewards in online reinforcement learning, allowing a model to learn from both initial responses and critique-guided refinements. ![]()