NLU Workshop Talk: Model-Aided Human Annotation at Scale
AuthorsHadas Kotek
NLU Workshop Talk: Model-Aided Human Annotation at Scale
AuthorsHadas Kotek
GRPO Beyond English: A Large-Scale Study of GRPO in Non-English and Multilingual Settings
August 18, 2026research area Speech and Natural Language Processing
Reinforcement Learning with Verifiable Rewards (RLVR), often optimized with Group Relative Policy Optimization (GRPO), has become a central recipe for improving the reasoning capabilities of pretrained language models but current studies remain heavily English-centric. We conduct a large-scale empirical study of multilingual and non-English GRPO across a wide range of base models, training languages, and different reasoning language rewards. We…
When Unlearning Is Free: Leveraging Low Influence Points to Reduce Computational Costs
August 13, 2026research area Data Science and Annotation, research area Privacy
As concerns around data privacy in machine learning grow, the ability to unlearn, or remove, specific data points from trained models becomes increasingly important. While state of the art unlearning methods have emerged in response, they typically treat all points in the forget set equally. In this work, we challenge this approach by asking whether points that have a negligible impact on the model’s learning need to be removed. Through a…