arxiv:2608.00220
Shaohang Wei
SylvainWei
AI & ML interests
NLP, LLM
Recent Activity
authored a paper 3 days ago
Verifier-Induced Support Reshaping in On-Policy Optimization authored a paper 3 days ago
Experience Augmented Policy Optimization for LLM Reasoning authored a paper 3 days ago
Sparse but Critical: A Token-Level Analysis of Distributional Shifts in RLVR Fine-Tuning of LLMs