arxiv:2504.15133
Xinle Deng
Linear-Matrix-Probability
AI & ML interests
Natural Language Processing, AI4Biology, Multimodal Learning, Machine Learning
Recent Activity
upvoted an article about 24 hours ago
From GRPO to DAPO and GSPO: What, Why, and How upvoted an article about 24 hours ago
A Guide to Reinforcement Learning Post-Training for LLMs: PPO, DPO, GRPO, and Beyond upvoted a paper 4 days ago
MobileMem: Learning from a Year of Mobile Experiences