World Action Modeling with Progressive Visual Planning Paper • 2610.02508 • Published 11 days ago • 97
GDSD: Reinforcement Learning as Guided Denoiser Self-Distillation for Diffusion Language Models Paper • 2605.29398 • Published May 28 • 5
LLM-WikiRace Benchmark: How Far Can LLMs Plan over Real-World Knowledge Graphs? Paper • 2602.16902 • Published Feb 18 • 10
diffusion-reasoning/LLaDA-8B-Instruct-wd1-acecode-iter180 Feature Extraction • 8B • Updated Nov 15, 2025 • 17
diffusion-reasoning/LLaDA-8B-Instruct-wd1-acecode-iter180 Feature Extraction • 8B • Updated Nov 15, 2025 • 17
diffusion-reasoning/LLaDA-8B-Instruct-wd1-acecode-iter100 Feature Extraction • 8B • Updated Nov 14, 2025 • 15
diffusion-reasoning/LLaDA-8B-Instruct-wd1-acecode-iter100 Feature Extraction • 8B • Updated Nov 14, 2025 • 15
diffusion-reasoning/LLaDA-8B-Instruct-wd1-acecode-iter60 Feature Extraction • 8B • Updated Nov 14, 2025 • 15
diffusion-reasoning/LLaDA-8B-Instruct-wd1-acecode-iter60 Feature Extraction • 8B • Updated Nov 14, 2025 • 15
xiaohangt/LLaDA-8B-Instruct-wd1ucllfinal_mdpoadv-numinas_checkpoint-20 Feature Extraction • 8B • Updated Oct 9, 2025 • 13
xiaohangt/LLaDA-8B-Instruct-wd1ucllfinal_mdpoadv-numinas_checkpoint-20 Feature Extraction • 8B • Updated Oct 9, 2025 • 13
xiaohangt/LLaDA-8B-Instruct-wd1d1-maths_checkpoint-30 Feature Extraction • 8B • Updated Oct 9, 2025 • 11
xiaohangt/LLaDA-8B-Instruct-wd1d1-maths_checkpoint-30 Feature Extraction • 8B • Updated Oct 9, 2025 • 11