-
Towards Scalable Pre-training of Visual Tokenizers for Generation
Paper • 2512.13687 • Published • 108 -
MMGR: Multi-Modal Generative Reasoning
Paper • 2512.14691 • Published • 121 -
Coupling Experts and Routers in Mixture-of-Experts via an Auxiliary Loss
Paper • 2512.23447 • Published • 100 -
LiveTalk: Real-Time Multimodal Interactive Video Diffusion via Improved On-Policy Distillation
Paper • 2512.23576 • Published • 66
Collections
Discover the best community collections!
Collections including paper arxiv:2607.13285
-
Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling
Paper • 2607.02980 • Published • 84 -
Gemma 4 Technical Report
Paper • 2607.02770 • Published • 80 -
SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe
Paper • 2607.03451 • Published • 35 -
TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training
Paper • 2607.05804 • Published • 20
-
Multi-Agent Computer Use
Paper • 2606.01533 • Published • 7 -
OpenSkill: Open-World Self-Evolution for LLM Agents
Paper • 2606.06741 • Published • 29 -
Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills
Paper • 2606.07412 • Published • 12 -
Bayesian-Agent: Posterior-Guided Skill Evolution for LLM Agent Harnesses
Paper • 2606.08348 • Published • 16
-
Agentic Reasoning for Large Language Models
Paper • 2601.12538 • Published • 208 -
From Code Foundation Models to Agents and Applications: A Practical Guide to Code Intelligence
Paper • 2511.18538 • Published • 307 -
Agent Learning via Early Experience
Paper • 2510.08558 • Published • 276 -
Weak-Driven Learning: How Weak Agents make Strong Agents Stronger
Paper • 2602.08222 • Published • 290
-
tencent/HiLS-Attention-7B
Text Generation • 7B • Updated • 365 • 24 -
Metacognition in LLMs: Foundations, Progress, and Opportunities
Paper • 2607.11881 • Published • 30 -
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable
Paper • 2607.13285 • Published • 235 -
Bonsai 27B WebGPU Kernels
🌳440Run a 1-bit 27B LLM locally in your browser on WebGPU
-
LoomVideo: Unifying Multimodal Inputs into Video Generation and Editing
Paper • 2606.06042 • Published • 24 -
Optimizing Visual Generative Models via Distribution-wise Rewards
Paper • 2607.02291 • Published • 18 -
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable
Paper • 2607.13285 • Published • 235 -
Self-Improvements in Modern Agentic Systems: A Survey
Paper • 2607.13104 • Published • 34
-
Beyond Semantic Similarity: Rethinking Retrieval for Agentic Search via Direct Corpus Interaction
Paper • 2605.05242 • Published • 127 -
MIGRATION-BENCH: Repository-Level Code Migration Benchmark from Java 8
Paper • 2505.09569 • Published • 7 -
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable
Paper • 2607.13285 • Published • 235
-
ComfyUI-R1: Exploring Reasoning Models for Workflow Generation
Paper • 2506.09790 • Published • 53 -
Saffron-1: Towards an Inference Scaling Paradigm for LLM Safety Assurance
Paper • 2506.06444 • Published • 73 -
DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents
Paper • 2506.11763 • Published • 74 -
Agentic Reasoning: Reasoning LLMs with Tools for the Deep Research
Paper • 2502.04644 • Published • 4
-
Towards Scalable Pre-training of Visual Tokenizers for Generation
Paper • 2512.13687 • Published • 108 -
MMGR: Multi-Modal Generative Reasoning
Paper • 2512.14691 • Published • 121 -
Coupling Experts and Routers in Mixture-of-Experts via an Auxiliary Loss
Paper • 2512.23447 • Published • 100 -
LiveTalk: Real-Time Multimodal Interactive Video Diffusion via Improved On-Policy Distillation
Paper • 2512.23576 • Published • 66
-
tencent/HiLS-Attention-7B
Text Generation • 7B • Updated • 365 • 24 -
Metacognition in LLMs: Foundations, Progress, and Opportunities
Paper • 2607.11881 • Published • 30 -
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable
Paper • 2607.13285 • Published • 235 -
Bonsai 27B WebGPU Kernels
🌳440Run a 1-bit 27B LLM locally in your browser on WebGPU
-
Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling
Paper • 2607.02980 • Published • 84 -
Gemma 4 Technical Report
Paper • 2607.02770 • Published • 80 -
SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe
Paper • 2607.03451 • Published • 35 -
TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training
Paper • 2607.05804 • Published • 20
-
LoomVideo: Unifying Multimodal Inputs into Video Generation and Editing
Paper • 2606.06042 • Published • 24 -
Optimizing Visual Generative Models via Distribution-wise Rewards
Paper • 2607.02291 • Published • 18 -
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable
Paper • 2607.13285 • Published • 235 -
Self-Improvements in Modern Agentic Systems: A Survey
Paper • 2607.13104 • Published • 34
-
Multi-Agent Computer Use
Paper • 2606.01533 • Published • 7 -
OpenSkill: Open-World Self-Evolution for LLM Agents
Paper • 2606.06741 • Published • 29 -
Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills
Paper • 2606.07412 • Published • 12 -
Bayesian-Agent: Posterior-Guided Skill Evolution for LLM Agent Harnesses
Paper • 2606.08348 • Published • 16
-
Beyond Semantic Similarity: Rethinking Retrieval for Agentic Search via Direct Corpus Interaction
Paper • 2605.05242 • Published • 127 -
MIGRATION-BENCH: Repository-Level Code Migration Benchmark from Java 8
Paper • 2505.09569 • Published • 7 -
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable
Paper • 2607.13285 • Published • 235
-
Agentic Reasoning for Large Language Models
Paper • 2601.12538 • Published • 208 -
From Code Foundation Models to Agents and Applications: A Practical Guide to Code Intelligence
Paper • 2511.18538 • Published • 307 -
Agent Learning via Early Experience
Paper • 2510.08558 • Published • 276 -
Weak-Driven Learning: How Weak Agents make Strong Agents Stronger
Paper • 2602.08222 • Published • 290
-
ComfyUI-R1: Exploring Reasoning Models for Workflow Generation
Paper • 2506.09790 • Published • 53 -
Saffron-1: Towards an Inference Scaling Paradigm for LLM Safety Assurance
Paper • 2506.06444 • Published • 73 -
DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents
Paper • 2506.11763 • Published • 74 -
Agentic Reasoning: Reasoning LLMs with Tools for the Deep Research
Paper • 2502.04644 • Published • 4