Apodex 1.1: Scaling Agentic Intelligence for Complex Work Paper • 2608.23283 • Published 17 days ago • 206
Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning Paper • 2609.03430 • Published 7 days ago • 172
Rethinking On-Policy Distillation of Large Language Models II: One Training Example Paper • 2609.04172 • Published 7 days ago • 84
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning Paper • 2608.09888 • Published about 1 month ago • 778
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement Paper • 2608.31046 • Published 10 days ago • 147
Repo-To-Skill: Distilling GitHub Repositories Into AI4AI Skills Paper • 2609.02749 • Published 8 days ago • 538
UserBench: An Interactive Gym Environment for User-Centric Agents Paper • 2507.22034 • Published Jul 29, 2025 • 31
view article Article Kimi K3 Model Overview: 2.8T Parameters, MXFP4 Quantization, and What the Open Weights Mean for the Community ResterChed • Jul 17 • 197
RLCSD: Reinforcement Learning with Contrastive On-Policy Self-Distillation Paper • 2606.11709 • Published Jun 10 • 1
Learning from Language Feedback via Variational Policy Distillation Paper • 2605.15113 • Published May 18 • 13
Dr-DCI: Scaling Direct Corpus Interaction via Dynamic Workspace Expansion Paper • 2606.14885 • Published Jun 12 • 13
Rethinking Continual Experience Internalization for Self-Evolving LLM Agents Paper • 2606.04703 • Published Jun 3 • 27