RSIAgent: Autonomous Exploration for Recursive Self-improvement in New Environments Paper • 2609.15364 • Published 7 days ago • 78
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 7 days ago • 239
Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation Paper • 2609.11115 • Published 11 days ago • 171
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search Paper • 2609.13356 • Published 10 days ago • 256
NCP-ArchPreview Technical Report: Moving towards Latent Space Language Models through Next Concept Prediction Paper • 2609.10715 • Published 12 days ago • 327
SenseNova-U1.5: Towards Native Unified Visual Intelligence Paper • 2609.11929 • Published 11 days ago • 270
Evaluating Multimodal LLMs as Generalist Vision-Language-Action Agents for Drone Control: Commanding, Approaching, Tracking and Searching Paper • 2609.01404 • Published 20 days ago • 28
Harness-of-Harness: Multi-Day Autonomous Software Development with Continual Improvement Paper • 2609.01481 • Published 20 days ago • 19
The Mechanics of Democratic Dominance: A System Dynamics Paradigm for Dynamic Consent Engineering Paper • 2608.27509 • Published 21 days ago • 5
Scaling Large Reasoning Models beyond Human Supervision: A Path toward Superintelligence Paper • 2608.31075 • Published 21 days ago • 29
Agentic Artifact Creation: Systems, Evaluation, Principles, and Opportunities Paper • 2608.28122 • Published 24 days ago • 66
LoopArena: Benchmarking Models as Runtime Controllers for Loop Engineering Paper • 2608.28281 • Published 24 days ago • 98
SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers Paper • 2609.01343 • Published 20 days ago • 89