Metacognition in LLMs: Foundations, Progress, and Opportunities Paper • 2607.11881 • Published 8 days ago • 27
Measuring the Gap Between Human and LLM Research Ideas Paper • 2607.01233 • Published 20 days ago • 20
Dockerless: Environment-Free Program Verifier for Coding Agents Paper • 2606.28436 • Published 25 days ago • 111
Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs Paper • 2606.32032 • Published 21 days ago • 29
GUI vs. CLI: Execution Bottlenecks in Screen-Only and Skill-Mediated Computer-Use Agents Paper • 2606.24551 • Published 29 days ago • 28
GUI vs. CLI: Execution Bottlenecks in Screen-Only and Skill-Mediated Computer-Use Agents Paper • 2606.24551 • Published 29 days ago • 28
Qwen-AgentWorld: Language World Models for General Agents Paper • 2606.24597 • Published 28 days ago • 149
VibeThinker-3B: Exploring the Frontier of Verifiable Reasoning in Small Language Models Paper • 2606.16140 • Published Jun 15 • 123
EurekAgent: Agent Environment Engineering is All You Need For Autonomous Scientific Discovery Paper • 2606.13662 • Published Jun 11 • 30
Benchmarking AI Agents for Addressing Scientific Challenges Across Scales Paper • 2606.12736 • Published Jun 10 • 5
VideoKR: Towards Knowledge- and Reasoning-Intensive Video Understanding Paper • 2606.05259 • Published Jun 3 • 39
VideoKR: Towards Knowledge- and Reasoning-Intensive Video Understanding Paper • 2606.05259 • Published Jun 3 • 39