EnvHarness: Awakening Static Worlds for Agent Learning Paper • 2608.19880 • Published 3 days ago • 245
Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill Paper • 2608.11924 • Published 11 days ago • 285
UEmbed: Unified Sparse and Dense Multimodal Embeddings Paper • 2608.02583 • Published 20 days ago • 50
TARS: Timestep-Aware Data Scaling for 3D-Free Video Re-Shooting Paper • 2607.28261 • Published 24 days ago • 116
K12-KGraph: A Curriculum-Aligned Knowledge Graph for Benchmarking and Training Educational LLMs Paper • 2605.09635 • Published about 1 month ago • 64
Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models Paper • 2607.12463 • Published Jul 14 • 108
Is One Layer Enough? Training A Single Transformer Layer Can Match Full-Parameter RL Training Paper • 2607.01232 • Published Jul 2 • 8
Multimodal Continuous Reasoning via Asymmetric Mutual Variational Learning Paper • 2607.00461 • Published Jul 1 • 29
Boundary-Aware Context Grounding for A Low-Channel EEG Agent Paper • 2606.26519 • Published Jun 25 • 2
Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models Paper • 2606.11324 • Published Jun 9 • 172
Skill is Not One-Size-Fits-All: Model-Aware Skill Alignment for LLM Agents Paper • 2605.30723 • Published May 29 • 17
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Paper • 2606.02031 • Published Jun 1 • 22
VisualThink-VLA: Visual Intermediate Reasoning for Effective and Low-Latency Vision-Language-Action Policies Paper • 2605.30011 • Published May 28 • 10
Evaluating Cognitive Age Alignment in Interactive AI Agents Paper • 2605.17894 • Published May 18 • 5
IndusAgent: Reinforcing Open-Vocabulary Industrial Anomaly Detection with Agentic Tools Paper • 2605.20682 • Published May 20 • 86
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Paper • 2605.21468 • Published May 20 • 51