Wake up for Touch! Mask-isolated Tactile Alignment Learning in MLLMs Paper • 2607.00302 • Published 27 days ago • 2
LLM-as-a-Verifier: A General-Purpose Verification Framework Paper • 2607.05391 • Published 22 days ago • 15
Lexical Consensus: Grounded Word Learning and Shared Meaning in Artificial Agents Paper • 2606.22207 • Published Jun 20 • 4
Retrospective Harness Optimization: Improving LLM Agents via Self-Preference over Trajectory Rollouts Paper • 2606.05922 • Published Jun 4 • 70
TVIR: Building Deep Research Agents Towards Text--Visual Interleaved Report Generation Paper • 2606.02320 • Published Jun 1 • 15
Convex Low-resource Accent-Robust Language Detection in Speech Recognition Paper • 2605.23235 • Published May 22 • 6
junbrro/n1_5_gr1_cotrain_optionY_dust_fk_temp_b636_bsz128x2_15000_layer12_lr3e-5_A_full_vitra_only Updated May 30 • 1
DelTA: Discriminative Token Credit Assignment for Reinforcement Learning from Verifiable Rewards Paper • 2605.21467 • Published May 20 • 207
OSCAR: Offline Spectral Covariance-Aware Rotation for 2-bit KV Cache Quantization Paper • 2605.17757 • Published May 18 • 66
Spreadsheet-RL: Advancing Large Language Model Agents on Realistic Spreadsheet Tasks via Reinforcement Learning Paper • 2605.22642 • Published May 21 • 35