Selection-Based Structured Reasoning: Toward Efficient Multimodal Search Agents Paper • 2610.01892 • Published 10 days ago • 32
What Gradients Add to Text Leakage in Split Language Models, Counted per Token and per Document Paper • 2610.04128 • Published 9 days ago • 12
Architect-Ant: Editable Automatic Furnishing of Architectural Floor Plans Paper • 2606.10953 • Published 11 days ago • 54
LEGO-Anything: Coding Agents for 3D Scene Reconstruction Paper • 2609.36380 • Published 13 days ago • 130
Just Ask Jev: Reinforcement Learning for Calibrated Decisions as a Zero-Shot Detector of AI Alignment Failures Paper • 2609.29429 • Published 17 days ago • 29
RayOrch: Programming and Executing Lineage-Controlled Multi-Grain Dataflows for Foundation-Model Data Preparation Paper • 2609.18703 • Published 25 days ago • 56
Tri-PvP: Exposing Modality Bias in Omni-Modal Large Language Models through Perceptual-Propositional Evidence Conflicts Paper • 2609.06011 • Published Sep 5 • 18
GameHorizon Suite: Multi-Horizon Data and Evaluation in Gameplay Paper • 2609.25001 • Published 20 days ago • 132
RecreationWorld: Scalable and Verifiable Environments for Hybrid Computer-Use Agents Paper • 2609.22000 • Published 23 days ago • 80
When EOS Tokens Disagree: Understanding Length Inflation in On-Policy Distillation Paper • 2609.20511 • Published 24 days ago • 115
Zing-0.5: Toward Playable Worlds with Real-Time Joint Action and Text Control Paper • 2609.17909 • Published 26 days ago • 47
HarnessVLN: Unifying Training-Free Embodied Navigation through an Agent Harness Paper • 2609.15195 • Published 27 days ago • 22
Another Blueprint In The Wall: How to Ask Frontier AI Like a Kid? Paper • 2609.14803 • Published 28 days ago • 12
ModularRSI: Modular and Generalizable Recursive Harness Self-Improvement Paper • 2609.14857 • Published 27 days ago • 215
Grouped Value Attention: Efficient KV Caching via On-Demand Key Reconstruction Paper • 2609.13285 • Published Sep 8 • 82
COBRA-Skills: Contextual Bandit-Guided Evolution for Agent Skill Optimization Paper • 2609.11682 • Published Sep 10 • 46
An Open Recipe for IMO Gold: Training Nemotron for Olympiad Mathematics Paper • 2609.10712 • Published Sep 9 • 45