LMBuild: Evaluating LLM Agents for Generating Buildable and Functional Structures Paper • 2610.04292 • Published 6 days ago • 38 • 3
Ego2Act: Evaluating Goal-Directed Manipulation in Egocentric Video Generation Paper • 2610.01092 • Published 8 days ago • 34 • 3
Scaling Trajectories for Complex Tasks through Recursive Self-Rewrite Paper • 2610.02826 • Published 7 days ago • 103 • 2
FrameMorrow: Future-guided Frame Selection with Prospective Tokens for Long-Horizon Video Generation Paper • 2609.38839 • Published 9 days ago • 113 • 4
Does Learning Protein Folding Generalize to Broader Reasoning? Paper • 2609.38879 • Published 9 days ago • 136 • 4
Adaptive Reward Routing: Dynamic Multi-Reward Optimization for Joint Audio-Video Diffusion via Forward-Process RL Paper • 2609.37200 • Published 10 days ago • 139 • 3
GraphForge: Training Working Agents with Graph-Anchored Workspace Synthesis Paper • 2609.38923 • Published 9 days ago • 149 • 3
It's Not What the Image Shows: Irrelevant Context Destabilises VLM Judges Without Informing Them Paper • 2609.37863 • Published 10 days ago • 39 • 3
Unmask the State: When Does State Adaptation Matter for Masked Diffusion Language Models Paper • 2609.33355 • Published 12 days ago • 51 • 3
TimeEvo: Failure-Driven Self-Evolution of a Time Series Agent Paper • 2609.27277 • Published 16 days ago • 33 • 3
In-Flight KV Cache with Clean Anchors for Faster Autoregressive Video Diffusion Paper • 2609.32540 • Published 13 days ago • 35 • 2
Learning from Teacher Continuations at Student States Paper • 2609.36246 • Published 11 days ago • 40 • 2
RSIGame: Autonomous Agentic Game Development with Recursive Self-improvement Paper • 2609.39045 • Published 9 days ago • 93 • 2
More Choices, Fewer Decisions: Ordinal-Scale Bias in JEV-like Direct-Decision Models Paper • 2609.38827 • Published 9 days ago • 61 • 3
Skill2Env: Capability-Oriented Environment Synthesis from Skills for General Agents Paper • 2609.33772 • Published 12 days ago • 35 • 7
EVOKE: Eliciting World Knowledge in Agents for Transferable Decision-Making Paper • 2609.38334 • Published 8 days ago • 80 • 3
AREX-2: Advancing Self-Improving Agents through Long-Horizon Reflective Tasks Paper • 2609.38288 • Published 10 days ago • 143 • 3
UniEvo-VL: An On-policy Self-Distillation Training Recipe for Multimodal Model Self-improvement Paper • 2609.38721 • Published 9 days ago • 295 • 3
WorldAuditBench: Interactive 3D World Auditing with Multimodal Agents Paper • 2609.40325 • Published 9 days ago • 104 • 3