Code2Games: Enabling Coding Agents for Gaming World Generation Paper • 2610.05033 • Published 3 days ago • 3
In-Distribution Forcing for Long Video Generation at Test Time Paper • 2610.03120 • Published 5 days ago • 25
RobotUse: Allocating Computation, Context, and Decisions Paper • 2610.04929 • Published 3 days ago • 11
When to Switch: Reliable Action-Chunk Extension for Vision-Language-Action Models Paper • 2610.05719 • Published 2 days ago • 9
CANOPY: Adaptive-Granularity Evidence Compression for Multimodal RAG Paper • 2610.00923 • Published 6 days ago • 14
The Tasteful Agent: Measuring and Improving Taste in Long-Horizon Tasks Paper • 2609.25804 • Published 15 days ago • 163
On-Policy or Off-Policy Learning? A Systematic Study of Distillation Dynamics Paper • 2609.35259 • Published 9 days ago • 190
Source Preference in the Wild: How LLM Agents Favor Items by Source, and How to Reduce It Paper • 2610.03195 • Published 5 days ago • 37
Science or Slop?: Benchmarking and Mitigating Scientific Slop in AI-Generated Papers Paper • 2610.00531 • Published 7 days ago • 54
Learning What to Recall: Adaptive Multi-Cue Episodic Memory for World Models Paper • 2609.34677 • Published 9 days ago • 12
Beyond the Current Scene: Event-Referential Grasping with Active View Selection Paper • 2609.39375 • Published 7 days ago • 51
World Observer: Joint Actor-Observer Generation for Persistent World Modeling Paper • 2610.02162 • Published 6 days ago • 83
Overcoming Scaling Limits in On-Policy Self-Distillation for LLM Reasoning Paper • 2609.37915 • Published 8 days ago • 8
A2Z GameSpec-Bench: How Faithfully Can Coding Agents Generate Games from Game Design Specifications? Paper • 2609.39564 • Published 7 days ago • 16
Imagine3D-LLM: Teaching MLLMs to Imagine 3D Scenes Before Answering Paper • 2609.38177 • Published 8 days ago • 71
RSIGame: Autonomous Agentic Game Development with Recursive Self-improvement Paper • 2609.39045 • Published 7 days ago • 90
The Teacher Is a Direction, Not a Destination: Extrapolating RL-Induced Representation Residuals in On-Policy Distillation Paper • 2609.36484 • Published 8 days ago • 578
Preference-Guided Adaptation for Open-Vocabulary Semantic Segmentation via Prompt Disagreement Paper • 2609.34528 • Published 9 days ago • 6
Anisotropic Representations Improve Planning in JEPA World Models Paper • 2609.37441 • Published 8 days ago • 34
Learning Beyond What You Sample: Off-Policy-Aware Cross-Model Trajectory Exchange for RLVR Paper • 2609.37868 • Published 8 days ago • 63