UltraText Bench: A Comprehensive Bilingual Benchmark for Evaluating Visual Text Rendering in Image Generation Paper • 2610.09823 • Published 4 days ago • 153
LoGRA: Scaling LLM Reinforcement Learning with Low-Rank Gradient Sketches Paper • 2610.06647 • Published 6 days ago • 148
ProgressCompass: Embodied Progress Reward Models Are Lost Without the Right Context Paper • 2609.36684 • Published 12 days ago • 20
VIEScore2: Unified Image Evaluation with Spatially Grounded Explanations Paper • 2610.00994 • Published 10 days ago • 26
MIMESIS: Learning User Simulators as Training Environments for Interactive Agents Paper • 2610.09484 • Published 4 days ago • 23
Rethinking Cross-Tokenizer On-Policy Distillation: From Alignment Coverage to Supervision Reliability Paper • 2610.08448 • Published 5 days ago • 203
Gains and Collapse in On-Policy Distillation:A Reinforcement Learning Perspective Paper • 2610.03185 • Published 9 days ago • 29
CADFather: Autonomous CAD Reconstruction through Coordinated Tool Use Paper • 2610.09127 • Published 5 days ago • 13
A self-learning scientific agent for X-ray diffraction Paper • 2610.07862 • Published 5 days ago • 17
Recursive Game Creator: An Agentic Product-Level Experience-Oriented Game Harness Paper • 2610.08621 • Published 5 days ago • 87
Questioning the Questions: Sustaining Self-Evolution in Reasoning Models Paper • 2610.04299 • Published 8 days ago • 70
Mechanics of Long-Context Hybrid Models Part 1.1: From Hybrid Attention to Hybrid Position Paper • 2610.10114 • Published 4 days ago • 31
Improving Proactive AI Assistance with Hierarchical Procedural Understanding Paper • 2610.06505 • Published 5 days ago • 9
Internalizing Agent Experience into Diffusion Model Weights via On-Policy Context Distillation Paper • 2610.07250 • Published 6 days ago • 10
From Pareto to Preference: Personalized Test-Time Scaling via Amortized Agentic Policy Discovery Paper • 2610.09684 • Published 4 days ago • 16
RobotWorld: Benchmarking Multimodal Agents for Robot Use Across Diverse Tasks and Embodiments Paper • 2610.10409 • Published 4 days ago • 32
We Query, Therefore We Compute: On Oracle Computation beyond the Machine, with an Application to Agents Paper • 2610.09243 • Published 4 days ago • 13
TRIAGE: Direction-Aware Mismatch Stabilization of Native NVFP4 Reinforcement Learning Paper • 2610.07043 • Published 6 days ago • 42