Are LLMs Ready for Scientific Discovery? A Capability-Oriented Benchmark for AI Scientists Paper • 2607.11079 • Published 7 days ago • 8
LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget Paper • 2607.14952 • Published 4 days ago • 177
MuseBench: Benchmarking Intent-Level Audiovisual Arts Understanding in MLLMs Paper • 2606.30026 • Published 21 days ago • 6
Lexical Consensus: Grounded Word Learning and Shared Meaning in Artificial Agents Paper • 2606.22207 • Published 29 days ago • 4
DataEvolver: Self-Evolving Multi-Agent Data Construction for Text-Rich Image Generation Paper • 2606.31537 • Published 20 days ago • 28
OmniDirector: General Multi-Shot Camera Cloning without Cross-Paired Data Paper • 2606.13432 • Published Jun 11 • 113
Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models Paper • 2606.03988 • Published Jun 3 • 126
electricsheepafrica/africa-owid-foreign-aid-given-per-capita-vs-gdp-per-capita Viewer • Updated Jun 3 • 1.8k • 35 • 1
Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players Paper • 2605.28816 • Published May 27 • 433
GRASP: Learning to Ground Social Reasoning in Multi-Person Non-Verbal Interactions Paper • 2605.15764 • Published May 15 • 4
Ethical Hyper-Velocity (EHV): A Provably Deterministic Governance-Aware JIT Compiler Architecture for Agentic Systems Paper • 2605.17909 • Published May 18 • 4