The Galaxy's Guide to the Tokenizer: A Benchmark for Scientific Foundation Models Paper • 2606.25610 • Published Jun 24 • 4
Learning to Fold: prizewinning solution at LeHome Challenge 2026 (1st place online, 2nd offline) Paper • 2606.27163 • Published Jun 25 • 6
Parallel Rollout Approximation for Pixel-Space Autoregressive Image Generation Paper • 2606.27978 • Published Jun 26 • 7
MemoBench: Benchmarking World Modeling in Dynamically Changing Environments Paper • 2606.27537 • Published Jun 25 • 7
Object-Centric Residual RL for Zero-Shot Sim-to-Real VLA Enhancement Paper • 2606.18953 • Published Jun 17 • 8
AgentOdyssey: Open-Ended Long-Horizon Text Game Generation for Test-Time Continual Learning Agents Paper • 2606.24893 • Published May 29 • 9
Ko-WideSearch: A Korean Breadth-Search Benchmark for Exhaustive Set Enumeration by Web Agents Paper • 2606.27595 • Published Jun 25 • 9
Towards Automating Scientific Review with Google's Paper Assistant Tool Paper • 2606.28277 • Published Jun 26 • 12
Thinking While Speaking: Inference-Time Knowledge Transfer for Responsive and Intelligent Conversational Voice Agents Paper • 2511.07397 • Published Jul 1 • 13
GBC: Gradient-Based Connections for Optimizing Multi-Agent Systems Paper • 2606.28187 • Published Jun 26 • 14
SimFoundry: Modular and Automated Scene Generation for Policy Learning and Evaluation Paper • 2606.28276 • Published Jun 26 • 18
Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System Paper • 2606.18112 • Published Jun 18 • 30
Translation as a Bridging Action: Transferring Manipulation Skills from Humans to Robots Paper • 2606.28133 • Published Jun 26 • 41
Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models Paper • 2606.17846 • Published Jun 17 • 34
Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs Paper • 2606.27378 • Published May 7 • 61
PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation Paper • 2606.28128 • Published Jun 26 • 54
When Does Combining Language Models Help? A Co-Failure Ceiling on Routing, Voting, and Mixture-of-Agents Across 67 Frontier Models Paper • 2606.27288 • Published Jun 25 • 5