Post-Training Frontier Text-to-Image Models by Composing Preference and Rubric Rewards Paper • 2610.02967 • Published 9 days ago • 26
From Traces to Agentic Worlds: Agentic Language World Models for Interactive Environment Simulation Paper • 2610.06100 • Published 6 days ago • 126
From Prompting to Composing: A Spatial Canvas Interface for Poster Generation Paper • 2610.12230 • Published 3 days ago • 8
Memento 3: Model-Based Recursive Self-Improvement through Reflective Rulebooks Paper • 2610.11794 • Published 3 days ago • 28
OuroWorld: Bringing Any 3D World Alive as Diverse, Endlessly Looping 3D Cinemagraphs Paper • 2610.12461 • Published 3 days ago • 37
MiMo-V2.6: Scaling Reinforcement Learning Towards Self-Improvement Paper • 2610.11959 • Published 3 days ago • 65
SuperNav: An Agentic Navigation System for Any Task in Any Scene Paper • 2610.12126 • Published 3 days ago • 67
TokenRouter: Efficient Serving System for Token-Level LLM Routing Paper • 2610.12242 • Published 3 days ago • 124
Learn2Play Bench: How Well Do LLM Agents Learn from Experience in Unfamiliar Environments? Paper • 2610.08215 • Published 3 days ago • 132
view article Article Transformers now runs llama.cpp quants +1 marcsun13, ArthurZ, lysandre • 19 days ago • 105
Running 259 The ultimate guide to RL environments: building and scaling them in the LLM era 📝 259 Building and scaling RL environments for LLM training
DecepEval: A Benchmark for Evaluating Deception in LLM Agents Paper • 2610.07967 • Published 5 days ago • 74
Questioning the Questions: Sustaining Self-Evolution in Reasoning Models Paper • 2610.04299 • Published 8 days ago • 70
UltraText Bench: A Comprehensive Bilingual Benchmark for Evaluating Visual Text Rendering in Image Generation Paper • 2610.09823 • Published 4 days ago • 153
The Past Frames the Future: Memory for Autoregressive Video Generation Paper • 2609.28466 • Published 18 days ago • 65