Skip to content

Uranus: Building the Next-Generation Simulation Infrastructure for Embodied AI

Sep 2026 · 0 citations
Computer Science

TL;DR

This work presents Uranus, a data-driven robot simulator built around a joint-trajectory-conditioned autoregressive diffusion model, providing a unified interface for synchronized multi-view generation across diverse robot embodiments and camera configurations.

Abstract

Scalable simulation is essential for robot data generation, policy training, evaluation, and safe iteration, yet real-world interaction is costly and conventional simulators require labor-intensive construction. We present Uranus, a data-driven robot simulator built around a joint-trajectory-conditioned autoregressive diffusion model. Uranus offers three key capabilities: (1) streaming, open-ended rollout, which receives future joint-position trajectories online and autoregressively generates one latent frame per step, corresponding to four RGB frames, without a fixed horizon; (2) low-latency generation, achieving 24 FPS after inference optimization; and (3) scalable, extensible robot control, providing a unified interface for synchronized multi-view generation across diverse robot embodiments and camera configurations. We conduct comprehensive quantitative and qualitative evaluations on both in-distribution and out-of-distribution data, providing an objective assessment of Uranus and clearly identifying its current limitations. We release the code and model weights to empower the community with practical tools and insights.

View source

Similar papers

Preprint Sep 2026

EmbodiRSI: Recursive Self-Improvement for Data-Efficient Robot Adaptation

Adapting robot manipulation policies to new tasks and environments remains highly data-intensive, while the data needed for further improvement depends on the policy's current capabilities and failure modes. We introduce EmbodiRSI, an agentic system for recursive self-improvement (RSI) in a real-to-sim-to-real setting,...

Hao-Ran Lang, Hao-Tao Lu, Shi-Yu Sang et al. · 1 citation
Open access Sep 2026

Fetch My Beer: Synthetic-to-Real Hierarchical Policy for Smooth Pick-and-Place

Many real-world robotic applications require dynamically sensitive manipulation, where success depends not only on reaching a target state but on maintaining stable object dynamics throughout execution. We study the stable transport of liquid-filled containers, where a robot must move objects to target locations while...

Ying-Yue Li, Chenyangguang Zhang, Rui-Da Zhang et al. · 0 citations
#artificial intelligence Preprint Sep 2026

FluxVLA Engine: A One-Stop VLA Engineering Platform for Embodied Intelligence

Vision-language-action (VLA) models, world-action models (WAMs), and offline reinforcement learning methods are rapidly expanding the design space of embodied policies, yet turning these algorithms into reliable robot systems remains constrained by fragmented data formats, training stacks, evaluation protocols, inferen...

Yin-Hao Li, Wei-Xin Mao, Zi-Han Lan et al. · 1 citation
Preprint Sep 2026

DynaForge: Planning-Guided Residual Learning for Dynamic Manipulation Demonstration Generation

DynaForge combines low-frequency global planning with high-frequency object-centric inverse kinematics across task phases, and applies a residual policy to correct actions during dynamic interaction, showing the ability of DynaForge for sim-to-real transfer.

Yi-Yang Jin, Yu Zheng, Xiao He et al. · 0 citations
#artificial intelligence Preprint Sep 2026

AeroManip-VLA: Scalable Vision-Language-Action Learning for Aerial Manipulation with RL-Generated Demonstrations

Aerial manipulators extend robotic manipulation into 3D workspaces that are difficult for ground-based robots to access, creating new opportunities for general-purpose manipulation. However, extending Vision-Language-Action (VLA) models to aerial robots introduces distinct challenges due to the tight coupling between m...

Rui Huang, Yan-Lin Mu, Li-Dong Li et al. · 0 citations

Related blog posts

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.