Recent interesting robotics papers
HIL-SERL: Precise and Dexterous Robotic Manipulation via Human-in-the-Loop Reinforcement Learning
- sparse reward classifier
- human in the loop online RL training
#arxiv #RL
Project Website
- sparse reward classifier
- human in the loop online RL training
#arxiv #RL
Project Website
IMLE Policy: Fast and Sample Efficient Visuomotor Policy Learning via Implicit Maximum Likelihood Estimation
- state of art BC comparing to diffusion
- faster inference speed
#RSS0225 #BC
X Overview, Project Website
- state of art BC comparing to diffusion
- faster inference speed
#RSS0225 #BC
X Overview, Project Website
Scalable Real2Sim: Physics-Aware Asset Generation Via Robotic Pick-and-Place Setups
- automatically generated simulation objects from the real world
- detailed analysis of object parameter estimation procedure
#arxiv #real2sim #reconstruction
X Overview, Project Website
- automatically generated simulation objects from the real world
- detailed analysis of object parameter estimation procedure
#arxiv #real2sim #reconstruction
X Overview, Project Website
PARC: Physics-based Augmentation with Reinforcement Learning for Character Controllers
- diffusion model in motion generation
- combine RL in data augmentation
#SIGGRAPH2025 #diffusion #RL
X Overview, Project Website
- diffusion model in motion generation
- combine RL in data augmentation
#SIGGRAPH2025 #diffusion #RL
X Overview, Project Website
Visuomotor Policies to Grasp Anything with Dexterous Hands
- grasping based on stero input
- engineering example of teacher-student, sim2real, and geometric fabrics
#arxiv #dexterous #grasp #sim2real
X Overview, Project Website, Geometric Fabrics
- grasping based on stero input
- engineering example of teacher-student, sim2real, and geometric fabrics
#arxiv #dexterous #grasp #sim2real
X Overview, Project Website, Geometric Fabrics
GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data
- combination of sim and real could help train robust grasping policy
- engineering example of pi-0 policy (VLM + flow matching)
#arxiv #grasp #simulation #sim2real #flow_matching
X Overview, Project Website, NVidia GraspGen
- combination of sim and real could help train robust grasping policy
- engineering example of pi-0 policy (VLM + flow matching)
#arxiv #grasp #simulation #sim2real #flow_matching
X Overview, Project Website, NVidia GraspGen
FunGrasp: Functional Grasping for Diverse Dexterous Hands
- detailed explanation of engineering
- demonstrate the importance of contact signal
#IROS2025 #dexterous #grasp
X Overview, Project Website
- detailed explanation of engineering
- demonstrate the importance of contact signal
#IROS2025 #dexterous #grasp
X Overview, Project Website
FALCON: Learning Force-Adaptive Humanoid Loco-Manipulation
- combines upper and lower body for policy training
- curriculum training for force adaptation
#arxiv #humanoid #locomanip
X Overview, Project Website
- combines upper and lower body for policy training
- curriculum training for force adaptation
#arxiv #humanoid #locomanip
X Overview, Project Website
Steerable Scene Generation with Post Training and Inference-Time Search
- use a diffusion-based framework to generate scalable scenes on procedural data
- use a Monte Carlo Tree Search to find physical feasibility and satisfy downstream objective
#arxiv #diffusion #generative
X Overview, Project Website
- use a diffusion-based framework to generate scalable scenes on procedural data
- use a Monte Carlo Tree Search to find physical feasibility and satisfy downstream objective
#arxiv #diffusion #generative
X Overview, Project Website
AMO:
Adaptive Motion Optimization for Hyper-Dexterous Humanoid Whole-Body Control
- framework that integrates sim-to-real reinforcement learning (RL) with trajectory optimization for real-time, adaptive whole-body control
#RSS2025 #humanoid
X overview, Project Website
Adaptive Motion Optimization for Hyper-Dexterous Humanoid Whole-Body Control
- framework that integrates sim-to-real reinforcement learning (RL) with trajectory optimization for real-time, adaptive whole-body control
#RSS2025 #humanoid
X overview, Project Website
Unified World Models: Coupling Video and Action Diffusion
for Pretraining on Large Robotic Datasets
- combine video and action diffusion in one transformer, can be train from robot trajectories and action-free videos.
- represents a policy, a forward dynamics model, an inverse dynamics model, and a video prediction model in a unified framework.
#RSS2025 #diffusion
X overview, Project Website
for Pretraining on Large Robotic Datasets
- combine video and action diffusion in one transformer, can be train from robot trajectories and action-free videos.
- represents a policy, a forward dynamics model, an inverse dynamics model, and a video prediction model in a unified framework.
#RSS2025 #diffusion
X overview, Project Website
TWIST Teleoperated Whole-Body Imitation System
- single network for whole-body control
#arxiv #humanoid
X Overview, Project Website
- single network for whole-body control
#arxiv #humanoid
X Overview, Project Website
PyRoki A Modular Toolkit for Robot Kinematic Optimization
- better performance than cuRobo
- GPU-assisted trajectory optimization toolbox
#arxiv #toolbox #trajopt
X Overview, Project Website
- better performance than cuRobo
- GPU-assisted trajectory optimization toolbox
#arxiv #toolbox #trajopt
X Overview, Project Website
VideoMimic Visual imitation enables contextual humanoid control
- used human video to learn policy for humanoid
- real-to-sim-to-real tendency
#arxiv #real2sim2real #humanoid #reconstruction
X Overview, Project Website
- used human video to learn policy for humanoid
- real-to-sim-to-real tendency
#arxiv #real2sim2real #humanoid #reconstruction
X Overview, Project Website
Instant Policy: In-Context Imitation Learning via Graph Diffusion
- formulate in-context imitation learning as a graph generation problem
- demonstrate impressive one-shot learning performance
#ICLR2025 #manipulation #in_context
X Overview, Project Website
- formulate in-context imitation learning as a graph generation problem
- demonstrate impressive one-shot learning performance
#ICLR2025 #manipulation #in_context
X Overview, Project Website
Vision Language Models are In-Context Value Learners
- formulating value learning as an autoregressive prediction task over shuffled sequence of the input video
- suitable across different tasks and in-context learning
#ICLR2025 #manipulation #in_context
X Overview, Project Website
- formulating value learning as an autoregressive prediction task over shuffled sequence of the input video
- suitable across different tasks and in-context learning
#ICLR2025 #manipulation #in_context
X Overview, Project Website
Taccel: Scaling Up Vision-based Tactile Robotics via High-performance GPU Simulation
- Combined ABD with IPC and implemented in Warp
- Good implementation of tactile sensor
#arxiv #tactile #warp
X Overview, Project Website
- Combined ABD with IPC and implemented in Warp
- Good implementation of tactile sensor
#arxiv #tactile #warp
X Overview, Project Website
Real-is-Sim: Bridging the Sim-to-Real Gap with a Dynamic Digital Twin for Real-World Robot Policy Evaluation
- Incorporate previous work embodied gaussians as a novel representation of the real world
- using gaussian splatting and warp to reconstruct the scene in simulation
#arxiv #simulation #reconstruction
X Overview, Project Website, Embodied Gaussians
- Incorporate previous work embodied gaussians as a novel representation of the real world
- using gaussian splatting and warp to reconstruct the scene in simulation
#arxiv #simulation #reconstruction
X Overview, Project Website, Embodied Gaussians
Efficient Dexterous Bimanual Manipulation Transfer via Residual Learning
- two stage training for in-hand dexterous manipulation
- using residual action to refine the manipulation
#CVPR2025 #dexterous
X Overview, Project Link
- two stage training for in-hand dexterous manipulation
- using residual action to refine the manipulation
#CVPR2025 #dexterous
X Overview, Project Link
PhysTwin: Physics-Informed Reconstruction and Simulation of Deformable Objects from Videos
- utilize scene reconstruction to create digital twin
- using warp as the simulation pipeline
#arxiv #deformable #reconstruction #simulation
X Overview, Project Link, Arxiv Link
- utilize scene reconstruction to create digital twin
- using warp as the simulation pipeline
#arxiv #deformable #reconstruction #simulation
X Overview, Project Link, Arxiv Link