ResearchResearch paperRobotics & Embodied AI1 source · Oct 7, 2026

From Digital Human Interactions to Physics-Based Humanoid Skills: Physics-Grounded Post-Training of Interaction Generators

In this paper, we introduce DIGHT, a co-adaptive framework that couples a Digital human Interaction Generator with a Humanoid Tracking policy.

Key points

  • Recent methods have made promising progress in generating interactions between two humanoids, largely relying on physics-based tracking policies to convert digital reference motions into executable trajectories.
  • Our DIGHT first executes multiple text-conditioned interaction candidates in simulation using a fixed tracker.
  • Rather than collapsing these signals into a single scalar reward for candidate ranking, we align the pretrained generator using physics-decoupled diffusion direct preference optimization (DPO), preserving criterion-specific supervision without differentiating through the simulator.
  • Additionally, to improve interaction fidelity, we propose to incorporate force feedback from simulator as a measure of contact fidelity and construct preferences over contact occurrence, location, duration, and force magnitude.

Sources (1)

Extractive summary: sentences quoted from the sources.

Before this

  1. Oct 7, 2026How to train your model organism
  2. Oct 7, 2026Iris-3B: Going Beyond the Latent with Pixel-Space Diffusion Training, Conversion and Fine-Tuning
  3. Oct 7, 2026Q-Learning with Scalar Adjoint Matching
  4. Oct 6, 2026AutodidactWAM: Cross-Modal Self-Distillation from Generated Video to Robot Actions
  5. Oct 6, 2026Frozen Models, Evolving Expertise: Model-Agnostic Learning from Deployment Experience for Multimodal Medical AI
  6. Sep 30, 2026Expanding AI Storage Access with NVIDIA cuObject and the NVIDIA SCADA Server SDK

Related