Human-Object Interaction via Automatically Designed VLM-Guided Motion Policy

Human-object interaction (HOI) synthesis is crucial for applications in animation, simulation, and robotics. However, existing approaches either rely on expensive motion capture data or require manual reward engineering, limiting their scalability and generalizability. In this work, we introduce the first unified physics-based HOI framework that leverages Vision-Language Models (VLMs) to enable…

Paper

Similar papers

© 2026 NYSGPT2525 LLC