Closer to Human: Hybrid Training of VR Agents Beyond Robotic Motion

1Citations
Citations of this article
9Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

The success of virtual reality (VR) agents is not solely defined by task completion; adaptability and the ability to move beyond repetitive, robotic behavior are equally critical for embodiment, social interaction, rehabilitation, and training. To address this challenge, we propose a hybrid training methodology that integrates imitation learning and reinforcement learning to develop agents that both master tasks and generalize beyond demonstrations. Our framework combines Behavioral Cloning (BC) for rapid policy initialization, Proximal Policy Optimization (PPO) for stable reinforcement-driven learning, and Generative Adversarial Imitation Learning (GAIL) for imitation-based rewards, resulting in policies that converge efficiently while avoiding repetitive imitation. We implemented this framework in a Unity-based VR environment, where participants performed a single-joint cube pick-and-place task using a headset and controller. Human demonstrations were collected and used to train and evaluate agents against multiple motion metrics, including Dynamic Time Warping (DTW), Fréchet distance, minimum jerk cost, Fitts’s law, and curvature–velocity coupling. Experimental results showed that hybrid-trained agents achieved stable convergence, consistent task mastery, and motion trajectories comparable to human demonstrations across temporal, spatial, and smoothness dimensions. Notably, the agents exhibited adaptive, non-repetitive behavior without explicitly optimizing for biomechanical fidelity, suggesting that human-like variability can emerge naturally from our customized hybrid IL + RL training. These findings provide a foundation for scaling the methodology toward multi-joint and full-body motion learning, with promising implications for future VR applications in rehabilitation, training, and embodied interaction.

Cite

CITATION STYLE

APA

Sobhi, M., Moosavi, M. S., Guillet, C., Meriénne, F., & Ebadzadeh, M. M. (2026). Closer to Human: Hybrid Training of VR Agents Beyond Robotic Motion. IEEE Access, 14, 18940–18951. https://doi.org/10.1109/ACCESS.2026.3658497

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free