VTutor: An Open-Source SDK for Generative AI-Powered Animated Pedagogical Agents with Multi-Media Output
Eason Chen, Chenyu Lin, Xinyi Tang, Aprille Xi, Canwen Wang, Jionghao, Lin, Kenneth R Koedinger

TL;DR
VTutor is an open-source SDK that integrates generative AI with advanced animation to create realistic, multi-media pedagogical agents for engaging and adaptive human-AI interactions in educational contexts.
Contribution
It introduces a comprehensive toolkit combining LLMs, animation, and web technologies to develop emotionally resonant, context-aware animated pedagogical agents for the first time.
Findings
Enhances learner engagement and feedback receptivity.
Supports real-time personalized feedback and natural speech synchronization.
Provides a scalable, open-source platform for research and development.
Abstract
The rapid evolution of large language models (LLMs) has transformed human-computer interaction (HCI), but the interaction with LLMs is currently mainly focused on text-based interactions, while other multi-model approaches remain under-explored. This paper introduces VTutor, an open-source Software Development Kit (SDK) that combines generative AI with advanced animation technologies to create engaging, adaptable, and realistic APAs for human-AI multi-media interactions. VTutor leverages LLMs for real-time personalized feedback, advanced lip synchronization for natural speech alignment, and WebGL rendering for seamless web integration. Supporting various 2D and 3D character models, VTutor enables researchers and developers to design emotionally resonant, contextually adaptive learning agents. This toolkit enhances learner engagement, feedback receptivity, and human-AI interaction while…
Peer Reviews
No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.
Videos
No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.
Taxonomy
TopicsHuman Motion and Animation
