GestureCoach: Rehearsing for Engaging Talks with LLM-Driven Gesture Recommendations
arXiv:2504.10706 · doi:10.1145/3746059.3747705
Abstract
This paper introduces GestureCoach, a system designed to help speakers deliver more engaging talks by guiding them to gesture effectively during rehearsal. GestureCoach combines an LLM-driven gesture recommendation model with a rehearsal interface that proactively cues speakers to gesture appropriately. Trained on experts' gesturing patterns from TED talks, the model consists of two modules: an emphasis proposal module, which predicts when to gesture by identifying gesture-worthy text segments in the presenter notes, and a gesture identification module, which determines what gesture to use by retrieving semantically appropriate gestures from a curated gesture database. Results of a model performance evaluation and user study (N=30) show that the emphasis proposal module outperforms off-the-shelf LLMs in identifying suitable gesture regions, and that participants rated the majority of these predicted regions and their corresponding gestures as highly appropriate. A subsequent user study (N=10) showed that rehearsing with GestureCoach encouraged speakers to gesture and significantly increased gesture diversity, resulting in more engaging talks. We conclude with design implications for future AI-driven rehearsal systems.
Accepted at UIST 2025
References in corpus (14)
- Studying the effect of AI Code Generators on Supporting Novice Learners in Introductory Programming
- Gesticulator: A framework for semantically-aware speech-driven gesture generation
- VISAR: A Human-AI Argumentative Writing Assistant with Visual Programming and Rapid Draft Prototyping
- A Comprehensive Review of Data-Driven Co-Speech Gesture Generation
- The GENEA Challenge 2022: A large evaluation of data-driven co-speech gesture generation
- Spellburst: A Node-based Interface for Exploratory Creative Coding with Natural Language Prompts
- Memoro: Using Large Language Models to Realize a Concise Interface for Real-Time Memory Augmentation
- Semantic Gesticulator: Semantics-Aware Co-Speech Gesture Synthesis
- VoiceCoach: Interactive Evidence-based Training for Voice Modulation Skills in Public Speaking
- ChameleonControl: Teleoperating Real Human Surrogates through Mixed Reality Gestural Guidance for Remote Hands-on Classrooms
- DanceGen: Supporting Choreography Ideation and Prototyping with Generative AI
- GestureLens: Visual Analysis of Gestures in Presentation Videos
- avaTTAR: Table Tennis Stroke Training with On-body and Detached Visualization in Augmented Reality
- SHAPE-IT: Exploring Text-to-Shape-Display for Generative Shape-Changing Behaviors with LLMs