3 papers
cs.SD2026
Whole-Piece Training for Symbolic Music Language Models via Full-Horizon Compressed Recurrence
Yungang Yi, Weihua Li, Matthew Kuo +2
For computational efficiency, modern language models are typically trained on independently sampled fixed-length sequences. Symbolic music language models largely inherit this para…
cs.CV2025
ESCA: Contextualizing Embodied Agents via Scene-Graph Generation
Jiani Huang, Amish Sethi, Matthew Kuo +6
Multi-modal large language models (MLLMs) are making rapid progress toward general-purpose embodied agents. However, existing MLLMs do not reliably capture fine-grained links betwe…
cs.AI2025
PerceiverS: A Multi-Scale Perceiver with Effective Segmentation for Long-Term Expressive Symbolic Music Generation
Yungang Yi, Weihua Li, Matthew Kuo +1
AI-based music generation has made significant progress in recent years. However, generating symbolic music that is both long-structured and expressive remains a significant challe…