3 papers
cs.CV2026
HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds
Team HY-World, Chenjie Cao, Xuhui Zuo +42
We introduce HY-World 2.0, a multi-modal world model framework that advances our prior project HY-World 1.0. HY-World 2.0 accommodates diverse input modalities, including text prom…
eess.AS2024
Device-Directed Speech Detection for Follow-up Conversations Using Large Language Models
Ognjen, Rudovic, Pranay Dighe +7
Follow-up conversations with virtual assistants (VAs) enable a user to seamlessly interact with a VA without the need to repeatedly invoke it using a keyword (after the first query…
cs.AI2024
KGLens: Towards Efficient and Effective Knowledge Probing of Large Language Models with Knowledge Graphs
Shangshang Zheng, He Bai, Yizhe Zhang +3
Large Language Models (LLMs) might hallucinate facts, while curated Knowledge Graph (KGs) are typically factually reliable especially with domain-specific knowledge. Measuring the…