most citedMEWL: Few-shot multimodal word learning with referential uncertainty

5 citations · 7 across the 6 of their papers we have counts for

collaborators

11 papers

cs.CV2024

AnySkill: Learning Open-Vocabulary Physical Skill for Interactive Agents

Jieming Cui, Tengyu Liu, Nian Liu +3

Traditional approaches in physics-based motion generation, centered around imitation learning and reward shaping, often struggle to adapt to new scenarios. To tackle this limitatio…

cs.AI20233 cited

Active Reasoning in an Open-World Environment

Manjie Xu, Guangyuan Jiang, Wei Liang +2

Recent advances in vision-language learning have achieved notable success on complete-information question-answering datasets through the integration of extensive world knowledge.…

cs.CV20232 cited

ProBio: A Protocol-guided Multimodal Dataset for Molecular Biology Lab

Jieming Cui, Ziren Gong, Baoxiong Jia +4

The challenge of replicating research results has posed a significant impediment to the field of molecular biology. The advent of modern intelligent systems has led to notable prog…

cs.CV20231 cited

Single-view 3D Scene Reconstruction with High-fidelity Shape and Texture

Yixin Chen, Junfeng Ni, Nan Jiang +3

Reconstructing detailed 3D scenes from single-view images remains a challenging task due to limitations in existing approaches, which primarily focus on geometric shape recovery, o…

cs.CV20236 cited

ChimpACT: A Longitudinal Dataset for Understanding Chimpanzee Behaviors

Xiaoxuan Ma, Stephan P. Kaufhold, Jiajun Su +6

Understanding the behavior of non-human primates is crucial for improving animal welfare, modeling social behavior, and gaining insights into distinctively human and phylogenetical…

cs.AI2023

X-VoE: Measuring eXplanatory Violation of Expectation in Physical Events

Bo Dai, Linge Wang, Baoxiong Jia +4

Intuitive physics is pivotal for human understanding of the physical world, enabling prediction and interpretation of events even in infancy. Nonetheless, replicating this level of…