activity
20232025
most citedUni-Med: A Unified Medical Generalist Foundation Model For Multi-Task Learning Via Connector-MoE

3 citations · 6 across the 5 of their papers we have counts for

collaborators

5 papers

cs.SD2025

Step-Audio-AQAA: a Fully End-to-End Expressive Large Audio Language Model

Ailin Huang, Bingxin Li, Bruce Wang +73

Large Audio-Language Models (LALMs) have significantly advanced intelligent human-computer interaction, yet their reliance on text-based outputs limits their ability to generate na…

cs.RO2024

Learning Robust Grasping Strategy Through Tactile Sensing and Adaption Skill

Yueming Hu, Mengde Li, Songhua Yang +3

Robust grasping represents an essential task in robotics, necessitating tactile feedback and reactive grasping adjustments for robust grasping of objects. Previous research has ext…

cs.CV20243 cited

Uni-Med: A Unified Medical Generalist Foundation Model For Multi-Task Learning Via Connector-MoE

Xun Zhu, Ying Hu, Fanbin Mo +2

Multi-modal large language models (MLLMs) have shown impressive capabilities as a general-purpose interface for various visual and linguistic tasks. However, building a unified MLL…

cs.CL20232 cited

THiFLY Research at SemEval-2023 Task 7: A Multi-granularity System for CTR-based Textual Entailment and Evidence Retrieval

Yuxuan Zhou, Ziyu Jin, Meiwei Li +4

The NLI4CT task aims to entail hypotheses based on Clinical Trial Reports (CTRs) and retrieve the corresponding evidence supporting the justification. This task poses a significant…

cs.CL20231 cited

Compressed Heterogeneous Graph for Abstractive Multi-Document Summarization

Miao Li, Jianzhong Qi, Jey Han Lau

Multi-document summarization (MDS) aims to generate a summary for a number of related documents. We propose HGSUM, an MDS model that extends an encoder-decoder architecture, to inc…