4 citations · 10 across the 8 of their papers we have counts for
8 papers
The Interspeech 2024 Challenge on Speech Processing Using Discrete Units
Xuankai Chang, Jiatong Shi, Jinchuan Tian +7
Representing speech and audio signals in discrete units has become a compelling alternative to traditional high-dimensional feature vectors. Numerous studies have highlighted the e…
Your Vision-Language Model Itself Is a Strong Filter: Towards High-Quality Instruction Tuning with Data Selection
Ruibo Chen, Yihan Wu, Lichang Chen +6
Data selection in instruction tuning emerges as a pivotal process for acquiring high-quality data and training instruction-following large language models (LLMs), but it is still a…
SpeechComposer: Unifying Multiple Speech Tasks with Prompt Composition
Yihan Wu, Soumi Maiti, Yifan Peng +6
Recent advancements in language models have significantly enhanced performance in multiple speech-related tasks. Existing speech language models typically utilize task-dependent pr…
GPT-4 Vision on Medical Image Classification -- A Case Study on COVID-19 Dataset
Ruibo Chen, Tianyi Xiong, Yihan Wu +6
This technical report delves into the application of GPT-4 Vision (GPT-4V) in the nuanced realm of COVID-19 image classification, leveraging the transformative potential of in-cont…
Unbiased Watermark for Large Language Models
Zhengmian Hu, Lichang Chen, Xidong Wu +3
The recent advancements in large language models (LLMs) have sparked a growing apprehension regarding the potential misuse. One approach to mitigating this risk is to incorporate w…
Shielding the Unseen: Privacy Protection through Poisoning NeRF with Spatial Deformation
Yihan Wu, Brandon Y. Feng, Heng Huang
In this paper, we introduce an innovative method of safeguarding user privacy against the generative capabilities of Neural Radiance Fields (NeRF) models. Our novel poisoning attac…