4 citations · 4 across the 6 of their papers we have counts for
Showing 2024Show all
3 papers · 1 filter
cs.CV2024
Do Current Video LLMs Have Strong OCR Abilities? A Preliminary Study
Yulin Fei, Yuhui Gao, Xingyuan Xian +3
With the rise of multimodal large language models, accurately extracting and understanding textual information from video content, referred to as video based optical character reco…
cs.CR2024
Fed-AugMix: Balancing Privacy and Utility via Data Augmentation
Haoyang Li, Wei Chen, Xiaojin Zhang
Gradient leakage attacks pose a significant threat to the privacy guarantees of federated learning. While distortion-based protection mechanisms are commonly employed to mitigate t…
cs.CL2024★ 4 cited
RSL-SQL: Robust Schema Linking in Text-to-SQL Generation
Zhenbiao Cao, Yuanlei Zheng, Zhihao Fan +3
Text-to-SQL generation aims to translate natural language questions into SQL statements. In Text-to-SQL based on large language models, schema linking is a widely adopted strategy…