1 citations · 2 across the 5 of their papers we have counts for
5 papers
A Multi-Level Framework for Accelerating Training Transformer Models
Longwei Zou, Han Zhang, Yangdong Deng
The fast growing capabilities of large-scale deep learning models, such as Bert, GPT and ViT, are revolutionizing the landscape of NLP, CV and many other domains. Training such mod…
Memory-based Cross-modal Semantic Alignment Network for Radiology Report Generation
Yitian Tao, Liyan Ma, Jing Yu +1
Generating radiology reports automatically reduces the workload of radiologists and helps the diagnoses of specific diseases. Many existing methods take this task as modality trans…
Uncertainty-Penalized Reinforcement Learning from Human Feedback with Diverse Reward LoRA Ensembles
Yuanzhao Zhai, Han Zhang, Yu Lei +5
Reinforcement learning from human feedback (RLHF) emerges as a promising paradigm for aligning large language models (LLMs). However, a notable challenge in RLHF is overoptimizatio…
Chat2Brain: A Method for Mapping Open-Ended Semantic Queries to Brain Activation Maps
Yaonai Wei, Tuo Zhang, Han Zhang +10
Over decades, neuroscience has accumulated a wealth of research results in the text modality that can be used to explore cognitive processes. Meta-analysis is a typical method that…
Pre-training Tasks for User Intent Detection and Embedding Retrieval in E-commerce Search
Yiming Qiu, Chenyu Zhao, Han Zhang +7
BERT-style models pre-trained on the general corpus (e.g., Wikipedia) and fine-tuned on specific task corpus, have recently emerged as breakthrough techniques in many NLP tasks: qu…