1 citations · 1 across the 2 of their papers we have counts for
6 papers
BatGPT-Chem: A Foundation Large Model For Retrosynthesis Prediction
Yifei Yang, Runhan Shi, Zuchao Li +4
Retrosynthesis analysis is pivotal yet challenging in drug discovery and organic chemistry. Despite the proliferation of computational tools over the past decade, AI-based systems…
SCANS: Mitigating the Exaggerated Safety for LLMs via Safety-Conscious Activation Steering
Zouying Cao, Yifei Yang, Hai Zhao
Safety alignment is indispensable for Large Language Models (LLMs) to defend threats from malicious instructions. However, recent researches reveal safety-aligned LLMs prone to rej…
Hypertext Entity Extraction in Webpage
Yifei Yang, Tianqiao Liu, Bo Shao +4
Webpage entity extraction is a fundamental natural language processing task in both research and applications. Nowadays, the majority of webpage entity extraction models are traine…
Head-wise Shareable Attention for Large Language Models
Zouying Cao, Yifei Yang, Hai Zhao
Large Language Models (LLMs) suffer from huge number of parameters, which restricts their deployment on edge devices. Weight sharing is one promising solution that encourages weigh…
LaCo: Large Language Model Pruning via Layer Collapse
Yifei Yang, Zouying Cao, Hai Zhao
Large language models (LLMs) based on transformer are witnessing a notable trend of size expansion, which brings considerable costs to both model training and inference. However, e…
AutoHall: Automated Factuality Hallucination Dataset Generation for Large Language Models
Zouying Cao, Yifei Yang, XiaoJing Li +1
Large language models (LLMs) have gained broad applications across various domains but still struggle with hallucinations. Currently, hallucinations occur frequently in the generat…