7 papers
Harmonised benchmarking of foundation models for single-cell and spatial transcriptomics reveals context-dependent generalisation
Sally Chen, Roxana Zahedi, Lucy Chhuo +11
Single-cell and spatial foundation models promise transferable biological representations, yet their generality remains largely untested across modalities, biological domains and a…
Structuring Reasoning for Complex Rules Beyond Flat Representations
Zhihao Yang, Ancheng Xu, Jingpeng Li +11
Large language models (LLMs) face significant challenges when processing complex rule systems, as they typically treat interdependent rules as unstructured textual data rather than…
Expanding before Inferring: Enhancing Factuality in Large Language Models through Premature Layers Interpolation
Dingwei Chen, Ziqiang Liu, Feiteng Fang +6
Large Language Models (LLMs) demonstrate remarkable capabilities in text understanding and generation. However, their tendency to produce factually inconsistent outputs, commonly r…
RxSafeBench: Identifying Medication Safety Issues of Large Language Models in Simulated Consultation
Jiahao Zhao, Luxin Xu, Minghuan Tan +4
Numerous medical systems powered by Large Language Models (LLMs) have achieved remarkable progress in diverse healthcare tasks. However, research on their medication safety remains…
Breaking the Block: Preserving Data Continuity to Train Superior SAEs for Instruct Models
Jiaming Li, Haoran Ye, Yukun Chen +5
Sparse Autoencoders (SAEs) are a cornerstone of mechanistic interpretability. Existing training methods inherit the Block Training paradigm from LLM pre-training, which introduces…
STORYTELLER: An Enhanced Plot-Planning Framework for Coherent and Cohesive Story Generation
Jiaming Li, Yukun Chen, Ziqiang Liu +10
Stories are central to human culture, serving to share ideas, preserve traditions, and foster connections. Automatic story generation, a key advancement in artificial intelligence…