Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
Multilingual Pretraining and Instruction Tuning Improve Cross-Lingual Knowledge Alignment, But Only Shallowly
Changjiang Gao, Hongda Hu, Peng Hu +3
Despite their strong ability to retrieve knowledge in English, current large language models show imbalance abilities in different languages. Two approaches are proposed to address…
cs.CL2023
Roles of Scaling and Instruction Tuning in Language Perception: Model vs. Human Attention
Changjiang Gao, Shujian Huang, Jixing Li +1
Recent large language models (LLMs) have revealed strong abilities to understand natural language. Since most of them share the same basic structure, i.e. the transformer block, po…