1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CL2026
Towards Compositional Generalization of LLMs via Skill Taxonomy Guided Data Synthesis
Yifan Wei, Li Du, Xiaoyan Yu +2
Large Language Models (LLMs) and agent-based systems often struggle with compositional generalization due to a data bottleneck in which complex skill combinations follow a long-tai…
stat.ML2025
GeoERM: Geometry-Aware Multi-Task Representation Learning on Riemannian Manifolds
Aoran Chen, Yang Feng
Multi-Task Learning (MTL) seeks to boost statistical power and learning efficiency by discovering structure shared across related tasks. State-of-the-art MTL representation methods…
cs.LG2024★ 1 cited
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration
Zhuofan Wen, Shangtong Gui, Yang Feng
Inference acceleration of large language models (LLMs) has been put forward in many application scenarios and speculative decoding has shown its advantage in addressing inference a…