activity
20242026
collaborators

9 papers

cs.AI2026

Bridging the Information Gap: Semantic Densification and Hindsight Distillation for Cold-Start Prediction

Hao Duong Le, Yifei Gao, Huan Li +4

New-user cold-start is a critical bottleneck for e-commerce platforms: predicting user lifetime value (LTV) and conversion rate (CVR) for users with sparse interaction history. Two…

cs.CL2026

TokenButler: Token Importance is Predictable

Yash Akhauri, Ahmed F AbouElhamayed, Yifei Gao +4

Large Language Models (LLMs) rely on the Key-Value (KV) Cache to store token history, enabling efficient decoding of tokens. As the KV-Cache grows, it becomes a major memory and co…

cs.LG2026

Near-Oracle KV Selection via Pre-hoc Sparsity for Long-Context Inference

Yifei Gao, Lei Wang, Rong-Cheng Tu +3

A core bottleneck in large language model (LLM) inference is the cost of attending over the ever-growing key-value (KV) cache. Although near-oracle top-k KV selection can preserve…

cs.HC2025

Evaluating Trust in AI, Human, and Co-produced Feedback Among Undergraduate Students

Audrey Zhang, Yifei Gao, Wannapon Suraworachet +2

As generative AI models, particularly large language models (LLMs), transform educational feedback practices in higher education (HE) contexts, understanding students' perceptions…

cs.CL2025

Compensate Quantization Errors+: Quantized Models Are Inquisitive Learners

Yifei Gao, Jie Ou, Lei Wang +2

The quantization of large language models (LLMs) has been a prominent research area aimed at enabling their lightweight deployment in practice. Existing research about LLM's quanti…

cs.GR2025

Beyond Existance: Fulfill 3D Reconstructed Scenes with Pseudo Details

Yifei Gao, Jun Huang, Lei Wang +2

The emergence of 3D Gaussian Splatting (3D-GS) has significantly advanced 3D reconstruction by providing high fidelity and fast training speeds across various scenarios. While rece…