3 papers
cs.CL2025
Haystack Engineering: Context Engineering for Heterogeneous and Agentic Long-Context Evaluation
Mufei Li, Dongqi Fu, Limei Wang +10
Modern long-context large language models (LLMs) perform well on synthetic "needle-in-a-haystack" (NIAH) benchmarks, but such tests overlook how noisy contexts arise from biased re…
cs.LG2025
Struc-EMB: The Potential of Structure-Aware Encoding in Language Embeddings
Shikun Liu, Haoyu Wang, Mufei Li +1
Text embeddings from Large Language Models (LLMs) have become foundational for numerous applications. However, these models typically operate on raw text, overlooking the rich stru…
cs.LG2025
Model Generalization on Text Attribute Graphs: Principles with Large Language Models
Haoyu Wang, Shikun Liu, Rongzhe Wei +1
Large language models (LLMs) have recently been introduced to graph learning, aiming to extend their zero-shot generalization success to tasks where labeled graph data is scarce. A…