From the 2 of 31 linked papers with an AI index.
10 papers · 1 filter
PolyQ: Codesigning End-to-End Quantization Framework for Scalable Edge CPU LLM Inference
Hyunwoo Oh, Suyeon Jang, Hanning Chen +4
PolyQ is a co-designed compiler and quantization framework that assigns per‑channel bit‑widths to LLM activations on CPUs, enabling fine‑grained fractional‑bit precision while keep…
-Musketeers: Reinforcement Learning Shapes Collaboration Among Language Models
Ryozo Masukawa, Sanggeon Yun, Hyunwoo Oh +8
Recent progress in reinforcement learning with verifiable rewards (RLVR) shows that small, specialized language models (SLMs) can exhibit structured reasoning without relying on la…
Internal Flow Signatures for Self-Checking and Refinement in LLMs
Sungheon Jeong, Sanggeon Yun, Ryozo Masukawa +3
Large language models can generate fluent answers that are unfaithful to the provided context, while many safeguards rely on external verification or a separate judge after generat…
Encoder-Free Knowledge-Graph Reasoning with LLMs via Hyperdimensional Path Retrieval
Yezi Liu, William Youngwoo Chung, Hanning Chen +2
Recent progress in large language models (LLMs) has made knowledge-grounded reasoning increasingly practical, yet KG-based QA systems often pay a steep price in efficiency and tran…
Cauchy-Schwarz Fairness Regularizer
Yezi Liu, Hanning Chen, Wenjun Huang +2
Group fairness in machine learning is often enforced by adding a regularizer that reduces the dependence between model predictions and sensitive attributes. However, existing regul…
Mitigating Bias in Graph Hyperdimensional Computing
Yezi Liu, William Youngwoo Chung, Yang Ni +2
Graph hyperdimensional computing (HDC) has emerged as a promising paradigm for cognitive tasks, emulating brain-like computation with high-dimensional vectors known as hypervectors…