From the 1 of 7 linked papers with an AI index.
7 papers
LAST: The Last Query Token Guides Visual Token Pruning for Edge-Cloud Collaborative MLLM Inference
Feng Yang, Xinrui Ju, Keyang Zhang +6
The paper introduces LAST, a training‑free method that uses the attention of the last query token to prune visual tokens on edge devices before sending them to a cloud multimodal L…
LUMI: Tokenizer-Agnostic LLM-Based Lossless Image Compression
Chris Xing Tian, Chengkai Wu, Ziyu Wang +6
Large language model (LLM)-based lossless image compression methods typically represent pixel data through the native text interface of a pretrained model, converting pixel values…
Domain-Specific Data Generation Framework for RAG Adaptation
Chris Xing Tian, Weihao Xie, Zhen Chen +5
Retrieval-Augmented Generation (RAG) combines the language understanding and reasoning power of large language models (LLMs) with external retrieval to enable domain-grounded respo…
Q-PART: Quasi-Periodic Adaptive Regression with Test-time Training for Pediatric Left Ventricular Ejection Fraction Regression
Jie Liu, Tiexin Qin, Hui Liu +5
In this work, we address the challenge of adaptive pediatric Left Ventricular Ejection Fraction (LVEF) assessment. While Test-time Training (TTT) approaches show promise for this t…
Large Language Models for Lossless Image Compression: Next-Pixel Prediction in Language Space is All You Need
Kecheng Chen, Pingping Zhang, Hui Liu +6
We have recently witnessed that ``Intelligence" and `` Compression" are the two sides of the same coin, where the language large model (LLM) with unprecedented intelligence is a ge…
Pixel-Inconsistency Modeling for Image Manipulation Localization
Chenqi Kong, Anwei Luo, Shiqi Wang +3
Digital image forensics plays a crucial role in image authentication and manipulation localization. Despite the progress powered by deep neural networks, existing forgery localizat…