From the 1 of 7 linked papers with an AI index.
7 papers
Progressive Cramming: Reliable Token Compression and What It Reveals
Dmitrii Tarasov, Timofei Lashukov, Elizaveta Goncharova +1
Token cramming compresses sequences into learned embeddings with near-perfect reconstruction, but fixed token budgets and 99\% accuracy thresholds leave it unclear whether residual…
Towards Robust Speech Deepfake Detection via Human-Inspired Reasoning
Artem Dvirniak, Evgeny Kushnir, Dmitrii Tarasov +5
The paper introduces HIR‑SDD, a speech deepfake detection framework that leverages large audio language models and chain‑of‑thought reasoning from a human‑annotated dataset to impr…
OCC-RAG: Optimal Cognitive Core for Faithful Question Answering
Maksim Savkin, Mikhail Goncharov, Alexander Gambashidze +7
Recent progress in the development of language models has been defined by scale, with each generation absorbing more of the world's knowledge into its weights. However, many practi…
Speech-to-LaTeX: New Models and Datasets for Converting Spoken Equations and Sentences
Dmitrii Korzh, Dmitrii Tarasov, Artyom Iudin +6
Conversion of spoken mathematical expressions is a challenging task that involves transcribing speech into a strictly structured symbolic representation while addressing the ambigu…
Sentence-Anchored Gist Compression for Long-Context LLMs
Dmitrii Tarasov, Elizaveta Goncharova, Kuznetsov Andrey
This work investigates context compression for Large Language Models (LLMs) using learned compression tokens to reduce the memory and computational demands of processing long seque…
Image Reconstruction as a Tool for Feature Analysis
Eduard Allakhverdov, Dmitrii Tarasov, Elizaveta Goncharova +1
Vision encoders are increasingly used in modern applications, from vision-only models to multimodal systems such as vision-language models. Despite their remarkable success, it rem…