2 papers
cs.CL2026
GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling
Alireza Dadgarnia, Soroush Tabesh, Mahdi Nikdan +4
Quantization has become a standard tool for efficient LLM deployment, especially for local inference, where models are now routinely served at 2-3 bits per parameter. The state of…
cs.CV2025
LRW-Persian: Lip-reading in the Wild Dataset for Persian Language
Zahra Taghizadeh, Mohammad Shahverdikondori, Arian Noori +1
Lipreading has emerged as an increasingly important research area for developing robust speech recognition systems and assistive technologies for the hearing-impaired. However, non…