Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Mercury: Ultra-Fast Language Models Based on Diffusion
Inception Labs, Samar Khanna, Siddhant Kharbanda +10
We present Mercury, a new generation of commercial-scale large language models (LLMs) based on diffusion. These models are parameterized via the Transformer architecture and traine…
cs.CL2024
Large Language Models are Geographically Biased
Rohin Manvi, Samar Khanna, Marshall Burke +2
Large Language Models (LLMs) inherently carry the biases contained in their training corpora, which can lead to the perpetuation of societal harm. As the impact of these foundation…