6 citations · 6 across the 4 of their papers we have counts for
6 papers
DACA-GRPO: Denoising-Aware Credit Assignment for Reinforcement Learning in Diffusion Language Models
Amin Karimi Monsefi, Dominic Culver, Nikhil Bhendawade +4
Diffusion large language models are a compelling alternative to autoregressive models, yet existing RL methods for diffusion treat all denoising steps as equally important and rely…
Trajectory as the Teacher: Few-Step Discrete Flow Matching via Energy-Navigated Distillation
Amin Karimi Monsefi, Dominic Culver, Nikhil Bhendawade +3
Discrete flow matching generates text by iteratively transforming noise tokens into coherent language, but may require hundreds of forward passes. Distillation uses the multi-step…
FS-DFM: Fast and Accurate Long Text Generation with Few-Step Diffusion Language Models
Amin Karimi Monsefi, Nikhil Bhendawade, Manuel Rafael Ciosici +3
Autoregressive language models (ARMs) deliver strong likelihoods, but are inherently serial: they generate one token per forward pass, which limits throughput and inflates latency…
Remember what you did so you know what to do next
Manuel R. Ciosici, Alex Hedges, Yash Kankanampati +3
We explore using a moderately sized large language model (GPT-J 6B parameters) to create a plan for a simulated robot to achieve 30 classes of goals in ScienceWorld, a text game si…
A reproduction of Apple's bi-directional LSTM models for language identification in short strings
Mads Toftrup, Søren Asger Sørensen, Manuel R. Ciosici +1
Language Identification is the task of identifying a document's language. For applications like automatic spell checker selection, language identification must use very short strin…
Unsupervised Abbreviation Disambiguation Contextual disambiguation using word embeddings
Manuel Ciosici, Tobias Sommer, Ira Assent
Abbreviations often have several distinct meanings, often making their use in text ambiguous. Expanding them to their intended meaning in context is important for Machine Reading t…