Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Nemotron-H: A Family of Accurate and Efficient Hybrid Mamba-Transformer Models
NVIDIA, :, Aaron Blakeman +198
As inference-time scaling becomes critical for enhanced reasoning capabilities, it is increasingly becoming important to build models that are efficient to infer. We introduce Nemo…
cs.CL2024
OMCAT: Omni Context Aware Transformer
Arushi Goel, Karan Sapra, Matthieu Le +3
Large Language Models (LLMs) have made significant strides in text generation and comprehension, with recent advancements extending into multimodal LLMs that integrate visual and a…