5 papers
PolyLingua: Margin-based Inter-class Transformer for Robust Cross-domain Language Detection
Ali Lotfi Rezaabad, Bikram Khanal, Shashwat Chaurasia +5
Language identification is a crucial first step in multilingual systems such as chatbots and virtual assistants, enabling linguistically and culturally accurate user experiences. E…
MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention
MiniMax, :, Aili Chen +125
We introduce MiniMax-M1, the world's first open-weight, large-scale hybrid-attention reasoning model. MiniMax-M1 is powered by a hybrid Mixture-of-Experts (MoE) architecture combin…
The Amazon Nova Family of Models: Technical Report and Model Card
Amazon AGI, Aaron Langford, Aayush Shah +783
We present Amazon Nova, a new generation of state-of-the-art foundation models that deliver frontier intelligence and industry-leading price performance. Amazon Nova Pro is a highl…
Uni-Renderer: Unifying Rendering and Inverse Rendering Via Dual Stream Diffusion
Zhifei Chen, Tianshuo Xu, Wenhang Ge +7
Rendering and inverse rendering are pivotal tasks in both computer vision and graphics. The rendering equation is the core of the two tasks, as an ideal conditional distribution tr…
CALICO: Conversational Agent Localization via Synthetic Data Generation
Andy Rosenbaum, Pegah Kharazmi, Ershad Banijamali +8
We present CALICO, a method to fine-tune Large Language Models (LLMs) to localize conversational agent training data from one language to another. For slots (named entities), CALIC…