1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.SE2026★ 1 cited
The Llama 4 Herd: Architecture, Training, Evaluation, and Deployment Notes
Redacted by arXiv
This document consolidates publicly reported technical details about Metas Llama 4 model family. It summarizes (i) released variants (Scout and Maverick) and the broader herd conte…
cs.CL2024
Confidence Preservation Property in Knowledge Distillation Abstractions
Dmitry Vengertsev, Elena Sherman
Social media platforms prevent malicious activities by detecting harmful content of posts and comments. To that end, they employ large-scale deep neural network language models for…