2 papers
cs.LG2026
Arcee Trinity Large Technical Report
Varun Singh, Lucas Krauss, Sami Jaghouar +23
We present the technical report for Arcee Trinity Large, a sparse Mixture-of-Experts model with 400B total parameters and 13B activated per token. Additionally, we report on Trinit…
cs.CL2025
Fine-tuning Language Models for Recipe Generation: A Comparative Analysis and Benchmark Study
Anneketh Vij, Changhao Liu, Rahul Anil Nair +3
This research presents an exploration and study of the recipe generation task by fine-tuning various very small language models, with a focus on developing robust evaluation metric…