most citedThe Llama 4 Herd: Architecture, Training, Evaluation, and Deployment Notes

1 citations · 1 across the 6 of their papers we have counts for

collaborators

6 papers

cs.SE20261 cited

The Llama 4 Herd: Architecture, Training, Evaluation, and Deployment Notes

Redacted by arXiv

This document consolidates publicly reported technical details about Metas Llama 4 model family. It summarizes (i) released variants (Scout and Maverick) and the broader herd conte…

cs.LG2025

Multi-Value Alignment for LLMs via Value Decorrelation and Extrapolation

Hefei Xu, Le Wu, Chen Cheng +1

With the rapid advancement of large language models (LLMs), aligning them with human values for safety and ethics has become a critical challenge. This problem is especially challe…

cs.AI2025

Towards Foundation Model on Temporal Knowledge Graph Reasoning

Jiaxin Pan, Mojtaba Nayyeri, Osama Mohammed +4

Temporal Knowledge Graphs (TKGs) store temporal facts with quadruple formats (s, p, o, t). Existing Temporal Knowledge Graph Embedding (TKGE) models perform link prediction tasks i…

cs.LG2025

Skywork Open Reasoner 1 Technical Report

Jujie He, Jiacai Liu, Chris Yuhao Liu +14

The success of DeepSeek-R1 underscores the significant role of reinforcement learning (RL) in enhancing the reasoning capabilities of large language models (LLMs). In this work, we…

cs.CL2025

AD-AGENT: A Multi-agent Framework for End-to-end Anomaly Detection

Tiankai Yang, Junjun Liu, Wingchun Siu +6

Anomaly detection (AD) is essential in areas such as fraud detection, network monitoring, and scientific research. However, the diversity of data modalities and the increasing numb…

cs.CL2025

A Comparative Study of Large Language Models and Human Personality Traits

Wang Jiaqi, Wang bo, Guo fa +2

Large Language Models (LLMs) have demonstrated human-like capabilities in language comprehension and generation, becoming active participants in social and cognitive domains. This…