NewEvery arXiv paper, its researchers & institutions — mapped.
papers

Publications (15)

cs.IR2025

Meta Lattice: Model Space Redesign for Cost-Effective Industry-Scale Ads Recommendations

Liang Luo, Yuxin Chen, Zhengyu Zhang +39

cs.AR2018

Best-Effort FPGA Programming: A Few Steps Can Go a Long Way

Jason Cong, Zhenman Fang, Yuchen Hao +4

cs.LG2024

Accelerating Communication in Deep Learning Recommendation Model Training with Dual-Level Adaptive Lossy Compression

Hao Feng, Boyuan Zhang, Fanjiang Ye +9

cs.LG2024

Disaggregated Multi-Tower: Topology-aware Modeling Technique for Efficient Large-Scale Recommendation

Liang Luo, Buyun Zhang, Michael Tsang +11

cs.DC2023

Software-Hardware Co-design for Fast and Scalable Training of Deep Learning Recommendation Models

Dheevatsa Mudigere, Yuchen Hao, Jianyu Huang +50

cs.LG2026

ROCS: Request-Oriented Compute Sharing for Efficient Large-Scale Recommendation

Yuxin Chen, Liang Luo, Buyun Zhang +44

The paper introduces ROCS, a request-oriented compute sharing framework that restructures recommendation inference to evaluate shared request features once per request rather than…

#recommendation#inference optimization#compute sharing#large-scale systems
cs.LG2026

LoKA: Low-precision Kernel Applications for Recommendation Models At Scale

Liang Luo, Yinbin Ma, Quanyu Zhu +21

cs.IR2022

DHEN: A Deep and Hierarchical Ensemble Network for Large-Scale Click-Through Rate Prediction

Buyun Zhang, Liang Luo, Xi Liu +14

cs.LG2024

Wukong: Towards a Scaling Law for Large-Scale Recommendation

Buyun Zhang, Liang Luo, Yuxin Chen +12

cs.AI2024

The Llama 3 Herd of Models

Aaron Grattafiori, Abhimanyu Dubey, Abhinav Jauhri +556

cs.IR2025

External Large Foundation Model: How to Efficiently Serve Trillions of Parameters for Online Ads Recommendation

Mingfu Liang, Xi Liu, Rong Jin +104

cs.DC2023

PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel

Yanli Zhao, Andrew Gu, Rohan Varma +15

cs.DC2026

ScaleAcross Explorer: Exploring Communication Optimization for Scale-Across AI Model Training

Minghao Li, Alicia Golden, Samuel Hsia +14

cs.DC2025

WLB-LLM: Workload-Balanced 4D Parallelism for Large Language Model Training

Zheng Wang, Anna Cai, Xinfeng Xie +9

cs.LG2023

Rankitect: Ranking Architecture Search Battling World-class Engineers at Meta Scale

Wei Wen, Kuang-Hung Liu, Igor Fedorov +19