1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.PF2025
Meta-Metrics and Best Practices for System-Level Inference Performance Benchmarking
Shweta Salaria, Zhuoran Liu, Nelson Mimura Gonzalez
Benchmarking inference performance (speed) of Foundation Models such as Large Language Models (LLM) involves navigating a vast experimental landscape to understand the complex inte…
cs.DC2024★ 1 cited
The infrastructure powering IBM's Gen AI model development
Talia Gershon, Seetharami Seelam, Brian Belgodere +143
AI Infrastructure plays a key role in the speed and cost-competitiveness of developing and deploying advanced AI models. The current demand for powerful AI infrastructure for model…