Showing cs.DCShow all
2 papers · 1 filter
cs.DC2026
Data Driven Optimization of GPU efficiency for Distributed LLM-Adapter Serving
Ferran Agullo, Joan Oliveras, Chen Wang +5
Large Language Model (LLM) adapters enable low-cost model specialization, but introduce complex caching and scheduling challenges in distributed serving systems where hundreds of a…
cs.DC2024
Enabling an OpenStack-based cloud on top of RISC-V hardware
Diego Marrón, Aaron Call, Josep Ll. Berral +1
The European Union's technological sovereignty strategy centers around the RISC-V Instruction Set Architecture, with the European Processor Initiative leading efforts to build prod…