3 papers
cs.DC2026
WANSpec: Leveraging Global Compute Capacity for LLM Inference
Noah Martin, Fahad Dogar
Data centers capable of running large language models (LLMs) are spread across the globe. Some have high end GPUs for running the most advanced models (100B+ parameters), and other…
cs.DC2025
LLMBridge: Reducing Costs to Access LLMs in a Prompt-Centric Internet
Noah Martin, Abdullah Bin Faisal, Hiba Eltigani +3
Today's Internet infrastructure is centered around content retrieval over HTTP, with middleboxes (e.g., HTTP proxies) playing a crucial role in performance, security, and cost-effe…
cs.HC2025
WaLLM -- Insights from an LLM-Powered Chatbot deployment via WhatsApp
Hiba Eltigani, Rukhshan Haroon, Asli Kocak +3
Recent advances in generative AI, such as ChatGPT, have transformed access to information in education, knowledge-seeking, and everyday decision-making. However, in many developing…