6 papers
Vortex: Hosting ML Inference and Knowledge Retrieval Services With Tight Latency and Throughput Requirements
Yuting Yang, Tiancheng Yuan, Jamal Hashim +6
There is growing interest in deploying ML inference and knowledge retrieval as services that could support both interactive queries by end users and more demanding request flows th…
zkToken: Empowering Holders to Limit Revocation Checks for Verifiable Credentials
Praveensankar Manimaran, Mayank Raikwar, Thiago Garrett +3
Systems managing Verifiable Credentials are becoming increasingly popular. Unfortunately, their support for revoking previously issued credentials allows verifiers to effectively m…
On Replacing Cryptopuzzles with Useful Computation in Blockchain Proof-of-Work Protocols
Andrea Merlina, Thiago Garrett, Roman Vitenberg
Proof-of-Work (PoW) blockchains have emerged as a robust and effective consensus mechanism in open environments, leading to widespread deployment with numerous cryptocurrency platf…
Privacy-preserving transactive energy systems: Key topics and open research challenges
Daniel Gerbi Duguma, Juliana Zhang, Meysam Aboutalebi +20
This manuscript aims to formalize and conclude the discussions initiated during the PriTEM workshop 22-23 March 2023. We present important ideas and discussion topics in the contex…
Keep Your Friends Close: Leveraging Affinity Groups to Accelerate AI Inference Workflows
Thiago Garrett, Weijia Song, Roman Vitenberg +1
AI inference workflows are typically structured as a pipeline or graph of AI programs triggered by events. As events occur, the AIs perform inference or classification tasks under…
Cascade: A Platform for Delay-Sensitive Edge Intelligence
Weijia Song, Thiago Garrett, Yuting Yang +6
Interactive intelligent computing applications are increasingly prevalent, creating a need for AI/ML platforms optimized to reduce per-event latency while maintaining high throughp…