2 papers
cs.LG2026
M*: A Modular, Extensible, Serving System for Multimodal Models
Atindra Jha, Naomi Sagan, Keisuke Kamahori +9
We are entering a new era of composite model architectures that integrate diverse components such as vision encoders, language backbones, diffusion and flow heads, audio codecs, ac…
cs.GR2024
Hybrid Voxel Formats for Efficient Ray Tracing
Russel Arbore, Jeffrey Liu, Aidan Wefel +2
Voxels are a geometric representation used for rendering volumes, multi-resolution models, and indirect lighting effects. Since the memory consumption of uncompressed voxel volumes…