Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Label-Free Reinforcement Learning via Cross-Model Entropy
Matt Gorbett, Hossein Shirazi
Post-training large language models with reinforcement learning is bottlenecked by the reward signal. Existing approaches require either ground-truth verifiable rewards, restrictin…
cs.LG2024
Tiled Bit Networks: Sub-Bit Neural Network Compression Through Reuse of Learnable Binary Vectors
Matt Gorbett, Hossein Shirazi, Indrakshi Ray
Binary Neural Networks (BNNs) enable efficient deep learning by saving on storage and computational costs. However, as the size of neural networks continues to grow, meeting comput…
cs.LG2024
Cross-Silo Federated Learning Across Divergent Domains with Iterative Parameter Alignment
Matt Gorbett, Hossein Shirazi, Indrakshi Ray
Learning from the collective knowledge of data dispersed across private sources can provide neural networks with enhanced generalization capabilities. Federated learning, a method…