Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Label-Free Reinforcement Learning via Cross-Model Entropy
Matt Gorbett, Hossein Shirazi
Post-training large language models with reinforcement learning is bottlenecked by the reward signal. Existing approaches require either ground-truth verifiable rewards, restrictin…
cs.LG2024
Tiled Bit Networks: Sub-Bit Neural Network Compression Through Reuse of Learnable Binary Vectors
Matt Gorbett, Hossein Shirazi, Indrakshi Ray
Binary Neural Networks (BNNs) enable efficient deep learning by saving on storage and computational costs. However, as the size of neural networks continues to grow, meeting comput…