2 papers
cs.LG2026
Performance Foundations of Parallel & Distributed Reasoning Language Models
Maciej Besta, Leonard Schmidt, Lara Nonino +7
Reinforcement Learning with Verifiable Rewards (RLVR) and other RL-style post-training paradigms have been used for aligning large language models (LLMs) with reasoning standards.…
cs.PF2025
Accelerating Sparse Ternary GEMM for Quantized ML on Apple Silicon
Baraq Lipshitz, Alessio Melone, Charalampos Maraziaris +1
Sparse Ternary General Matrix-Matrix Multiplication (GEMM) remains under-optimized in existing libraries for Apple Silicon CPUs. We present a Sparse Ternary GEMM kernel optimized s…