2 papers
cs.LG2026
One-Step Gradient Delay is Not a Barrier for Large-Scale Asynchronous Pipeline Parallel LLM Pretraining
Philip Zmushko, Egor Petrov, Nursultan Abdullaev +2
Modern large-scale LLM pretraining benefits from utilizing Pipeline Parallelism; however, synchronous implementations leave GPUs idle during pipeline bubbles, wasting computational…
eess.SP2026
A methodology to rank importance of frequencies and channels in electromyography data with Decision Tree classifiers
Albert A. Nasybullin, Nursultan Abdullaev, Maksim A. Baranov +2
This study presents a methodology for identifying the most informative frequencies and channels in electromyography (EMG) data to evaluate muscle recovery using Decision Tree class…