1 paper
Sihwa Lee, Janghwan Lee, Donghoon Yoo +4
Large reasoning models (LRMs) improve complex problem-solving by generating long intermediate reasoning traces, but this substantially increases inference costs. NVFP4 inference of…