1 citations · 1 across the 1 of their papers we have counts for
2 papers
eess.AS2025
Frame-Stacked Local Transformers For Efficient Multi-Codebook Speech Generation
Roy Fejgin, Paarth Neekhara, Xuesong Yang +6
Speech generation models based on large language models (LLMs) typically operate on discrete acoustic codes, which differ fundamentally from text tokens due to their multicodebook…
eess.IV2021★ 1 cited
Automatic calibration of time of flight based non-line-of-sight reconstruction
Subhash Chandra Sadhu, Abhishek Singh, Tomohiro Maeda +4
Time of flight based Non-line-of-sight (NLOS) imaging approaches require precise calibration of illumination and detector positions on the visible scene to produce reasonable resul…