5 citations · 10 across the 2 of their papers we have counts for
3 papers
cs.CL2018★ 5 cited
Exploring Hyper-Parameter Optimization for Neural Machine Translation on GPU Architectures
Robert Lim, Kenneth Heafield, Hieu Hoang +2
Neural machine translation (NMT) has been accelerated by deep learning neural networks over statistical-based approaches, due to the plethora and programmability of commodity heter…
cs.DC2017
A Similarity Measure for GPU Kernel Subgraph Matching
Robert Lim, Boyana Norris, Allen Malony
Accelerator architectures specialize in executing SIMD (single instruction, multiple data) in lockstep. Because the majority of CUDA applications are parallelized loops, control fl…
cs.DC2017★ 5 cited
Autotuning GPU Kernels via Static and Predictive Analysis
Robert V. Lim, Boyana Norris, Allen D. Malony
Optimizing the performance of GPU kernels is challenging for both human programmers and code generators. For example, CUDA programmers must set thread and block parameters for a ke…