9 citations · 9 across the 2 of their papers we have counts for
2 papers
cs.CL2023★ 9 cited
The Wisdom of Hindsight Makes Language Models Better Instruction Followers
Tianjun Zhang, Fangchen Liu, Justin Wong +2
Reinforcement learning has seen wide success in finetuning large language models to better align with instructions via human feedback. The so-called algorithm, Reinforcement Learni…
cs.CV2022
Context-Aware Streaming Perception in Dynamic Environments
Gur-Eyal Sela, Ionel Gog, Justin Wong +9
Efficient vision works maximize accuracy under a latency budget. These works evaluate accuracy offline, one image at a time. However, real-time vision applications like autonomous…