18 citations · 72 across the 34 of their papers we have counts for
3 papers · 1 filter
Agri-LLaVA: Knowledge-Infused Large Multimodal Assistant on Agricultural Pests and Diseases
Liqiong Wang, Teng Jin, Jinyu Yang +3
In the general domain, large multimodal models (LMMs) have achieved significant advancements, yet challenges persist in applying them to specific fields, especially agriculture. As…
bit2bit: 1-bit quanta video reconstruction via self-supervised photon prediction
Yehe Liu, Alexander Krull, Hector Basevi +2
Quanta image sensors, such as SPAD arrays, are an emerging sensor technology, producing 1-bit arrays representing photon detection events over exposures as short as a few nanosecon…
Improving Object Detection via Local-global Contrastive Learning
Danai Triantafyllidou, Sarah Parisot, Ales Leonardis +1
Visual domain gaps often impact object detection performance. Image-to-image translation can mitigate this effect, where contrastive approaches enable learning of the image-to-imag…