22 citations · 94 across the 19 of their papers we have counts for
16 papers
All in an Aggregated Image for In-Image Learning
Lei Wang, Wanyu Xu, Zhiqiang Hu +5
This paper introduces a new in-context learning (ICL) mechanism called In-Image Learning (IL) that combines demonstration examples, visual cues, and chain-of-thought reasoning…
MemeCraft: Contextual and Stance-Driven Multimodal Meme Generation
Han Wang, Roy Ka-Wei Lee
Online memes have emerged as powerful digital cultural artifacts in the age of social media, offering not only humor but also platforms for political discourse, social critique, an…
Modularized Networks for Few-shot Hateful Meme Detection
Rui Cao, Roy Ka-Wei Lee, Jing Jiang
In this paper, we address the challenge of detecting hateful memes in the low-resource setting where only a few labeled examples are available. Our approach leverages the compositi…
SBTRec- A Transformer Framework for Personalized Tour Recommendation Problem with Sentiment Analysis
Ngai Lam Ho, Roy Ka-Wei Lee, Kwan Hui Lim
When traveling to an unfamiliar city for holidays, tourists often rely on guidebooks, travel websites, or recommendation systems to plan their daily itineraries and explore popular…
Language Guided Visual Question Answering: Elevate Your Multimodal Language Model Using Knowledge-Enriched Prompts
Deepanway Ghosal, Navonil Majumder, Roy Ka-Wei Lee +2
Visual question answering (VQA) is the task of answering questions about an image. The task assumes an understanding of both the image and the question to provide a natural languag…
BTRec: BERT-Based Trajectory Recommendation for Personalized Tours
Ngai Lam Ho, Roy Ka-Wei Lee, Kwan Hui Lim
An essential task for tourists having a pleasant holiday is to have a well-planned itinerary with relevant recommendations, especially when visiting unfamiliar cities. Many tour re…