1 paper
Alexander Shim, Khalil Saieh, Samuel Clarke
This research analyzed and compared the multi-modal approach in the Vision Transformer(EVA-ViT) based image encoder with the LlaMA or ChatGPT LLM to reduce the hallucination proble…