4 papers
Rewrite-to-Rank: Optimizing Ad Visibility via Retrieval-Aware Text Rewriting
Chloe Ho, Ishneet Sukhvinder Singh, Diya Sharma +4
Search algorithms and user query relevance have given LLMs the ability to return relevant information, but the effect of content phrasing on ad visibility remains underexplored. We…
NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts
Abhay Gupta, Michael Lu, Kevin Zhu +2
Current large language models (LLMs) struggle to answer questions that span tens of thousands of tokens, especially when multi-hop reasoning is involved. While prior benchmarks exp…
Distill CLIP (DCLIP): Enhancing Image-Text Retrieval via Cross-Modal Transformer Distillation
Daniel Csizmadia, Andrei Codreanu, Victor Sim +5
We present Distill CLIP (DCLIP), a fine-tuned variant of the CLIP model that enhances multimodal image-text retrieval while preserving the original model's strong zero-shot classif…
Rosetta-PL: Propositional Logic as a Benchmark for Large Language Model Reasoning
Shaun Baek, Shaun Esua-Mensah, Cyrus Tsui +6
Large Language Models (LLMs) are primarily trained on high-resource natural languages, limiting their effectiveness in low-resource settings and in tasks requiring deep logical rea…