5 papers
Learning to Detect Cross-Modal Negation: An Analysis of Latent Representations and an Attention-Based Solution
Ali AbuSaleh, Leon Hammerla, Alexander Mehler
Detecting high-level semantic concepts like negation across modalities remains a challenge for current multimodal systems. We analyze this as a fundamental representation learning…
MMTM: Tri-Modal Topic Modeling for Long-Form Video via Similarity-Gated Fusion
Ali Abusaleh, Bhuvanesh Verma, Alexander Mehler
We introduce MMTM, a modular pipeline for topic discovery in long-form video that integrates speech recognition, audio and visual embeddings, and BERTopic clustering through a dete…
From Early Encoding to Late Suppression: Interpreting LLMs on Character Counting Tasks
Ayan Datta, Mounika Marreddy, Alexander Mehler +2
Large language models (LLMs) exhibit failures on elementary symbolic tasks such as character counting in a word, despite excelling on complex benchmarks. Although this limitation h…
The Spatial Semantics of Iconic Gesture
Andy Lücking, Alexander Henlein, Alexander Mehler
The current multimodal turn in linguistic theory leaves a crucial question unanswered: what is the meaning of iconic gestures, and how does it compose with speech meaning? We argue…
You Shall Know a Tool by the Traces it Leaves: The Predictability of Sentiment Analysis Tools
Daniel Baumartz, Mevlüt Bagci, Alexander Henlein +3
If sentiment analysis tools were valid classifiers, one would expect them to provide comparable results for sentiment classification on different kinds of corpora and for different…