5 papers
TUR-DPO: Topology- and Uncertainty-Aware Direct Preference Optimization
Abdulhady Abas Abdullah, Fatemeh Daneshfar, Seyedali Mirjalili +1
Aligning large language models (LLMs) with human preferences is commonly done via reinforcement learning from human feedback (RLHF) with Proximal Policy Optimization (PPO) or, more…
Beyond Wave Variables: A Data-Driven Ensemble Approach for Enhanced Teleoperation Transparency and Stability
Nour Mitiche, Farid Ferguene, Mourad Oussalah
Time delays in communication channels present significant challenges for bilateral teleoperation systems, affecting both transparency and stability. Although traditional wave varia…
AnomalyExplainer Explainable AI for LLM-based anomaly detection using BERTViz and Captum
Prasasthy Balasubramanian, Dumindu Kankanamge, Ekaterina Gilman +1
Conversational AI and Large Language Models (LLMs) have become powerful tools across domains, including cybersecurity, where they help detect threats early and improve response tim…
PaPaformer: Language Model from Pre-trained Parallel Paths
Joonas Tapaninaho, Mourad Oussala
The training of modern large-language models requires an increasingly amount of computation power and time. Even smaller variants, such as small-language models (SLMs), take severa…
Towards an Automated Multimodal Approach for Video Summarization: Building a Bridge Between Text, Audio and Facial Cue-Based Summarization
Md Moinul Islam, Sofoklis Kakouros, Janne Heikkilä +1
The increasing volume of video content in educational, professional, and social domains necessitates effective summarization techniques that go beyond traditional unimodal approach…