13 citations · 13 across the 1 of their papers we have counts for
3 papers
cs.CL2023★ 13 cited
Spam-T5: Benchmarking Large Language Models for Few-Shot Email Spam Detection
Maxime Labonne, Sean Moran
This paper investigates the effectiveness of large language models (LLMs) in email spam detection by comparing prominent models from three distinct families: BERT-like, Sentence Tr…
cs.LG2023
Estimating class separability of text embeddings with persistent homology
Kostis Gourgoulias, Najah Ghalyan, Maxime Labonne +3
This paper introduces an unsupervised method to estimate the class separability of text datasets from a topological point of view. Using persistent homology, we demonstrate how tra…
cs.LG2023
A Benchmark Generative Probabilistic Model for Weak Supervised Learning
Georgios Papadopoulos, Fran Silavong, Sean Moran
Finding relevant and high-quality datasets to train machine learning models is a major bottleneck for practitioners. Furthermore, to address ambitious real-world use-cases there is…