5 citations · 13 across the 5 of their papers we have counts for
6 papers
Agentic AI for Human Resources: LLM-Driven Candidate Assessment
Kamer Ali Yuksel, Abdul Basit Anees, Ashraf Elneima +3
In this work, we present a modular and interpretable framework that uses Large Language Models (LLMs) to automate candidate assessment in recruitment. The system integrates diverse…
ITEm: Unsupervised Image-Text Embedding Learning for eCommerce
Baohao Liao, Michael Kozielski, Sanjika Hewavitharana +3
Product embedding serves as a cornerstone for a wide range of applications in eCommerce. The product embedding learned from multiple modalities shows significant improvement over t…
Mask More and Mask Later: Efficient Pre-training of Masked Language Models by Disentangling the [MASK] Token
Baohao Liao, David Thulke, Sanjika Hewavitharana +2
The pre-training of masked language models (MLMs) consumes massive computation to achieve good results on downstream NLP tasks, resulting in a large carbon footprint. In the vanill…
Back-translation for Large-Scale Multilingual Machine Translation
Baohao Liao, Shahram Khadivi, Sanjika Hewavitharana
This paper illustrates our approach to the shared task on large-scale multilingual machine translation in the sixth conference on machine translation (WMT-21). This work aims to bu…
Word-based Domain Adaptation for Neural Machine Translation
Shen Yan, Leonard Dahlmann, Pavel Petrushkov +2
In this paper, we empirically investigate applying word-level weights to adapt neural machine translation to e-commerce domains, where small e-commerce datasets and large out-of-do…
Towards Semantic Query Segmentation
Ajinkya Kale, Thrivikrama Taula, Sanjika Hewavitharana +1
Query Segmentation is one of the critical components for understanding users' search intent in Information Retrieval tasks. It involves grouping tokens in the search query into mea…