8 papers
A System for Name and Address Parsing with Large Language Models
Adeeba Tarannum, Muzakkiruddin Ahmed Mohammed, Mert Can Cakmak +2
Reliable transformation of unstructured person and address text into structured data remains a key challenge in large-scale information systems. Traditional rule-based and probabil…
Case Count Metric for Comparative Analysis of Entity Resolution Results
John R. Talburt, Muzakkiruddin Ahmed Mohammed, Mert Can Cakmak +4
This paper describes a new process and software system, the Case Count Metric System (CCMS), for systematically comparing and analyzing the outcomes of two different ER clustering…
Policy-Aware Generative AI for Safe, Auditable Data Access Governance
Shames Al Mandalawi, Muzakkiruddin Ahmed Mohammed, Hendrika Maclean +2
Enterprises need access decisions that satisfy least privilege, comply with regulations, and remain auditable. We present a policy aware controller that uses a large language model…
A Keyframe-Based Approach for Auditing Bias in YouTube Shorts Recommendations
Mert Can Cakmak, Nitin Agarwal
YouTube Shorts and other short-form video platforms now influence how billions engage with content, yet their recommendation systems remain largely opaque. Small shifts in promoted…
Efficient Data Retrieval and Comparative Bias Analysis of Recommendation Algorithms for YouTube Shorts and Long-Form Videos
Selimhan Dagtas, Mert Can Cakmak, Nitin Agarwal
The growing popularity of short-form video content, such as YouTube Shorts, has transformed user engagement on digital platforms, raising critical questions about the role of recom…
Investigating Algorithmic Bias in YouTube Shorts
Mert Can Cakmak, Nitin Agarwal, Diwash Poudel
The rapid growth of YouTube Shorts, now serving over 2 billion monthly users, reflects a global shift toward short-form video as a dominant mode of online content consumption. This…