◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Taslim Mahbub

2 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author1

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.CL1
  • cs.LG1

identity via Semantic Scholar / OpenAlex

most citedDomain Specific Benchmarks for Evaluating Multimodal Large Language Models

2 citations · 2 across the 2 of their papers we have counts for

collaborators

2 papers

cs.CL2025

Mitigating Self-Preference by Authorship Obfuscation

Taslim Mahbub, Shi Feng

Language models (LMs) judges are widely used to evaluate the quality of LM outputs. Despite many advantages, LM judges display concerning biases that can impair their integrity in…

cs.LG2025★ 2 cited

Domain Specific Benchmarks for Evaluating Multimodal Large Language Models

Khizar Anjum, Muhammad Arbab Arshad, Kadhim Hayawi +10

Large language models (LLMs) are increasingly being deployed across disciplines due to their advanced reasoning and problem solving capabilities. To measure their effectiveness, va…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.