◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Sercan Ö. Arik

Google

39 papers hereh-index 4717.3k citations137 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author21
  • last author15

Across the 36 of 39 papers where every author was matched, so the position is known.

fields
  • cs.CL15
  • cs.LG14
  • cs.AI5
  • cs.CV3
  • cs.CR1
  • cs.DB1
affiliations
  • Google
Homepage

identity via Semantic Scholar / OpenAlex

activity
20242026
most citedGemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

47 citations · 49 across the 9 of their papers we have counts for

collaborators
Showing cs.CVShow all

3 papers · 1 filter

cs.CV2025

VISTA: A Test-Time Self-Improving Video Generation Agent

Do Xuan Long, Xingchen Wan, Hootan Nakhost +3

Despite rapid advances in text-to-video synthesis, generated video quality remains critically dependent on precise user prompts. Existing test-time optimization methods, successful…

cs.CV2025

Mitigating Object Hallucination in MLLMs via Data-augmented Phrase-level Alignment

Pritam Sarkar, Sayna Ebrahimi, Ali Etemad +3

Despite their significant advancements, Multimodal Large Language Models (MLLMs) often generate factually inaccurate information, referred to as hallucination. In this work, we add…

cs.CV2024

CROME: Cross-Modal Adapters for Efficient Multimodal LLM

Sayna Ebrahimi, Sercan O. Arik, Tejas Nama +1

Multimodal Large Language Models (MLLMs) demonstrate remarkable image-language capabilities, but their widespread use faces challenges in cost-effective training and adaptation. Ex…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.