◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yen-Chun Chen

4 papers hereh-index 81k citations12 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author3
  • last author1

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CV3
  • cs.CL1
same name
  • Yen-Chun Chen — 8 papers
  • Yen-Chun Chen — 2 papers
  • Yen-Chun Chen — 1 paper, h 6
  • Yen-Chun Chen — 1 paper
  • Yen-Chun Chen — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedVALUE: A Multi-Task Benchmark for Video-and-Language Understanding Evaluation

38 citations · 50 across the 3 of their papers we have counts for

collaborators

4 papers

cs.CV2021★ 38 cited

VALUE: A Multi-Task Benchmark for Video-and-Language Understanding Evaluation

Linjie Li, Jie Lei, Zhe Gan +12

Most existing video-and-language (VidL) research focuses on a single dataset, or multiple datasets of a single task. In reality, a truly useful VidL system is expected to be easily…

cs.CL2021★ 6 cited

LightningDOT: Pre-training Visual-Semantic Embeddings for Real-Time Image-Text Retrieval

Siqi Sun, Yen-Chun Chen, Linjie Li +3

Multimodal pre-training has propelled great advancement in vision-and-language research. These large-scale pre-trained models, although successful, fatefully suffer from slow infer…

cs.CV2020★ 6 cited

Self-Prediction for Joint Instance and Semantic Segmentation of Point Clouds

Jinxian Liu, Minghui Yu, Bingbing Ni +1

We develop a novel learning scheme named Self-Prediction for 3D instance and semantic segmentation of point clouds. Distinct from most existing methods that focus on designing conv…

cs.CV2020

Large-Scale Adversarial Training for Vision-and-Language Representation Learning

Zhe Gan, Yen-Chun Chen, Linjie Li +3

We present VILLA, the first known effort on large-scale adversarial training for vision-and-language (V+L) representation learning. VILLA consists of two training stages: (i) task-…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.