◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Zhao-Yang Fu

3 papers hereh-index 391 citations5 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author3

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.RO2
  • cs.LG1

identity via Semantic Scholar / OpenAlex

most citedSemi-Supervised Multi-Modal Multi-Instance Multi-Label Deep Network with Optimal Transport

17 citations · 17 across the 1 of their papers we have counts for

collaborators

3 papers

cs.RO2026

In-Context World Modeling for Robotic Control

Siyin Wang, Junhao Shi, Senyu Fei +4

Modern Vision-Language-Action (VLA) models often fail to generalize to novel setups, such as altered camera viewpoints or robot morphologies, because they are typically conditioned…

cs.RO2026

Two Bridges, One Pathway: From VLMs to Generalizable VLAs with Embodied Trajectory-Coupled Data

Linqi Yin, Shiduo Zhang, Shenling Qiu +11

Vision-language models (VLMs) are powerful general-purpose reasoners, yet converting them into robot control policies (VLAs) is surprisingly difficult. The root cause is a two-fold…

cs.LG2021★ 17 cited

Semi-Supervised Multi-Modal Multi-Instance Multi-Label Deep Network with Optimal Transport

Yang Yang, Zhao-Yang Fu, De-Chuan Zhan +2

Complex objects are usually with multiple labels, and can be represented by multiple modal representations, e.g., the complex articles contain text and image information as well as…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.