◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Zhiyong Wang

4 papers hereh-index 559 citations8 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author2

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.LG4
same name
  • Zhiyong Wang — 20 papers, h 8
  • Zhiyong Wang — 17 papers, h 4
  • Zhiyong Wang — 11 papers, h 3
  • Zhiyong Wang — 8 papers, h 1
  • Zhiyong Wang — 7 papers, h 3
  • Zhiyong Wang — 7 papers, h 4

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20232025
most citedOnline Clustering of Bandits with Misspecified User Models

2 citations · 2 across the 3 of their papers we have counts for

collaborators

4 papers

cs.LG2025

Cascading Bandits Robust to Adversarial Corruptions

Jize Xie, Cheng Chen, Zhiyong Wang +1

Online learning to rank sequentially recommends a small list of items to users from a large candidate set and receives the users' click feedback. In many real-world scenarios, user…

cs.LG2024

Combinatorial Multivariant Multi-Armed Bandits with Applications to Episodic Reinforcement Learning and Beyond

Xutong Liu, Siwei Wang, Jinhang Zuo +7

We introduce a novel framework of combinatorial multi-armed bandits (CMAB) with multivariant and probabilistically triggering arms (CMAB-MT), where the outcome of each arm is a d…

cs.LG2023

Online Corrupted User Detection and Regret Minimization

Zhiyong Wang, Jize Xie, Tong Yu +2

In real-world online web systems, multiple users usually arrive sequentially into the system. For applications like click fraud and fake reviews, some users can maliciously perform…

cs.LG2023★ 2 cited

Online Clustering of Bandits with Misspecified User Models

Zhiyong Wang, Jize Xie, Xutong Liu +2

The contextual linear bandit is an important online learning problem where given arm features, a learning agent selects an arm at each round to maximize the cumulative rewards in t…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.