◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

D. Cai

4 papers hereh-index 316 citations4 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author1
  • last author1

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.LG4
same name
  • D. Cai — 12 papers, h 26
  • D. Cai — 4 papers, h 11
  • D. Cai — 3 papers, h 10
  • D. Cai — 1 paper, h 3
  • D. Cai — 1 paper, h 8
  • D. Cai — 1 paper, h 8

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedVariational Bayesian Inference for Crowdsourcing Predictions

1 citations · 1 across the 4 of their papers we have counts for

collaborators
Showing cs.LGShow all

4 papers · 1 filter

cs.LG2021

Decentralized Deterministic Multi-Agent Reinforcement Learning

Antoine Grosnit, Desmond Cai, Laura Wynter

[Zhang, ICML 2018] provided the first decentralized actor-critic algorithm for multi-agent reinforcement learning (MARL) that offers convergence guarantees. In that work, policies…

cs.LG2021

Efficient Reinforcement Learning in Resource Allocation Problems Through Permutation Invariant Multi-task Learning

Desmond Cai, Shiau Hong Lim, Laura Wynter

One of the main challenges in real-world reinforcement learning is to learn successfully from limited training samples. We show that in certain settings, the available data can be…

cs.LG2021

Probabilistic Inference for Learning from Untrusted Sources

Duc Thien Nguyen, Shiau Hoong Lim, Laura Wynter +1

Federated learning brings potential benefits of faster learning, better solutions, and a greater propensity to transfer when heterogeneous data from different parties increases div…

cs.LG2020★ 1 cited

Variational Bayesian Inference for Crowdsourcing Predictions

Desmond Cai, Duc Thien Nguyen, Shiau Hong Lim +1

Crowdsourcing has emerged as an effective means for performing a number of machine learning tasks such as annotation and labelling of images and other data sets. In most early sett…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.