◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Ronghui Mu

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author3

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CV2
  • cs.CR1
  • cs.LG1

identity via Semantic Scholar / OpenAlex

activity
20222024
most citedSafeguarding Large Language Models: A Survey

8 citations · 10 across the 4 of their papers we have counts for

collaborators

4 papers

cs.CR2024★ 8 cited

Safeguarding Large Language Models: A Survey

Yi Dong, Ronghui Mu, Yanghao Zhang +9

In the burgeoning field of Large Language Models (LLMs), developing a robust safety mechanism, colloquially known as "safeguards" or "guardrails", has become imperative to ensure t…

cs.CV2024

Towards Fairness-Aware Adversarial Learning

Yanghao Zhang, Tianle Zhang, Ronghui Mu +2

Although adversarial training (AT) has proven effective in enhancing the model's robustness, the recently revealed issue of fairness in robustness has not been well addressed, i.e.…

cs.LG2023★ 2 cited

Randomized Adversarial Training via Taylor Expansion

Gaojie Jin, Xinping Yi, Dengyu Wu +2

In recent years, there has been an explosion of research into developing more robust deep neural networks against adversarial examples. Adversarial training appears as one of the m…

cs.CV2022

3DVerifier: Efficient Robustness Verification for 3D Point Cloud Models

Ronghui Mu, Wenjie Ruan, Leandro S. Marcolino +1

3D point cloud models are widely applied in safety-critical scenes, which delivers an urgent need to obtain more solid proofs to verify the robustness of models. Existing verificat…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.