activity
20232026
most citedBlackboxBench: A Comprehensive Benchmark of Black-box Adversarial Attacks

2 citations · 5 across the 15 of their papers we have counts for

collaborators
Showing 2026Show all

5 papers · 1 filter

cs.CV2026

Distributed Implicit Harm: A Compositional Safety Blind Spot in MLLM-Based Video Moderation

Ruotong Wang, Zihao Zhu, Siwei Lyu +2

Despite their growing use in video moderation, multimodal large language models (MLLMs) exhibit a compositional safety blind spot: videos composed of seemingly benign components ca…

cs.CL2026

Can Induced Emotion Bias LLM Behaviors in Sequential Decision Making?

Minh Khoi Ho, Zihao Zhu, Runchuan Zhu +4

As Large Language Models (LLMs) are increasingly deployed as autonomous agents in high-stakes domains, understanding contextual factors that may modulate their decision-making beco…

cs.AI2026

A Statistical Framework for Auditing Behavioral Dependence and Induced Bias in LLM Judges

Chenchen Kuai, Jiwan Jiang, Zihao Zhu +8

The rapid growth of the large language model (LLM) ecosystem raises a critical question: are seemingly diverse models truly independent? Shared pretraining data, distillation, and…

cs.CV2026

BrandFusion: A Multi-Agent Framework for Seamless Brand Integration in Text-to-Video Generation

Zihao Zhu, Ruotong Wang, Siwei Lyu +2

The rapid advancement of text-to-video (T2V) models has revolutionized content creation, yet their commercial potential remains largely untapped. We introduce, for the first time,…

cs.LG2026

Unveiling Covert Toxicity in Multimodal Data via Toxicity Association Graphs: A Graph-Based Metric and Interpretable Detection Framework

Guanzong Wu, Zihao Zhu, Siwei Lyu +1

Detecting toxicity in multimodal data remains a significant challenge, as harmful meanings often lurk beneath seemingly benign individual modalities: only emerging when modalities…