89 citations · 93 across the 5 of their papers we have counts for
6 papers
UNIVID: Unified Vision-Language Model for Video Moderation
Kejuan Yang, Yizhuo Zhang, Mingyuan Du +6
Global-scale video moderation faces a dual challenge: the need for fine-grained multi-modal reasoning and the demand for interpretable outputs to support downstream enforcement. Tr…
GazeSummary: Exploring Gaze as an Implicit Prompt for Personalization in Text-based LLM Tasks
Jiexin Ding, Yizhuo Zhang, Xinyun Liu +4
Smart glasses are accelerating progress toward more seamless and personalized LLM-based assistance by integrating multimodal inputs. Yet, these inputs rely on obtrusive explicit pr…
Generalizable LLM Learning of Graph Synthetic Data with Post-training Alignment
Yizhuo Zhang, Heng Wang, Shangbin Feng +3
Previous research has sought to enhance the graph reasoning capabilities of LLMs by supervised fine-tuning on synthetic graph data. While these led to specialized LLMs better at so…
Can LLM Graph Reasoning Generalize beyond Pattern Memorization?
Yizhuo Zhang, Heng Wang, Shangbin Feng +4
Large language models (LLMs) demonstrate great potential for problems with implicit graphical structures, while recent works seek to enhance the graph reasoning capabilities of LLM…
Less Is More: Fast Multivariate Time Series Forecasting with Light Sampling-oriented MLP Structures
Tianping Zhang, Yizhuo Zhang, Wei Cao +4
Multivariate time series forecasting has seen widely ranging applications in various domains, including finance, traffic, energy, and healthcare. To capture the sophisticated tempo…
On Rotation Gains Within and Beyond Perceptual Limitations for Seated VR
Chen Wang, Song-Hai Zhang, Yizhuo Zhang +2
Head tracking in head-mounted displays (HMDs) enables users to explore a 360-degree virtual scene with free head movements. However, for seated use of HMDs such as users sitting on…