collaborators

9 papers

cs.CL2026

FinReportBench: Measuring and Improving Institution-Grade Financial Report Generation

Yinghao Tang, Tan Zhenwei, Yiyao Wang +4

Large language models can produce fluent financial analysis, but fluency alone does not establish whether a report is suitable for institutional delivery. We introduce FinReportBen…

cs.AI2026

LiveEvalBench: Toward Open-World Evaluation for Web Generation

Yiyao Wang, Zhen Wen, Yinghao Tang +5

Large language models are increasingly capable of synthesizing executable frontend projects, yet existing benchmarks still treat web generation as a static evaluation problem. We a…

cs.HC2026

RAGExplorer: A Visual Analytics System for the Comparative Diagnosis of RAG Systems

Haoyu Tian, Yingchaojie Feng, Zhen Wen +3

The advent of Retrieval-Augmented Generation (RAG) has significantly enhanced the ability of Large Language Models (LLMs) to produce factually accurate and up-to-date responses. Ho…

cs.CV2026

V-Zero: Self-Improving Multimodal Reasoning with Zero Annotation

Han Wang, Yi Yang, Jingyuan Hu +2

Recent advances in multimodal learning have significantly enhanced the reasoning capabilities of vision-language models (VLMs). However, state-of-the-art approaches rely heavily on…

cs.CL2025

Multimodal DeepResearcher: Generating Text-Chart Interleaved Reports From Scratch with Agentic Framework

Zhaorui Yang, Bo Pan, Han Wang +8

Visualizations play a crucial part in effective communication of concepts and information. Recent advances in reasoning and retrieval augmented generation have enabled Large Langua…

cs.CL2025

ConceptViz: A Visual Analytics Approach for Exploring Concepts in Large Language Models

Haoxuan Li, Zhen Wen, Qiqi Jiang +7

Large language models (LLMs) have achieved remarkable performance across a wide range of natural language tasks. Understanding how LLMs internally represent knowledge remains a sig…