16 citations · 16 across the 3 of their papers we have counts for
3 papers
cs.CV2026
JieZi: A Large-Scale Expert-Audited Dataset and Benchmark for Ancient Chinese Character Exegesis
Ran Li, Huiguo He, Jiahuan Cao +3
The scholarly exegesis of ancient Chinese characters demands integrating visual observation, linguistic analysis, and historical context. However, existing computational approaches…
cs.CV2025
Capturing More: Learning Multi-Domain Representations for Robust Online Handwriting Verification
Peirong Zhang, Kai Ding, Lianwen Jin
In this paper, we propose SPECTRUM, a temporal-frequency synergistic model that unlocks the untapped potential of multi-domain representation learning for online handwriting verifi…
cs.CL2024★ 16 cited
Datasets for Large Language Models: A Comprehensive Survey
Yang Liu, Jiahuan Cao, Chongyu Liu +2
This paper embarks on an exploration into the Large Language Model (LLM) datasets, which play a crucial role in the remarkable advancements of LLMs. The datasets serve as the found…