1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2025
LLaVA-c: Continual Improved Visual Instruction Tuning
Wenzhuo Liu, Fei Zhu, Haiyang Guo +2
Multimodal models like LLaVA-1.5 achieve state-of-the-art visual understanding through visual instruction tuning on multitask datasets, enabling strong instruction-following and mu…
cs.CL2024
From Statistical Methods to Pre-Trained Models; A Survey on Automatic Speech Recognition for Resource Scarce Urdu Language
Muhammad Sharif, Zeeshan Abbas, Jiangyan Yi +1
Automatic Speech Recognition (ASR) technology has witnessed significant advancements in recent years, revolutionizing human-computer interactions. While major languages have benefi…
cs.CV2024★ 1 cited
A comprehensive survey of oracle character recognition: challenges, benchmarks, and beyond
Jing Li, Xueke Chi, Qiufeng Wang +4
Oracle character recognition-an analysis of ancient Chinese inscriptions found on oracle bones-has become a pivotal field intersecting archaeology, paleography, and historical cult…