activity
20222024
most citedHierarchical Audio-Visual Information Fusion with Multi-label Joint Decoding for MER 2023

3 citations · 4 across the 7 of their papers we have counts for

collaborators

7 papers

cs.CL2024

SRFUND: A Multi-Granularity Hierarchical Structure Reconstruction Benchmark in Form Understanding

Jiefeng Ma, Yan Wang, Chenyu Liu +6

Accurately identifying and organizing textual content is crucial for the automation of document processing in the field of form understanding. Existing datasets, such as FUNSD and…

cs.SD2024

A Study of Dropout-Induced Modality Bias on Robustness to Missing Video Frames for Audio-Visual Speech Recognition

Yusheng Dai, Hang Chen, Jun Du +5

Advanced Audio-Visual Speech Recognition (AVSR) systems have been observed to be sensitive to missing video frames, performing even worse than single-modality models. While applyin…

cs.CV2023

Bidirectional Trained Tree-Structured Decoder for Handwritten Mathematical Expression Recognition

Hanbo Cheng, Chenyu Liu, Pengfei Hu +3

The Handwritten Mathematical Expression Recognition (HMER) task is a critical branch in the field of OCR. Recent studies have demonstrated that incorporating bidirectional context…

eess.AS20233 cited

Hierarchical Audio-Visual Information Fusion with Multi-label Joint Decoding for MER 2023

Haotian Wang, Yuxuan Xi, Hang Chen +11

In this paper, we propose a novel framework for recognizing both discrete and dimensional emotions. In our framework, deep features extracted from foundation models are used as rob…

cs.CV2023

Count, Decode and Fetch: A New Approach to Handwritten Chinese Character Error Correction

Pengfei Hu, Jiefeng Ma, Zhenrong Zhang +2

Recently, handwritten Chinese character error correction has been greatly improved by employing encoder-decoder methods to decompose a Chinese character into an ideographic descrip…

cs.CL20231 cited

HRDoc: Dataset and Baseline Method Toward Hierarchical Reconstruction of Document Structures

Jiefeng Ma, Jun Du, Pengfei Hu +4

The problem of document structure reconstruction refers to converting digital or scanned documents into corresponding semantic structures. Most existing works mainly focus on split…