2 papers
cs.AI2022
Sequence-to-Sequence Models for Extracting Information from Registration and Legal Documents
Ramon Pires, Fábio C. de Souza, Guilherme Rosa +2
A typical information extraction pipeline consists of token- or span-level classification models coupled with a series of pre- and post-processing scripts. In a production pipeline…
cs.CL2019
Portuguese Named Entity Recognition using BERT-CRF
Fábio Souza, Rodrigo Nogueira, Roberto Lotufo
Recent advances in language representation using neural networks have made it viable to transfer the learned internal states of a trained model to downstream natural language proce…