activity
20192021
collaborators

6 papers

cs.CL2021

Going Full-TILT Boogie on Document Understanding with Text-Image-Layout Transformer

Rafał Powalski, Łukasz Borchmann, Dawid Jurkiewicz +3

We address the challenging problem of Natural Language Comprehension beyond plain-text documents by introducing the TILT neural network architecture which simultaneously learns lay…

cs.CL2020

From Dataset Recycling to Multi-Property Extraction and Beyond

Tomasz Dwojak, Michał Pietruszka, Łukasz Borchmann +2

This paper investigates various Transformer architectures on the WikiReading Information Extraction and Machine Reading Comprehension dataset. The proposed dual-source model outper…

cs.LG2020

Successive Halving Top-k Operator

Michał Pietruszka, Łukasz Borchmann, Filip Graliński

We propose a differentiable successive halving method of relaxing the top-k operator, rendering gradient-based optimization possible. The need to perform softmax iteratively on the…

cs.CL2020

On the Multi-Property Extraction and Beyond

Tomasz Dwojak, Michał Pietruszka, Łukasz Borchmann +2

In this paper, we investigate the Dual-source Transformer architecture on the WikiReading information extraction and machine reading comprehension dataset. The proposed model outpe…

cs.CL2020

ApplicaAI at SemEval-2020 Task 11: On RoBERTa-CRF, Span CLS and Whether Self-Training Helps Them

Dawid Jurkiewicz, Łukasz Borchmann, Izabela Kosmala +1

This paper presents the winning system for the propaganda Technique Classification (TC) task and the second-placed system for the propaganda Span Identification (SI) task. The purp…

cs.CL2019

Contract Discovery: Dataset and a Few-Shot Semantic Retrieval Challenge with Competitive Baselines

Łukasz Borchmann, Dawid Wiśniewski, Andrzej Gretkowski +7

We propose a new shared task of semantic retrieval from legal texts, in which a so-called contract discovery is to be performed, where legal clauses are extracted from documents, g…