201 citations · 240 across the 13 of their papers we have counts for
9 papers · 1 filter
Massively Multilingual Corpus of Sentiment Datasets and Multi-faceted Sentiment Classification Benchmark
Łukasz Augustyniak, Szymon Woźniak, Marcin Gruza +4
Despite impressive advancements in multilingual corpora collection and model training, developing large-scale deployments of multilingual models still presents a significant challe…
This is the way: designing and compiling LEPISZCZE, a comprehensive NLP benchmark for Polish
Łukasz Augustyniak, Kamil Tagowski, Albert Sawczyn +9
The availability of compute and data to train larger and larger language models increases the demand for robust methods of benchmarking the true progress of LM training. Recent yea…
Assessment of Massively Multilingual Sentiment Classifiers
Krzysztof Rajda, Łukasz Augustyniak, Piotr Gramacki +3
Models are increasing in size and complexity in the hunt for SOTA. But what if those 2\% increase in performance does not make a difference in a production use case? Maybe benefits…
Political Advertising Dataset: the use case of the Polish 2020 Presidential Elections
Łukasz Augustyniak, Krzysztof Rajda, Tomasz Kajdanowicz +1
Political campaigns are full of political ads posted by candidates on social media. Political advertisements constitute a basic form of campaigning, subjected to various social req…
Extracting Aspects Hierarchies using Rhetorical Structure Theory
Łukasz Augustyniak, Tomasz Kajdanowicz, Przemysław Kazienko
We propose a novel approach to generate aspect hierarchies that proved to be consistently correct compared with human-generated hierarchies. We present an unsupervised technique us…
Aspect Detection using Word and Char Embeddings with (Bi)LSTM and CRF
Łukasz Augustyniak, Tomasz Kajdanowicz, Przemysław Kazienko
We proposed a~new accurate aspect extraction method that makes use of both word and character-based embeddings. We have conducted experiments of various models of aspect extraction…