12 citations · 59 across the 20 of their papers we have counts for
25 papers
Detecting Depression in Thai Blog Posts: a Dataset and a Baseline
Mika Hämäläinen, Pattama Patpong, Khalid Alnajjar +2
We present the first openly available corpus for detecting depression in Thai. Our corpus is compiled by expert verified cases of depression in several online blogs. We experiment…
Finnish Dialect Identification: The Effect of Audio and Text
Mika Hämäläinen, Khalid Alnajjar, Niko Partanen +1
Finnish is a language with multiple dialects that not only differ from each other in terms of accent (pronunciation) but also in terms of morphological forms and lexical choice. We…
The Current State of Finnish NLP
Mika Hämäläinen, Khalid Alnajjar
There are a lot of tools and resources available for processing Finnish. In this paper, we survey recent papers focusing on Finnish NLP related to many different subcategories of N…
When a Computer Cracks a Joke: Automated Generation of Humorous Headlines
Khalid Alnajjar, Mika Hämäläinen
Automated news generation has become a major interest for new agencies in the past. Oftentimes headlines for such automatically generated news articles are unimaginative as they ha…
How Cute is Pikachu? Gathering and Ranking Pokémon Properties from Data with Pokémon Word Embeddings
Mika Hämäläinen, Khalid Alnajjar, Niko Partanen
We present different methods for obtaining descriptive properties automatically for the 151 original Pokémon. We train several different word embeddings models on a crawled Pokémon…
Human Evaluation of Creative NLG Systems: An Interdisciplinary Survey on Recent Papers
Mika Hämäläinen, Khalid Alnajjar
We survey human evaluation in papers presenting work on creative natural language generation that have been published in INLG 2020 and ICCC 2020. The most typical human evaluation…