5 papers
Tokens, the oft-overlooked appetizer: Large language models, the distributional hypothesis, and meaning
Julia Witte Zimmerman, Denis Hudon, Kathryn Cramer +9
Tokenization is a necessary component within the current architecture of many language mod-els, including the transformer-based large language models (LLMs) of Generative AI, yet i…
From Flowers to Fascism? The Cottagecore to Tradwife Pipeline on Tumblr
Oliver Mel Allen, Yi Zu, Milo Z. Trujillo +1
In this work we collected and analyzed social media posts to investigate aesthetic-based radicalization where users searching for Cottagecore content may find Tradwife content co-o…
Invisible Labor: The Backbone of Open Source Software
Robin A. Lange, Anna Gibson, Milo Z. Trujillo +1
Invisible labor is an intrinsic part of the modern workplace, and includes labor that is undervalued or unrecognized such as creating collaborative atmospheres. Open source softwar…
Invisible Labor in Open Source Software Ecosystems
John Meluso, Amanda Casari, Katie McLaughlin +1
Invisible labor is work that is either not fully visible or not appropriately compensated. In open source software (OSS) ecosystems, essential tasks that do not involve code (like…
A blind spot for large language models: Supradiegetic linguistic information
Julia Witte Zimmerman, Denis Hudon, Kathryn Cramer +5
Large Language Models (LLMs) like ChatGPT reflect profound changes in the field of Artificial Intelligence, achieving a linguistic fluency that is impressively, even shockingly, hu…