2 papers
cs.CL2024
Prompt-Time Symbolic Knowledge Capture with Large Language Models
Tolga Çöplü, Arto Bendiken, Andrii Skomorokhov +3
Augmenting large language models (LLMs) with user-specific knowledge is crucial for real-world applications, such as personal AI assistants. However, LLMs inherently lack mechanism…
cs.LG2023
A Performance Evaluation of a Quantized Large Language Model on Various Smartphones
Tolga Çöplü, Marc Loedi, Arto Bendiken +3
This paper explores the feasibility and performance of on-device large language model (LLM) inference on various Apple iPhone models. Amidst the rapid evolution of generative AI, o…