1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CL2024★ 1 cited
SCAR: Sparse Conditioned Autoencoders for Concept Detection and Steering in LLMs
Ruben Härle, Felix Friedrich, Manuel Brack +3
Large Language Models (LLMs) have demonstrated remarkable capabilities in generating human-like text, but their output may not be aligned with the user or even produce harmful cont…
cs.AI2022
LogicRank: Logic Induced Reranking for Generative Text-to-Image Systems
Björn Deiseroth, Patrick Schramowski, Hikaru Shindo +2
Text-to-image models have recently achieved remarkable success with seemingly accurate samples in photo-realistic quality. However as state-of-the-art language models still struggl…