2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CL2023
Red Teaming for Large Language Models At Scale: Tackling Hallucinations on Mathematics Tasks
Aleksander Buszydlik, Karol Dobiczek, Michał Teodor Okoń +3
We consider the problem of red teaming LLMs on elementary calculations and algebraic tasks to evaluate how various prompting techniques affect the quality of outputs. We present a…
cs.HC2023★ 2 cited
How do you feel? Measuring User-Perceived Value for Rejecting Machine Decisions in Hate Speech Detection
Philippe Lammerts, Philip Lippmann, Yen-Chia Hsu +2
Hate speech moderation remains a challenging task for social media platforms. Human-AI collaborative systems offer the potential to combine the strengths of humans' reliability and…