1 paper · 1 filter
Or Bachar, Or Levi, Sardhendu Mishra +6
As LLMs are increasingly integrated into human-in-the-loop content moderation systems, a central challenge is deciding when their outputs can be trusted versus when escalation for…