1 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CL2024★ 1 cited
Towards Inference-time Category-wise Safety Steering for Large Language Models
Amrita Bhattacharjee, Shaona Ghosh, Traian Rebedea +1
While large language models (LLMs) have seen unprecedented advancements in capabilities and applications across a variety of use-cases, safety alignment of these models is still an…
cs.CL2022★ 1 cited
Prompt Learning for Domain Adaptation in Task-Oriented Dialogue
Makesh Narsimhan Sreedhar, Christopher Parisien
Conversation designers continue to face significant obstacles when creating production quality task-oriented dialogue systems. The complexity and cost involved in schema developmen…