8 citations · 10 across the 4 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025★ 1 cited
Know Thy Judge: On the Robustness Meta-Evaluation of LLM Safety Judges
Francisco Eiras, Eliott Zemour, Eric Lin +1
Large Language Model (LLM) based judges form the underpinnings of key safety evaluation processes such as offline benchmarking, automated red-teaming, and online guardrailing. This…
cs.LG2023★ 8 cited
Does fine-tuning GPT-3 with the OpenAI API leak personally-identifiable information?
Albert Yu Sun, Eliott Zemour, Arushi Saxena +4
Machine learning practitioners often fine-tune generative pre-trained models like GPT-3 to improve model performance at specific tasks. Previous works, however, suggest that fine-t…