1 paper
Rahul Bissa, Abhishek Vyas, Yash Jain
We benchmark three supervised fine-tuned models against frontier zero-shot baselines on a 661-row held-out slice of PiSAR (Persona, intent, Screen, Action, Rationale), a 12,929-tup…