papers
Publications (4)
cs.LG2025
Apple Intelligence Foundation Language Models: Tech Report 2025
Ethan Li, Anders Boesen Lindbo Larsen, Chen Zhang +395
cs.AI2026
AgentAtlas: Beyond Outcome Leaderboards for LLM Agents
Parsa Mazaheri, Kasra Mazaheri
cs.SE2026
REPOT: Recoverable Program-of-Thought via Checkpoint Repair
Parsa Mazaheri
cs.CL2024
Don't Believe Everything You Read: Enhancing Summarization Interpretability through Automatic Identification of Hallucinations in Large Language Models
Priyesh Vakharia, Devavrat Joshi, Meenal Chavan +3