From the 1 of 4 linked papers with an AI index.
4 citations · 4 across the 3 of their papers we have counts for
4 papers
From Critic to Confidence: PPO for Language-Based Quantitative Prediction with Confidence Estimation
Mehak Dhaliwal, Rasta Tadayon, Andong Hua +2
The paper introduces CARE-PPO, a reinforcement‑learning framework that fine‑tunes large language models to make accurate numeric predictions while simultaneously learning confidenc…
English is Not All You Need: Systematically Exploring the Role of Multilinguality in LLM Post-Training
Mehak Dhaliwal, Shashwat Chaurasia, Yao Qin +2
Despite the widespread multilingual deployment of large language models, post-training pipelines remain predominantly English-centric, contributing to performance disparities acros…
NutriBench: A Dataset for Evaluating Large Language Models on Nutrition Estimation from Meal Descriptions
Andong Hua, Mehak Preet Dhaliwal, Laya Pullela +2
Accurate nutrition estimation helps people make informed dietary choices and is essential in the prevention of serious health complications. We present NutriBench, the first public…
A Shocking Amount of the Web is Machine Translated: Insights from Multi-Way Parallelism
Brian Thompson, Mehak Preet Dhaliwal, Peter Frisch +2
We show that content on the web is often translated into many languages, and the low quality of these multi-way translations indicates they were likely created using Machine Transl…