4 papers · 1 filter
DRIFT: Detecting Representational Inconsistencies for Factual Truthfulness
Rohan Bhatnagar, Youran Sun, Chi Andrew Zhang +2
LLMs often produce fluent but incorrect answers, yet detecting such hallucinations typically requires multiple sampling passes or post-hoc verification, adding significant latency…
OptimAI: Optimization from Natural Language Using LLM-Powered AI Agents
Raghav Thind, Youran Sun, Ling Liang +1
Optimization plays a vital role in scientific research and practical applications. However, formulating a concrete optimization problem described in natural language into a mathema…
LLMs Meet Finance: Fine-Tuning Foundation Models for the Open FinLLM Leaderboard
Varun Rao, Youran Sun, Mahendra Kumar +3
This paper investigates the application of large language models (LLMs) to financial tasks. We fine-tuned foundation models using the Open FinLLM Leaderboard as a benchmark. Buildi…
rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Xinyu Guan, Li Lyna Zhang, Yifei Liu +5
We present rStar-Math to demonstrate that small language models (SLMs) can rival or even surpass the math reasoning capability of OpenAI o1, without distillation from superior mode…