6 papers
DiPS: Dialogue Policy Selection for High-Stakes Persuasion Agents
Tianyi Zhang, Mousumi Das, Abrar Anwar +2
Large Language Models (LLMs) often struggle with persuasion in high-stakes scenarios. People's individual personalities and concerns require tailored strategies rather than a one-s…
RISK: A Framework for GUI Agents in E-commerce Risk Management
Renqi Chen, Zeyin Tao, Jianming Guo +5
E-commerce risk management requires aggregating diverse, deeply embedded web data through multi-step, stateful interactions, which traditional scraping methods and most existing Gr…
FML-bench: Benchmarking Machine Learning Agents for Scientific Research
Qiran Zou, Hou Hei Lam, Wenhao Zhao +7
Large language models (LLMs) have sparked growing interest in machine learning research agents that can autonomously propose ideas and conduct experiments. However, existing benchm…
Personalized Chain-of-Thought Summarization of Financial News for Investor Decision Support
Tianyi Zhang, Mu Chen
Financial advisors and investors struggle with information overload from financial news, where irrelevant content and noise obscure key market signals and hinder timely investment…
A new approach for fine-tuning sentence transformers for intent classification and out-of-scope detection tasks
Tianyi Zhang, Atta Norouzian, Aanchan Mohan +1
In virtual assistant (VA) systems it is important to reject or redirect user queries that fall outside the scope of the system. One of the most accurate approaches for out-of-scope…
Human Latency Conversational Turns for Spoken Avatar Systems
Derek Jacoby, Tianyi Zhang, Aanchan Mohan +1
A problem with many current Large Language Model (LLM) driven spoken dialogues is the response time. Some efforts such as Groq address this issue by lightning fast processing of th…