works on

From the 1 of 14 linked papers with an AI index.

most citedAn AI Co-Data-Scientist for Prioritizing Candidate Biomarkers from Wearable Sensor Data

2 citations · 2 across the 2 of their papers we have counts for

collaborators

14 papers

cs.CL2026

The Hidden Puppet Master: Predicting Human Belief Change in Manipulative LLM Dialogues

Jocelyn Shen, Amina Luvsanchultem, Jessica Kim +6

The paper presents PUPPET, a taxonomy and dataset of over a thousand human‑LLM advice interactions, and studies how well language models can predict the amount of belief change the…

cs.AI20262 cited

An AI Co-Data-Scientist for Prioritizing Candidate Biomarkers from Wearable Sensor Data

Yubin Kim, Salman Rahman, Samuel Schmidgall +33

Wearable devices generate continuous physiological and behavioral data, but converting these signals into clinically reviewable biomarker hypotheses remains labor-intensive. We int…

cs.HC2026

Social Human Robot Embodied Conversation (SHREC) Dataset: Benchmarking Foundational Models' Social Reasoning

Dong Won Lee, Yubin Kim, Denison Guvenoz +5

Our work focuses on the social reasoning capabilities of foundation models for real-world human-robot interactions. We introduce the Social Human Robot Embodied Conversation (SHREC…

cs.AI2026

TeamBench: Evaluating Agent Coordination under Enforced Role Separation

Yubin Kim, Chanwoo Park, Taehan Kim +9

Agent systems often decompose a task across multiple roles, but these roles are typically specified by prompts rather than enforced by access controls. Without enforcement, a team…

cs.AI2026

InvThink: Premortem Reasoning for Safer Language Models

Yubin Kim, Taehan Kim, Eugene Park +4

We present InvThink, a training and prompting framework that requires the model to enumerate, analyze, and constrain potential failures before generating its final response. Unlike…

cs.CL2026

Aligning Dialogue Agents with Global Feedback via Large Language Model Multimodal Reward Decomposition

Dong Won Lee, Hae Won Park, Cynthia Breazeal +1

We propose a large language model based reward decomposition framework for aligning dialogue agents using only a single session-level feedback signal. We leverage the reasoning cap…