3 papers
cs.CL2025
TimeSense:Making Large Language Models Proficient in Time-Series Analysis
Zhirui Zhang, Changhua Pei, Tianyi Gao +7
In the time-series domain, an increasing number of works combine text with temporal data to leverage the reasoning capabilities of large language models (LLMs) for various downstre…
cs.CL2025
Few-shot Policy (de)composition in Conversational Question Answering
Kyle Erwin, Guy Axelrod, Maria Chang +8
The task of policy compliance detection (PCD) is to determine if a scenario is in compliance with respect to a set of written policies. In a conversational setting, the results of…
cs.LG2024
Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone
Max Sobol Mark, Tian Gao, Georgia Gabriela Sampaio +4
Recent advances in learning decision-making policies can largely be attributed to training expressive policy models, largely via imitation learning. While imitation learning discar…