1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.LG2025
The Evolving Landscape of LLM- and VLM-Integrated Reinforcement Learning
Sheila Schoepp, Masoud Jafaripour, Yingyue Cao +6
Reinforcement learning (RL) has shown impressive results in sequential decision-making tasks. Meanwhile, Large Language Models (LLMs) and Vision-Language Models (VLMs) have emerged…
cs.LG2023★ 1 cited
LaFFi: Leveraging Hybrid Natural Language Feedback for Fine-tuning Language Models
Qianxi Li, Yingyue Cao, Jikun Kang +4
Fine-tuning Large Language Models (LLMs) adapts a trained model to specific downstream tasks, significantly improving task-specific performance. Supervised Fine-Tuning (SFT) is a c…