2 citations · 2 across the 2 of their papers we have counts for
3 papers
cs.LG2025
SRPO: A Cross-Domain Implementation of Large-Scale Reinforcement Learning on LLM
Xiaojiang Zhang, Jinghui Wang, Zifei Cheng +14
Recent advances of reasoning models, exemplified by OpenAI's o1 and DeepSeek's R1, highlight the significant potential of Reinforcement Learning (RL) to enhance the reasoning capab…
cs.CR2024★ 2 cited
Security of AI Agents
Yifeng He, Ethan Wang, Yuyang Rong +2
AI agents have been boosted by large language models. AI agents can function as intelligent assistants and complete tasks on behalf of their users with access to tools and the abil…
cs.RO2023
DAVIS-Ag: A Synthetic Plant Dataset for Prototyping Domain-Inspired Active Vision in Agricultural Robots
Taeyeong Choi, Dario Guevara, Zifei Cheng +5
In agricultural environments, viewpoint planning can be a critical functionality for a robot with visual sensors to obtain informative observations of objects of interest (e.g., fr…