1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.LG2024
ParMod: A Parallel and Modular Framework for Learning Non-Markovian Tasks
Ruixuan Miao, Xu Lu, Cong Tian +2
The commonly used Reinforcement Learning (RL) model, MDPs (Markov Decision Processes), has a basic premise that rewards depend on the current state and action only. However, many r…
cs.LG2024
Preventing Catastrophic Overfitting in Fast Adversarial Training: A Bi-level Optimization Perspective
Zhaoxin Wang, Handing Wang, Cong Tian +1
Adversarial training (AT) has become an effective defense method against adversarial examples (AEs) and it is typically framed as a bi-level optimization problem. Among various AT…
cs.SE2024★ 1 cited
Towards Practical Requirement Analysis and Verification: A Case Study on Software IP Components in Aerospace Embedded Systems
Zhi Ma, Cheng Wen, Jie Su +4
IP-based software design is a crucial research field that aims to improve efficiency and reliability by reusing complex software components known as intellectual property (IP) comp…