3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.CL2026
TL-GRPO: Turn-Level RL for Reasoning-Guided Iterative Optimization
Peiji Li, Linyang Li, Handa Sun +15
Large language models have demonstrated strong reasoning capabilities in complex tasks through tool integration, which is typically framed as a Markov Decision Process and optimize…
cs.AR2024★ 3 cited
AnalogGym: An Open and Practical Testing Suite for Analog Circuit Synthesis
Jintao Li, Haochang Zhi, Ruiyu Lyu +9
Recent advances in machine learning (ML) for automating analog circuit synthesis have been significant, yet challenges remain. A critical gap is the lack of a standardized evaluati…