2 papers
cs.RO2026
FRAMES: Failure Recovery And Monitoring of Embodied Skills for Humanoid Loco-Manipulation
Ajay Vikram Periasami, Xinyuan Luo, Haoyu Li +1
Large language model (LLM) planners can decompose natural-language instructions and select reusable robot skills, but choosing the correct skill does not guarantee successful physi…
cs.CV2026
Vision2Code: A Multi-Domain Benchmark for Evaluating Image-to-Code Generation
Ajay Vikram Periasami, Junlin Wang, Bhuwan Dhingra
Image-to-code generation tests whether a vision-language model (VLM) can recover the structure of an image enough to express it as executable code. Existing benchmarks either focus…