2 papers
cs.LG2025
AccidentBench: Benchmarking Multimodal Understanding and Reasoning in Vehicle Accidents and Beyond
Shangding Gu, Xiaohan Wang, Donghao Ying +9
Rapid advances in multimodal models demand benchmarks that rigorously evaluate understanding and reasoning in safety-critical, dynamic real-world settings. We present AccidentBench…
cs.LG2025
Few-Shot Test-Time Optimization Without Retraining for Semiconductor Recipe Generation and Beyond
Shangding Gu, Donghao Ying, Ming Jin +4
We introduce Model Feedback Learning (MFL), a novel test-time optimization framework for optimizing inputs to pre-trained AI models or deployed hardware systems without requiring a…