Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Do Image-Text Metrics Respect Semantic Invariances?
Amit Agarwal, Hitesh Laxmichand Patel, Meizhu Liu +9
Reference-free image-to-text evaluators are now standard for scoring image-caption alignment, yet it is unclear whether they respect semantic invariances. We present an invariance…
cs.CV2026
Lightweight and Production-Ready PDF Visual Element Parsing
Meizhu Liu, Yassi Abbasi, Matthew Rowe +2
PDF documents contain critical visual elements such as figures, tables, and forms whose accurate extraction is essential for document understanding and multimodal retrieval-augment…