1 citations · 2 across the 2 of their papers we have counts for
2 papers
astro-ph.IM2025★ 1 cited
AstroMMBench: A Benchmark for Evaluating Multimodal Large Language Models Capabilities in Astronomy
Jinghang Shi, Xiaoyu Tang, Yang Huang +4
Astronomical image interpretation presents a significant challenge for applying multimodal large language models (MLLMs) to specialized scientific tasks. Existing benchmarks focus…
cs.CL2025★ 1 cited
Statutory Construction and Interpretation for Artificial Intelligence
Luxi He, Nimra Nadeem, Michel Liao +4
AI systems are increasingly governed by natural language principles, yet a key challenge arising from reliance on language remains underexplored: interpretive ambiguity. As in lega…