cross-lingual modeling 1hierarchical classification 1language-agnostic codebooks 1LLM agents 1multilingual sign language translation 1multimodal translation 1paper organization 1personalized routing 1reference management 1vector quantization 1
From the 2 of 16 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
PhysUniBench: A Multi-Modal Physics Reasoning Benchmark at Undergraduate Level
Lintao Wang, Encheng Su, Jiaqi Liu +11
Physics problem-solving is a challenging domain for AI models, requiring integration of conceptual understanding, mathematical reasoning, and interpretation of physical diagrams. E…
cs.AI2026
SciIF: Benchmarking Scientific Instruction Following Towards Rigorous Scientific Intelligence
Encheng Su, Jianyu Wu, Chen Tang +9
As large language models (LLMs) transition from general knowledge retrieval to complex scientific discovery, their evaluation standards must also incorporate the rigorous norms of…