From the 4 of 103 linked papers with an AI index.
16 papers · 1 filter
Regime-Conditional Verification: Correctness Estimation for Adapting and Monitoring Safety Classifiers
Thiago Sandoval, Ufuk Topcu
Safety classifiers deployed with large language models often fail for two reasons: their decisions reflect the policy learned during training rather than the deployer's desired pol…
What We are Missing in Multimodal LLM Evaluation?
Po-han Li, Shenghui Chen, Sandeep Chinchali +1
Multimodal large language models (MLLMs) can process diverse inputs, e.g., text, images, audio, and video, and generate textual responses. While their capabilities have advanced ra…
Physically Viable World Models: A Case for Query-Conditioned Embodied AI
Adam J. Thorpe, Stepan Tretiakov, Cheng-Hsi Hsiao +6
World models for embodied AI must be physically viable: constructed to answer intervention queries by representing the physical structure governing action outcomes, rather than mer…
Foundation Models for Logistics: Toward Certifiable, Conversational Planning Interfaces
Yunhao Yang, Neel P. Bhatt, Christian Ellis +4
Logistics operators, from battlefield coordinators re-routing airlifts ahead of a storm to warehouse managers juggling late trucks, need to make mission-critical decisions. Prevail…
Neurosymbolic LoRA: Why and When to Tune Weights vs. Rewrite Prompts
Kevin Wang, Neel P. Bhatt, Cong Liu +7
Large language models (LLMs) can be adapted either through numerical updates that alter model parameters or symbolic manipulations that work on discrete prompts or logical constrai…
Joint Verification and Refinement of Language Models for Safety-Constrained Planning
Yunhao Yang, Neel P. Bhatt, William Ward +3
Large language models possess impressive capabilities in generating programs (e.g., Python) from natural language descriptions to execute robotic tasks. However, these generated pr…