7 papers
MultiView-Bench: A Diagnostic Benchmark for World-Centric Multi-View Integration in VLMs
Hantao Zhang, Jinru Sui, Ed Li +2
Recent benchmarks for VLMs largely assess single- or limited-view perception, leaving untested the core cognitive ability to integrate observations across viewpoints into a coheren…
Scaling Multiagent Systems with Process Rewards
Ed Li, Junyu Ren, Cat Yan
While multiagent systems have shown promise for tackling complex tasks via specialization, finetuning multiple agents simultaneously faces two key challenges: (1) credit assignment…
Llama-3.1-FoundationAI-SecurityLLM-Reasoning-8B Technical Report
Zhuoran Yang, Ed Li, Jianliang He +18
We present Foundation-Sec-8B-Reasoning, the first open-source native reasoning model for cybersecurity. Built upon our previously released Foundation-Sec-8B base model (derived fro…
Copyright Detection in Large Language Models: An Ethical Approach to Generative AI Development
David Szczecina, Senan Gaffori, Edmond Li
The widespread use of Large Language Models (LLMs) raises critical concerns regarding the unauthorized inclusion of copyrighted content in training data. Existing detection framewo…
See it. Say it. Sorted: Agentic System for Compositional Diagram Generation
Hantao Zhang, Jingyang Liu, Ed Li
We study sketch-to-diagram generation: converting rough hand sketches into precise, compositional diagrams. Diffusion models excel at photorealism but struggle with the spatial pre…
Build Your Personalized Research Group: A Multiagent Framework for Continual and Interactive Science Automation
Ed Li, Junyu Ren, Xintian Pan +4
The automation of scientific discovery represents a critical milestone in Artificial Intelligence (AI) research. However, existing agentic systems for science suffer from two funda…