3 papers
cs.AI2026
Representation Affects Retrieval: A Case Study of Skill Discovery and Routing in a Multimodal Agent Harness
Kevin Dela Rosa
A production agent harness must discover and rank, from a growing library of skills, the one most appropriate for a user's task. At small scale this selection happens in context: t…
cs.CV2025
Smart Routing for Multimodal Video Retrieval: When to Search What
Kevin Dela Rosa
We introduce ModaRoute, an LLM-based intelligent routing system that dynamically selects optimal modalities for multimodal video retrieval. While dense text captions can achieve 75…
cs.IR2025
RAVEN: An Agentic Framework for Multimodal Entity Discovery from Large-Scale Video Collections
Kevin Dela Rosa
We present RAVEN an adaptive AI agent framework designed for multimodal entity discovery and retrieval in large-scale video collections. Synthesizing information across visual, aud…