3 papers
cs.CV2026
Token-Efficient Multimodal Reasoning via Image Prompt Packaging
Joong Ho Choi, Jiayang Zhao, Avani Appalla +3
Deploying large multimodal language models at scale is constrained by token-based inference costs, yet the cost-performance behavior of visual prompting strategies remains poorly c…
cs.AI2025
CompactPrompt: A Unified Pipeline for Prompt Data Compression in LLM Workflows
Joong Ho Choi, Jiayang Zhao, Jeel Shah +5
Large Language Models (LLMs) deliver powerful reasoning and generation capabilities but incur substantial run-time costs when operating in agentic workflows that chain together len…
cs.LG2025
Generative Large-Scale Pre-trained Models for Automated Ad Bidding Optimization
Yu Lei, Jiayang Zhao, Yilei Zhao +4
Modern auto-bidding systems are required to balance overall performance with diverse advertiser goals and real-world constraints, reflecting the dynamic and evolving needs of the i…