2 papers
cs.CV2026
GameReplica: A Benchmark for Black-Box Visual Game Replication by Vision-Language Agents
Boyu Qiao, Zixin Tang, Xiaoshuai Hao +1
Coding-agent benchmarks usually evaluate implementation after the target behavior has been specified in text, code, or demonstrations. Existing research has extensively evaluated t…
cs.LG2026
Not All Ranks Are Equal: Budget-Aware LoRA Merging Across Tasks
Avinash Amballa, Yashas Malur Saidutta, Wenbo Li +2
Merging low-rank adapters (LoRAs) promises to eliminate the overhead of swapping task-specific weights at inference time. However, existing merging methods assume every layer needs…