7 citations · 13 across the 13 of their papers we have counts for
4 papers · 1 filter
CulturalMenuBench: Probing the Knowledge-Application Gap in Multimodal Culinary Reasoning
Bo Zeng, Linfeng Gao, Peiqin Lin +9
Multimodal language models achieve near-ceiling scores on food recognition benchmarks, yet it remains unclear whether this success reflects genuine cultural understanding or mere v…
What Matters for Aggressive Decoding-Time KV Eviction? Temporal Aggregation and Ranking Preservation
Bo Zeng, Yu Zhao, Yefeng Liu +3
Decoding-time KV cache compression research focuses heavily on designing better token scoring functions, while the temporal rule that aggregates scores across decode steps is often…
LISA: Linear-Indexed Sparse Attention for Efficient Long-Context Reasoning
Yu Zhao, Zekun Zhang, Fan Jiang +6
Recent advances in long chain-of-thought reasoning models such as DeepSeek-R1 have led to increasingly longer inference context lengths under the test-time scaling paradigm. Howeve…
Difficulty-Estimated Policy Optimization
Yu Zhao, Fan Jiang, Tianle Liu +4
Recent advancements in Large Reasoning Models (LRMs), exemplified by DeepSeek-R1, have underscored the potential of scaling inference-time compute through Group Relative Policy Opt…