From the 1 of 3 linked papers with an AI index.
1 citations · 1 across the 3 of their papers we have counts for
3 papers
ReLope: KL-Regularized LoRA Probes for Multimodal LLM Routing
Yaopei Zeng, Congchao Wang, Blake JianHang Chen +1
The paper proposes two methods—a attention‑based probe and a KL‑regularized LoRA probe (ReLope)—to improve routing decisions in multimodal large language models by extracting more…
Think Deep, Not Just Long: Measuring LLM Reasoning Effort via Deep-Thinking Tokens
Wei-Lin Chen, Liqian Peng, Tian Tan +5
Large language models (LLMs) have demonstrated impressive reasoning capabilities by scaling test-time compute via long Chain-of-Thought (CoT). However, recent findings suggest that…
Gemma 4 Technical Report
Gemma Team, Sherif El Abd, Vaibhav Aggarwal +320
We introduce Gemma 4, a new generation of open-weight, natively multimodal language models in the Gemma model family. Designed to advance compute efficiency and reasoning, the Gemm…