Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
HEAL: Hindsight Entropy-Assisted Learning for Reasoning Distillation
Wenjing Zhang, Jiangze Yan, Jieyun Huang +7
Distilling reasoning capabilities from Large Reasoning Models (LRMs) into smaller models is typically constrained by the limitation of rejection sampling. Standard methods treat th…
cs.AI2025
Quantifying the Capability Boundary of DeepSeek Models: An Application-Driven Performance Analysis
Kaikai Zhao, Zhaoxiang Liu, Xuejiao Lei +12
DeepSeek-R1, known for its low training cost and exceptional reasoning capabilities, has achieved state-of-the-art performance on various benchmarks. However, detailed evaluations…
cs.AI2024
Piculet: Specialized Models-Guided Hallucination Decrease for MultiModal Large Language Models
Kohou Wang, Xiang Liu, Zhaoxiang Liu +2
Multimodal Large Language Models (MLLMs) have made significant progress in bridging the gap between visual and language modalities. However, hallucinations in MLLMs, where the gene…