papers
Publications (3)
cs.CV2024
SASE: A Searching Architecture for Squeeze and Excitation Operations
Hanming Wang, Yunlong Li, Zijun Wu +2
In the past few years, channel-wise and spatial-wise attention blocks have been widely adopted as supplementary modules in deep neural networks, enhancing network representational…
cs.CL2024
MIMIR: A Streamlined Platform for Personalized Agent Tuning in Domain Expertise
Chunyuan Deng, Xiangru Tang, Yilun Zhao +5
Recently, large language models (LLMs) have evolved into interactive agents, proficient in planning, tool use, and task execution across a wide variety of tasks. However, without s…
cs.LG2022
MBGDT:Robust Mini-Batch Gradient Descent
Hanming Wang, Haozheng Luo, Yue Wang
In high dimensions, most machine learning method perform fragile even there are a little outliers. To address this, we hope to introduce a new method with the base learner, such as…