3 papers
cs.LG2024
LAMPO: Large Language Models as Preference Machines for Few-shot Ordinal Classification
Zhen Qin, Junru Wu, Jiaming Shen +2
We introduce LAMPO, a novel paradigm that leverages Large Language Models (LLMs) for solving few-shot multi-class ordinal classification tasks. Unlike conventional methods, which c…
cs.LG2024
Principled Architecture-aware Scaling of Hyperparameters
Wuyang Chen, Junru Wu, Zhangyang Wang +1
Training a high-quality deep neural network requires choosing suitable hyperparameters, which is a non-trivial and expensive process. Current works try to automatically optimize or…
cs.CV2021
Auto-X3D: Ultra-Efficient Video Understanding via Finer-Grained Neural Architecture Search
Yifan Jiang, Xinyu Gong, Junru Wu +3
Efficient video architecture is the key to deploying video recognition systems on devices with limited computing resources. Unfortunately, existing video architectures are often co…