3 citations · 3 across the 6 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Alignment-Aware Model Adaptation via Feedback-Guided Optimization
Gaurav Bhatt, Aditya Chinchure, Jiawei Zhou +1
Fine-tuning is the primary mechanism for adapting foundation models to downstream tasks; however, standard approaches largely optimize task objectives in isolation and do not accou…
cs.LG2025
Features Emerge as Discrete States: The First Application of SAEs to 3D Representations
Albert Miao, Chenliang Zhou, Jiawei Zhou +1
Sparse Autoencoders (SAEs) are a powerful dictionary learning technique for decomposing neural network activations, translating the hidden state into human ideas with high semantic…