6 citations · 7 across the 10 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
LLM-JEPA: Large Language Models Meet Joint Embedding Predictive Architectures
Hai Huang, Yann LeCun, Randall Balestriero
Large Language Model (LLM) pretraining, finetuning, and evaluation rely on input-space reconstruction and generative capabilities. Yet, it has been observed in vision that embeddin…
cs.CL2025
Next Token Perception Score: Analytical Assessment of your LLM Perception Skills
Yu-Ang Cheng, Leyang Hu, Hai Huang +1
Autoregressive pretraining has become the de facto paradigm for learning general-purpose representations in large language models (LLMs). However, linear probe performance across d…