9 papers
Belief-reality separation lives in routing over a shared value slot in language models
Oliver Steele, Jiangtao Wen, Yuxing Han
Capable language models hold what a character believes apart from what is true: told "Anna believes the cup is blue; in reality it is red," they answer blue about Anna and red abou…
One mechanism for many mental spaces: a shared router over a value slot in language models
Oliver Steele, Jiangtao Wen, Yuxing Han
Language builds discourse contexts other than the actual: a painting, a belief, a memory, a hypothetical. Each is a mental space in which the same entity can take a different value…
It Takes a Good Model to Train a Good Model: Generalized Gaussian Priors for Optimized LLMs
Jun Wu, Patrick Huang, Jiangtao Wen +1
Despite rapid progress in large language models (LLMs), the statistical structure of their weights, activations, and gradients-and its implications for initialization, training dyn…
Efficient Bayer-Domain Video Computer Vision with Fast Motion Estimation and Learned Perception Residual
Haichao Wang, Jiangtao Wen, Yuxing Han
Video computer vision systems face substantial computational burdens arising from two fundamental challenges: eliminating unnecessary processing and reducing temporal redundancy in…
ReSeFlow: Rectifying SE(3)-Equivariant Policy Learning Flows
Zhitao Wang, Yanke Wang, Jiangtao Wen +2
Robotic manipulation in unstructured environments requires the generation of robust and long-horizon trajectory-level policy with conditions of perceptual observations and benefits…
Vision without Images: End-to-End Computer Vision from Single Compressive Measurements
Fengpu Pan, Heting Gao, Jiangtao Wen +1
Snapshot Compressed Imaging (SCI) offers high-speed, low-bandwidth, and energy-efficient image acquisition, but remains challenged by low-light and low signal-to-noise ratio (SNR)…