2 citations · 2 across the 3 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2025
Unsupervised decoding of encoded reasoning using language model interpretability
Ching Fang, Samuel Marks
As large language models become increasingly capable, there is growing concern that they may develop reasoning processes that are encoded or hidden from human oversight. To investi…
cs.AI2025
From Memories to Maps: Mechanisms of In-Context Reinforcement Learning in Transformers
Ching Fang, Kanaka Rajan
Humans and animals show remarkable learning efficiency, adapting to new environments with minimal experience. This capability is not well captured by standard reinforcement learnin…
cs.AI2023
Predictive auxiliary objectives in deep RL mimic learning in the brain
Ching Fang, Kimberly L Stachenfeld
The ability to predict upcoming events has been hypothesized to comprise a key aspect of natural and machine cognition. This is supported by trends in deep reinforcement learning (…