Tagged articles

temporal difference learning

1 articles · Page 1 of 1
Data Party THU
Data Party THU
Oct 1, 2026 · Artificial Intelligence

Stanford & Tsinghua Find Two Neuron Types Encoding Reward in LLMs

Researchers from Stanford and Tsinghua identify sparse value and dopamine neurons in LLMs that encode state value and temporal difference errors, forming a causal reward subsystem enabling confidence estimation and process reward modeling for improved reasoning.

LLM interpretabilityconfidence estimationdopamine neurons
0 likes · 7 min read
Stanford & Tsinghua Find Two Neuron Types Encoding Reward in LLMs