Rohan Paul
@rohanpaul_ai
You can predict where an LLM's internal state is heading, and that targeted edits can pull it back on course.
On a frozen Hermes 8B, 8 numbers, taken from the model's hidden states at one layer, predict 68 measurements of what happens several layers later.
The prediction error comes out about 69-76% lower than a simple average baseline.