Rohan Paul

@rohanpaul_ai

You can predict where an LLM's internal state is heading, and that targeted edits can pull it back on course. On a frozen Hermes 8B, 8 numbers, taken from the model's hidden states at one layer, predict 68 measurements of what happens several layers later. The prediction error comes out about 69-76% lower than a simple average baseline.
打开原帖#511482
  1. Industry

    Databricks: Two new frontier models just landed on Databricks
  2. Industry

    Replit ⠕: Analyze your app funnel directly in Replit
  3. Industry

    Replit ⠕: Models keep getting smarter