Rajan Agarwal

Writing

Notes and essays

  1. Training a small model to read typing as a stream and learn when not to act.

    Jun 2026
  2. A postmortem on world models, failed assumptions, and research habits.

    Jan 2026
  3. Turning compression into an RL problem and watching models invent compact languages.

    Nov 2025
  4. What hidden-information games reveal about deception learned from reward.

    Nov 2025
  5. Moving an unusual distributed RL setup onto a research API.

    Nov 2025