Rajan Agarwal

Research

Models, training and evals

  1. Post-training a small model to follow a live stream of typing and learn when to act or stay quiet.

    Jun 2026
  2. An ultra long-horizon benchmark for coding agents doing engineering and research work.

  3. Reinforcement-learning work for the browser-use model as a research intern.

    Dec 2025 · Research internship at Amazon AGI, Mentioned in: Amazon Science
  4. An RL setup in which language models learn their own compact codebooks.

    Nov 2025 · Mentioned in: Deep Learning with Yacine
  5. Hidden-information games for studying deceptive behavior learned from reward.

    Nov 2025
  6. Cross-lingual alignment through encoder injection for low-resource languages.

    Oct 2025 · With Cohere Labs