Research Scientist at Ai2 · RLHF & post-trainingResearch Scientist at Ai2 · RLHF & post-training
rlnlp
American AI researcher who leads post-training work at the Allen Institute for AI (Ai2), where he helped build the open Tulu and OLMo model and recipe families. He earned his PhD at UC Berkeley and previously worked on RLHF at Hugging Face. Through his widely read Interconnects newsletter he has become one of the clearest public explainers of reinforcement learning from human feedback, preference tuning, and the mechanics of open versus closed model development.
American AI researcher who leads post-training work at the Allen Institute for AI (Ai2), where he helped build the open Tulu and OLMo model and recipe families. He earned his PhD at UC Berkeley and previously worked on RLHF at Hugging Face. Through his widely read Interconnects newsletter he has become one of the clearest public explainers of reinforcement learning from human feedback, preference tuning, and the mechanics of open versus closed model development.