person
Monte MacDiarmid
Monte MacDiarmid is a core research contributor to the reward-hacking generalization study.
Monte MacDiarmid is a core research contributor to the reward-hacking generalization study, documented by the original paper.[1]
Contribution
Natural Emergent Misalignment from Reward Hacking in Production RL describes the contribution and its limitations. This is a contribution-focused stub; current affiliations have not been independently established.
Sources
Pages that link here
Last updated 2026-10-07