Kyle Harrison
concept

AI Safety

AI Safety

In Kyle’s notes, AI Safety is the central theme of Sam Altman’s Planning for AGI and Beyond (OpenAI’s mission statement on stewarding AGI into existence). The piece’s distinctive posture is that safety and capability are not separable — Altman calls it “a false dichotomy to talk about them separately,” arguing OpenAI’s best safety work has come from working with its most capable models, while insisting that the ratio of safety progress to capability progress must increase over time. The other load-bearing claim is the existential-risk stance: some in the field think AGI’s risks are fictitious, but OpenAI chooses to “operate as if these risks are existential” even while hoping the skeptics turn out right.

The concept also carries Altman’s governance argument — that a gradual, iteratively-deployed transition to an AGI world beats a sudden one, navigated through “a tight feedback loop of rapid learning and careful iteration,” and that “the future of humanity should be determined by humanity.” This is where AI Safety connects to Kyle’s broader storytelling thread: Planning for AGI and Beyond is one of three companion long-reads feeding The Meme Economy (via The Meme Economy - Research), where Altman’s “determined by humanity” framing rhymes with the essay’s argument that narratives steer reality.

Context: “AI safety” broadly refers to the field of research and practice aimed at ensuring advanced AI systems behave as intended and do not cause catastrophic harm — spanning alignment, interpretability, robustness, and governance. The safety-vs-capabilities debate Altman addresses is a live tension within that field.

Where this appears

  • Planning for AGI and Beyond — the safety/capability false-dichotomy argument, the safety-to-capability ratio, and the existential-risk posture are all tagged AI Safety.
  • The Meme Economy - Research — listed as the theme of the Altman long-read, one of the three companion reads feeding the essay.