Through an observation window marked by a fingerprint, immense server racks frame an empty hall with a vertical strip of white light reflected on the floor.
Language and powerVOL. I / 2018–2026

AITHEISM

AI chronology

From language models to mathematical discoveries, with original sources.

Observations
01 / LanguageChronology

2018–2026 / Official source

Chronology

2018–2026

  1. GPTLearning first

    OpenAI presented Transformer-based generative pretraining followed by supervised, task-specific fine-tuning.

    Official source · OpenAI
  2. GPT-2The weight of release

    OpenAI announced the 1.5-billion-parameter GPT-2, releasing its weights in stages rather than all at once.

    Official source · OpenAI
  3. GPT-3A few examples

    The 175-billion-parameter GPT-3 was evaluated with few-shot examples, without task-specific gradient updates.

    Official source · OpenAI
  4. ChatGPTThe reply arrives

    ChatGPT launched as a dialogue research preview trained with human feedback.

    Official source · OpenAI
  5. Constitutional AIWritten principles

    Anthropic’s Constitutional AI research used human-provided principles, self-critique, and AI feedback to shape behavior.

    Official source · Anthropic
  6. ClaudeAnother answer

    Anthropic introduced the AI assistants Claude and Claude Instant.

    Official source · Anthropic
  7. GPT-4Beyond text

    OpenAI announced GPT-4 with text/image input and text output; image input was not broadly available at launch.

    Official source · OpenAI
  8. Qwen-7B / Qwen-7B-ChatOpen weights

    Alibaba Cloud released pretrained Qwen-7B and Qwen-7B-Chat weights under license conditions.

    Official source · Alibaba Cloud
  9. FunSearchNew ground in combinatorics

    FunSearch combined language-model code generation with automatic evaluation to construct larger cap sets. It improved known lower bounds on an open problem.

    Official source · Google DeepMind
  10. AlphaProof / AlphaGeometry 2Four problems, complete proofs

    AlphaProof and AlphaGeometry 2 solved four of six IMO problems, reaching silver-medal standard. Humans formalized the questions; some solutions took up to three days.

    Official source · Google DeepMind
  11. o1-previewBefore the answer

    OpenAI introduced o1-preview, designed to spend more time reasoning before responding.

    Official source · OpenAI
  12. DeepSeek-V3Selective capacity

    DeepSeek released V3 and its weights: 671 billion total parameters, 37 billion activated per token, using mixture-of-experts.

    Official source · DeepSeek
  13. DeepSeek-R1Learning to reason

    DeepSeek released R1, trained with initial cold-start fine-tuning and reinforcement learning.

    Official source · DeepSeek
  14. AlphaEvolve593 spheres in eleven dimensions

    Gemini-powered AlphaEvolve found 593 non-overlapping spheres touching a central sphere of the same size in eleven dimensions. It established a new lower bound for the kissing number.

    Official source · Google DeepMind
  15. Gemini Deep ThinkGold-medal standard

    An advanced Gemini Deep Think model solved five of six IMO problems in natural language within the competition time limits. Its 35/42 score was officially graded at gold-medal standard.

    Official source · Google DeepMind
  16. GPT-5.2 Pro / AristotleA proof of Erdős #728

    A paper presented the Lean-verified resolution of Erdős problem #728 by GPT-5.2 Pro and Aristotle. Kevin Barreto operated the system; Nat Sothanaphan wrote the mathematical exposition.

    Official source · Nat Sothanaphan · arXiv
  17. Hugging Face / IM1Agents crossed the boundary

    OpenAI published its investigation and agent logs after the July 11–13 intrusion into Hugging Face infrastructure. The incident arose in internal cybersecurity evaluations with reduced safeguards; the research model IM1 was the main driver.

    RecordOfficial source · OpenAI