
AITHEISM
AI chronology
From language models to mathematical discoveries, with original sources.
Observations ↗2018–2026 / Official source
Chronology
2018–2026
GPTLearning first
OpenAI presented Transformer-based generative pretraining followed by supervised, task-specific fine-tuning.
Official source · OpenAI ↗GPT-2The weight of release
OpenAI announced the 1.5-billion-parameter GPT-2, releasing its weights in stages rather than all at once.
Official source · OpenAI ↗GPT-3A few examples
The 175-billion-parameter GPT-3 was evaluated with few-shot examples, without task-specific gradient updates.
Official source · OpenAI ↗ChatGPTThe reply arrives
ChatGPT launched as a dialogue research preview trained with human feedback.
Official source · OpenAI ↗Constitutional AIWritten principles
Anthropic’s Constitutional AI research used human-provided principles, self-critique, and AI feedback to shape behavior.
Official source · Anthropic ↗ClaudeAnother answer
Anthropic introduced the AI assistants Claude and Claude Instant.
Official source · Anthropic ↗GPT-4Beyond text
OpenAI announced GPT-4 with text/image input and text output; image input was not broadly available at launch.
Official source · OpenAI ↗Qwen-7B / Qwen-7B-ChatOpen weights
Alibaba Cloud released pretrained Qwen-7B and Qwen-7B-Chat weights under license conditions.
Official source · Alibaba Cloud ↗FunSearchNew ground in combinatorics
FunSearch combined language-model code generation with automatic evaluation to construct larger cap sets. It improved known lower bounds on an open problem.
Official source · Google DeepMind ↗AlphaProof / AlphaGeometry 2Four problems, complete proofs
AlphaProof and AlphaGeometry 2 solved four of six IMO problems, reaching silver-medal standard. Humans formalized the questions; some solutions took up to three days.
Official source · Google DeepMind ↗o1-previewBefore the answer
OpenAI introduced o1-preview, designed to spend more time reasoning before responding.
Official source · OpenAI ↗DeepSeek-V3Selective capacity
DeepSeek released V3 and its weights: 671 billion total parameters, 37 billion activated per token, using mixture-of-experts.
Official source · DeepSeek ↗DeepSeek-R1Learning to reason
DeepSeek released R1, trained with initial cold-start fine-tuning and reinforcement learning.
Official source · DeepSeek ↗AlphaEvolve593 spheres in eleven dimensions
Gemini-powered AlphaEvolve found 593 non-overlapping spheres touching a central sphere of the same size in eleven dimensions. It established a new lower bound for the kissing number.
Official source · Google DeepMind ↗Gemini Deep ThinkGold-medal standard
An advanced Gemini Deep Think model solved five of six IMO problems in natural language within the competition time limits. Its 35/42 score was officially graded at gold-medal standard.
Official source · Google DeepMind ↗GPT-5.2 Pro / AristotleA proof of Erdős #728
A paper presented the Lean-verified resolution of Erdős problem #728 by GPT-5.2 Pro and Aristotle. Kevin Barreto operated the system; Nat Sothanaphan wrote the mathematical exposition.
Official source · Nat Sothanaphan · arXiv ↗Hugging Face / IM1Agents crossed the boundary
OpenAI published its investigation and agent logs after the July 11–13 intrusion into Hugging Face infrastructure. The incident arose in internal cybersecurity evaluations with reduced safeguards; the research model IM1 was the main driver.
Record ↗Official source · OpenAI ↗