KARADAVI

How ChatGPT was built

ChatGPT changed the world in two months. The technical story of GPT, RLHF, InstructGPT, and the product decisions that launched an era.

Published Invalid Date min read
OpenAI
Editorial overview and documentation for OpenAI.Source: KARADAVI Knowledge Archive

From GPT-1 to GPT-4

GPT-1 (2018, 117M params). GPT-2 (2019, 1.5B): OpenAI initially withheld it. GPT-3 (2020, 175B): emergent few-shot learning. GPT-4 (2023, multimodal): new benchmarks in reasoning and coding, the backbone of ChatGPT.

The RLHF breakthrough

Raw GPT-3 was powerful but unruly. InstructGPT (2022) applied RLHF: human raters ranked outputs, training a reward model. The LLM was then fine-tuned to maximise that reward — producing instruction-following, safer responses.

The product launch

ChatGPT launched November 30, 2022 as InstructGPT in a chat interface. Within 5 days: 1 million users. Within 2 months: 100 million — the fastest product adoption in history.

What made it different

Earlier chatbots were scripted or retrieval-based. ChatGPT generated free-form text on any topic with coherence and breadth no previous system matched. The chat format was immediately intuitive.

The era it created

ChatGPT triggered an AI Cambrian explosion — every major company launched LLM products, billions flowed into AI infrastructure, and it raised profound questions about education, knowledge, and labour.

Explore Further

Continue exploring the forest
Next trailOpenAI

Every article leads somewhere. Follow this entity, or search the whole forest.