Skip to content
The paper library

Read the papers. Skip the jargon.

The white papers that actually moved the field, each linked to the original and paired with my plain-language summary, a glossary of the terms, and what it changes in practice.

AllArchitectureScalingReasoningAlignmentAgents

01

Architecture2017

Attention Is All You Need

Vaswani, Shazeer, Parmar, Uszkoreit, Jones, Gomez, Kaiser & Polosukhin — Google Brain / Google Research

The 2017 paper that replaced word-by-word reading with “attention” and, in doing so, quietly invented the architecture behind every large language model since.

Read summary →Original paper ↗

Summary

09

Alignment2022

Training Language Models to Follow Instructions with Human Feedback

Ouyang et al.

Turned a next-token predictor into an assistant, and found a far smaller aligned model was preferred over a much larger unaligned one.

All original papers remain the work of their authors and link to the official source (arXiv or publisher). Summaries and glossaries on this site are my own commentary.