Skip to content
The paper library

Read the papers. Skip the jargon.

The white papers that actually moved the field, each linked to the original and paired with my plain-language summary, a glossary of the terms, and what it changes in practice.

AllArchitectureScalingReasoningAlignmentAgents

02

Alignment2022

Training Language Models to Follow Instructions with Human Feedback

Ouyang et al.

Turned a next-token predictor into an assistant, and found a far smaller aligned model was preferred over a much larger unaligned one.

All original papers remain the work of their authors and link to the official source (arXiv or publisher). Summaries and glossaries on this site are my own commentary.