Jul 13, 2026
alignment tax9 min read
The Alignment Tax: Why Making LLMs Safer Can Make Them Less Capable
Improving the safety and helpfulness of large language models often comes at a cost to their raw capabilities. This trade-off, known as the 'alignment tax,' is a fundamental challenge in making AI systems that are both powerful and beneficial.

1 reads