AI Interview Prep

AI Interview Prep

LLM System Design Interview #11 - The Alignment Tax

When RL makes your model obedient but dumb - and how a KL divergence leash keeps it smart, creative, and aligned.

Hao Hoang's avatar
Hao Hoang
Nov 09, 2025
∙ Paid

You’re in an AI Engineer interview at OpenAI and the interviewer asks:

“You’ve successfully fine-tuned a model with RL. It’s now excellent at following instructions, but it’s become ‘dumber’ at general knowledge and creative writing. What is this phenomenon called, and what specific term would you add to your loss function to prevent this?”

Thanks for rea…

User's avatar

Continue reading this post for free, courtesy of Hao Hoang.

Or purchase a paid subscription.
© 2026 Hao Hoang · Privacy ∙ Terms ∙ Collection notice
Start your SubstackGet the app
Substack is the home for great culture