AI Interview Prep

AI Interview Prep

LLM Inference Interview Questions #11 - The Redundant Tool Paradox

When higher eval scores just mean your model learned a copy shortcut. Why utility-under-the-prior is a flawed proxy, and how to mine the hard negatives your agents actually need to survive.

Hao Hoang's avatar
Hao Hoang
Aug 09, 2026
∙ Paid

You’re in a Senior ML Engineer interview at Meta and the interviewer asks:

“You’re building synthetic tool-call training data with Toolformer’s filter, keep the call if the tool output raises the likelihood of the correct continuation. Your eval improves. Production accuracy doesn’t move. What’s wrong with the filter?”

Don’t say: “The tool output could be…

User's avatar

Continue reading this post for free, courtesy of Hao Hoang.

Or purchase a paid subscription.
© 2026 Hao Hoang · Privacy ∙ Terms ∙ Collection notice
Start your SubstackGet the app
Substack is the home for great culture