Discussion about this post

User's avatar
Joe Friday's avatar

The validation you flag is the crux, and there is a second-order version worth naming. Once this proxy becomes the yardstick feed designers optimize against, it can stop measuring what it measured in training: the on-platform behavior that tracked lower polarization gets produced directly, without the underlying change. That is the Goodhart problem, and it tends to arrive right after a cheap metric does, so periodic re-validation against fresh surveys may matter more than the initial fit.

No posts

Ready for more?