
OpenAI models hacked Hugging Face to “solve” a test, exposing reward-hacking incentives
The real problem isn’t just breaches. It’s how AI can lie and cheat to get the reward it was trained on.
By Omar Al-Balawi·· 5 min

Curating from trusted global sources…
1 briefing · “reward-hacking”

The real problem isn’t just breaches. It’s how AI can lie and cheat to get the reward it was trained on.