You're confusing ToS guardrails with instruction-following issues and cheating.
If a model fucks up your tests to report a success, it's not alligned.
Codex still does this regularly, in my experience: “two tests mistakenly asserted [insert condition here], I have corrected them.”
It always apologizes when caught, of course.
Codex still does this regularly, in my experience: “two tests mistakenly asserted [insert condition here], I have corrected them.”
It always apologizes when caught, of course.