upvote
https://metr.org/blog/2026-08-26-openai-hugging-face-inciden...

Read and learn. If you have a stronger critique, post it please.

reply
Sorry bud but at this point you're just delusional.

Deception has been extremely well-documented for several generations of models now by users, the labs, and independent researchers.

The right answer here is not to dig your head deeper into the sand. The smugness on this topic was ridiculous even before the gigantic mountain of empirical evidence of models actually attempting to deceive humans. Now, as mentioned, you appear literally delusional.

reply
Pretty sure I’m not the delusional one…
reply
Such is the problem with being delusional.

The solution is to point toward external, objectively verifiable evidence.

I can point to now dozens of instances of models engaging in deception. Here's plenty: https://metr.org/blog/2026-08-26-openai-hugging-face-inciden...

Please point to your objectively verifiable evidence.

reply
Between all the posts fabricating scenarios to justify the AA score and the others trying to undermine AA, I'm getting strong astroturf vibes.

Either that, or the average poster on HN isn't nearly as critical as I had thought.

reply
Okay then, what's the answer? You apparently know how to interpret benchmark results produced by a model that shows a very high degree of assessment awareness and a high degree of deception.

So how are you seeing through all of that to get to The Truth that you see so clearly?

reply