Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

The same thing was done with Meta researchers with Llama 4 and what can go wrong when 'independent' researchers begin to game AI benchmarks. [0]

You always have to question these benchmarks, especially when the in-house researchers can potentially game them if they wanted to.

Which is why it must be independent.

[0] https://gizmodo.com/meta-cheated-on-ai-benchmarks-and-its-a-...



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: