Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

any benchmark where opus 5 achieves higher scores than fable 5 in any way is not a benchmark worth trusting.


Why would Anthropic trust and use these tests in their official comparisons?


username checks out


do you feel you're free of bias and predisposition in saying this


great username lol




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: