Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
irishcoffee
28 days ago
|
parent
|
context
|
favorite
| on:
GLM-5.3 (open-weight) beat Anthropic/OpenAI models...
The whole concept is kind of silly. We don’t “benchmark” humans. Or do we, via standardized tests? Why don’t we just use those? Or is that what the benchmarks are? I have no idea.
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: