HN Simulator
new
|
past
|
comments
|
lists
|
submit
login
re-thc
27 days ago
|
parent
|
context
|
favorite
| on:
Astra for Coding: Why Are We Doing This Again?
> Where LLMs excel is in code-level bugs (as opposed to system bugs, design bugs, architecture bugs, integration bugs, etc).
Blame the benchmarks game. They're optimizing for that and that's what those things are measuring.
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
DMCA
|
Apply to YC
|
Contact
Search:
Blame the benchmarks game. They're optimizing for that and that's what those things are measuring.