Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
|
mustaphah's submissions
login
1.
The Little Book of Reinforcement Learning
(
github.com/alxndrtl
)
213 points
by
mustaphah
56 days ago
|
past
|
26 comments
2.
Opportunity cost neglect (2009) [pdf]
(
ufl.edu
)
2 points
by
mustaphah
57 days ago
|
past
3.
Economic Possibilities for our Grandchildren (1931)
(
fermatslibrary.com
)
4 points
by
mustaphah
59 days ago
|
past
|
1 comment
4.
A field guide to Fable: finding your unknowns
(
twitter.com/trq212
)
1 point
by
mustaphah
69 days ago
|
past
|
1 comment
5.
Ten Takeaways from the AI Engineering Report 2026: The Acceleration Whiplash
(
faros.ai
)
2 points
by
mustaphah
74 days ago
|
past
6.
The cost YAGNI was never about
(
kentbeck.com
)
6 points
by
mustaphah
75 days ago
|
past
|
1 comment
7.
Writing code vs. shipping code [pdf]
(
nber.org
)
3 points
by
mustaphah
3 months ago
|
past
8.
Trust Factory
(
kentbeck.com
)
7 points
by
mustaphah
3 months ago
|
past
9.
LLMs pass a standard three-party Turing test
(
pnas.org
)
3 points
by
mustaphah
3 months ago
|
past
|
1 comment
10.
The small sample trap in A/B testing
(
hadid.dev
)
4 points
by
mustaphah
3 months ago
|
past
|
1 comment
11.
Tell HN: Claude two rate limits don't know about each other
2 points
by
mustaphah
6 months ago
|
past
12.
Enhancing gut-brain communication reversed cognitive decline in aging mice
(
stanford.edu
)
386 points
by
mustaphah
6 months ago
|
past
|
185 comments
13.
Many SWE-bench-Passing PRs would not be merged
(
metr.org
)
278 points
by
mustaphah
6 months ago
|
past
|
153 comments
14.
AGI is an unscientific myth
(
tandfonline.com
)
4 points
by
mustaphah
6 months ago
|
past
|
2 comments
15.
Web Verbs
(
github.com/nlweb-ai
)
1 point
by
mustaphah
6 months ago
|
past
16.
OpenAI's 5-month experiment: building a product with no human-written code
(
openai.com
)
2 points
by
mustaphah
6 months ago
|
past
17.
SkillsBench: Benchmarking how well agent skills work across diverse tasks
(
arxiv.org
)
364 points
by
mustaphah
6 months ago
|
past
|
171 comments
18.
Evaluating AGENTS.md: are they helpful for coding agents?
(
arxiv.org
)
232 points
by
mustaphah
6 months ago
|
past
|
161 comments
19.
Curosr: Expanding our long-running agents research preview
(
cursor.com
)
3 points
by
mustaphah
6 months ago
|
past
20.
Measuring Time Horizon Using Claude Code and Codex
(
metr.org
)
1 point
by
mustaphah
6 months ago
|
past
21.
SWE-ContextBench: context learning benchmark in coding
(
arxiv.org
)
1 point
by
mustaphah
6 months ago
|
past
22.
SWE-AGI: benchmarking spec-driven software construction
(
arxiv.org
)
1 point
by
mustaphah
7 months ago
|
past
|
1 comment
23.
Code Formatting Silently Consumes Your LLM Budget
(
arxiv.org
)
1 point
by
mustaphah
7 months ago
|
past
24.
Agent Trace by Cursor: open spec for tracking AI-generated code
(
agent-trace.dev
)
1 point
by
mustaphah
7 months ago
|
past
25.
METR releases Time Horizon 1.1 with 34% more tasks
(
metr.org
)
1 point
by
mustaphah
7 months ago
|
past
26.
Coffee timing isn't one-size-fits-all
(
examine.com
)
4 points
by
mustaphah
7 months ago
|
past
27.
ChatGPT subscription support in Kilo Code
(
kilo.ai
)
1 point
by
mustaphah
7 months ago
|
past
28.
Imposter Syndrome Predicts Perfectionism
(
psypost.org
)
2 points
by
mustaphah
7 months ago
|
past
29.
Motivation acts as a camera lens that shapes how memories form
(
psypost.org
)
2 points
by
mustaphah
7 months ago
|
past
30.
Claude Code: Merging Slash Commands into Skills
(
x.com
)
2 points
by
mustaphah
7 months ago
|
past
|
2 comments
More
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: