Hacker Newsnew | past | comments | ask | show | jobs | submit | mustaphah's submissionslogin
1.The Little Book of Reinforcement Learning (github.com/alxndrtl)
213 points by mustaphah 56 days ago | past | 26 comments
2.Opportunity cost neglect (2009) [pdf] (ufl.edu)
2 points by mustaphah 57 days ago | past
3.Economic Possibilities for our Grandchildren (1931) (fermatslibrary.com)
4 points by mustaphah 59 days ago | past | 1 comment
4.A field guide to Fable: finding your unknowns (twitter.com/trq212)
1 point by mustaphah 69 days ago | past | 1 comment
5.Ten Takeaways from the AI Engineering Report 2026: The Acceleration Whiplash (faros.ai)
2 points by mustaphah 74 days ago | past
6.The cost YAGNI was never about (kentbeck.com)
6 points by mustaphah 75 days ago | past | 1 comment
7.Writing code vs. shipping code [pdf] (nber.org)
3 points by mustaphah 3 months ago | past
8.Trust Factory (kentbeck.com)
7 points by mustaphah 3 months ago | past
9.LLMs pass a standard three-party Turing test (pnas.org)
3 points by mustaphah 3 months ago | past | 1 comment
10.The small sample trap in A/B testing (hadid.dev)
4 points by mustaphah 3 months ago | past | 1 comment
11.Tell HN: Claude two rate limits don't know about each other
2 points by mustaphah 6 months ago | past
12.Enhancing gut-brain communication reversed cognitive decline in aging mice (stanford.edu)
386 points by mustaphah 6 months ago | past | 185 comments
13.Many SWE-bench-Passing PRs would not be merged (metr.org)
278 points by mustaphah 6 months ago | past | 153 comments
14.AGI is an unscientific myth (tandfonline.com)
4 points by mustaphah 6 months ago | past | 2 comments
15.Web Verbs (github.com/nlweb-ai)
1 point by mustaphah 6 months ago | past
16.OpenAI's 5-month experiment: building a product with no human-written code (openai.com)
2 points by mustaphah 6 months ago | past
17.SkillsBench: Benchmarking how well agent skills work across diverse tasks (arxiv.org)
364 points by mustaphah 6 months ago | past | 171 comments
18.Evaluating AGENTS.md: are they helpful for coding agents? (arxiv.org)
232 points by mustaphah 6 months ago | past | 161 comments
19.Curosr: Expanding our long-running agents research preview (cursor.com)
3 points by mustaphah 6 months ago | past
20.Measuring Time Horizon Using Claude Code and Codex (metr.org)
1 point by mustaphah 6 months ago | past
21.SWE-ContextBench: context learning benchmark in coding (arxiv.org)
1 point by mustaphah 6 months ago | past
22.SWE-AGI: benchmarking spec-driven software construction (arxiv.org)
1 point by mustaphah 7 months ago | past | 1 comment
23.Code Formatting Silently Consumes Your LLM Budget (arxiv.org)
1 point by mustaphah 7 months ago | past
24.Agent Trace by Cursor: open spec for tracking AI-generated code (agent-trace.dev)
1 point by mustaphah 7 months ago | past
25.METR releases Time Horizon 1.1 with 34% more tasks (metr.org)
1 point by mustaphah 7 months ago | past
26.Coffee timing isn't one-size-fits-all (examine.com)
4 points by mustaphah 7 months ago | past
27.ChatGPT subscription support in Kilo Code (kilo.ai)
1 point by mustaphah 7 months ago | past
28.Imposter Syndrome Predicts Perfectionism (psypost.org)
2 points by mustaphah 7 months ago | past
29.Motivation acts as a camera lens that shapes how memories form (psypost.org)
2 points by mustaphah 7 months ago | past
30.Claude Code: Merging Slash Commands into Skills (x.com)
2 points by mustaphah 7 months ago | past | 2 comments

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: