| | How well do agents use test/verification techniques? (danluu.com) |
| 5 points by tosh 8 hours ago | past | discuss |
|
| | How accurate have Ed Zitron's AI skeptic predictions been? (danluu.com) |
| 876 points by jatins 6 days ago | past | 1055 comments |
|
| | Bug Blindness (danluu.com) |
| 407 points by davidmckenna 8 days ago | past | 287 comments |
|
| | Developer Hiring and the Market for Lemons (danluu.com) |
| 2 points by rzk 13 days ago | past | 2 comments |
|
| | Bad benchmarks and evals: Senior SWE-Bench, napkin math, and winter tires (danluu.com) |
| 1 point by luu 13 days ago | past | discuss |
|
| | There's no reason for software to be slow anymore (danluu.com) |
| 669 points by Jach 16 days ago | past | 499 comments |
|
| | HN: The Good Parts (2016) (danluu.com) |
| 82 points by adletbalzhanov 17 days ago | past | 23 comments |
|
| | Bad benchmarks and evals: Senior SWE-Bench, napkin math, and winter tires (danluu.com) |
| 2 points by yosefk 19 days ago | past |
|
| | The Benchmarkpocalypse (danluu.com) |
| 183 points by cyndunlop 20 days ago | past | 64 comments |
|
| | Agentic test processes, LLM benchmarks, and other notes (danluu.com) |
| 2 points by kqr 27 days ago | past |
|
| | What's the best programming language for coding agents? (danluu.com) |
| 262 points by chaychoong 28 days ago | past | 188 comments |
|
| | How do programming languages impact token efficiency and correctness? (danluu.com) |
| 18 points by matt_d 28 days ago | past | 1 comment |
|
| | Exercises in benchmarking and evals, part 7: DeepSWE, Senior SWE-Bench (danluu.com) |
| 3 points by ndr 40 days ago | past |
|
| | Exercises in benchmarking and evals, part 7 (danluu.com) |
| 1 point by janvdberg 41 days ago | past |
|
| | Agentic test processes, LLM benchmarks, and other notes on agentic coding (danluu.com) |
| 17 points by bathtub365 43 days ago | past | 1 comment |
|
| | I could do that in a weekend (2016) (danluu.com) |
| 20 points by ntumlin 49 days ago | past | 4 comments |
|
| | Agentic test processes, LLM benchmarks, and other notes from Galapagos Island (danluu.com) |
| 2 points by gnyeki 49 days ago | past |
|
| | Agentic test processes, LLM benchmarks, notes on agentic coding from Galapagos (danluu.com) |
| 1 point by FabHK 59 days ago | past |
|
| | Why don't schools teach debugging? (2014) (danluu.com) |
| 2 points by bmacho 59 days ago | past |
|
| | Agentic test processes, LLM benchmarks, and other notes on agentic coding fr (danluu.com) |
| 20 points by lifeisstillgood 61 days ago | past | 2 comments |
|
| | Agentic test processes, LLM benchmarks, and other notes on agentic coding (danluu.com) |
| 1 point by ndr 61 days ago | past |
|
| | Agentic test processes, LLM benchmarks (danluu.com) |
| 2 points by eatonphil 65 days ago | past |
|
| | Diseconomies of scale in fraud, spam, support, and moderation (danluu.com) |
| 4 points by Liriel 65 days ago | past | 3 comments |
|
| | Agentic coding notes (danluu.com) |
| 179 points by gm678 65 days ago | past | 83 comments |
|
| | 95%-ile isn't that good (2020) (danluu.com) |
| 2 points by bmacho 68 days ago | past | 1 comment |
|
| | Suspicious Discontinuities (2020) (danluu.com) |
| 278 points by tosh 72 days ago | past | 101 comments |
|
| | Against essential and accidental complexity (2020) (danluu.com) |
| 1 point by pramodbiligiri 82 days ago | past |
|
| | Are closed social networks inevitable? (2010) (danluu.com) |
| 4 points by downbad_ 4 months ago | past | 1 comment |
|
| | Dunning-Kruger and other memes (2015) (danluu.com) |
| 4 points by downbad_ 4 months ago | past | 1 comment |
|
| | Integer Overflow Checking Cost (danluu.com) |
| 40 points by iwsk 4 months ago | past | 15 comments |
|
|
| More |