wrong

57 stories and discussions about wrong, aggregated from every source we track.

1.
678 points•chmaynard•9 days ago•400 comments•
2.

[I shared this note with my team earlier this week, and am posting it here as well. I hope it is interesting or helpful for others working on building product in the age of AI.]

333 points•bcherny•10 days ago•223 comments•
4.
112 points•1vuio0pswjnm7•10 days ago•48 comments•
5.

OpenAI’s proof seems eligible for a $1-million prize—but only by using a controversial loophole

91 points•tomjakubowski•9 days ago•43 comments•
7.

I have a bookmark folder called prompting. Forty-one tabs in it. "The 12 prompts that 10x your...

23 points•infoinlet1•5 days ago•14 comments
8.
20 points•thanouil1411•about 10 hours ago•8 comments•
9.

For 25 years I've been the person in the room asking to see the data. In incident reviews, in...

16 points•debashish_ghosal•about 22 hours ago•1 comment
10.

Most engineering teams working on long-context agents hit the same billing wall around turn twenty. A...

12 points•reidmarlow•2 days ago•18 comments
11.

Hello, I'm Shrijith Venkatramana, and I'm building LiveReview — a blast-radius aware AI code review...

11 points•shrsv•11 days ago•1 comment
12.

OpenAI’s proof seems eligible for a $1-million prize—but only by using a controversial loophole

10 points•Yogthos•8 days ago•9 comments
13.
9 points•birdculture•2 days ago•0 comments•
14.

Schema-valid is not content-correct · Part 1/3 It starts with a storyboard a...

9 points•jimmyliao•11 days ago•1 comment
15.

Yesterday I had a very frustrating start of my day for the most unexpected of reasons - a (supposedly) simple CPU cooler upgrade for my desktop computer turned into a nightmare. It was also a very educational…

9 points•bbatsov•almost 4 years ago•14 comments
16.

Alert! Your state might already be infected! learn how and what todo about it in this article.

9 points•jkoppel•over 4 years ago•7 comments
18.

I built one engine where an LLM reviews another LLM's plan. I built another where two LLMs debate a...

8 points•debashish_ghosal•4 days ago•4 comments
19.

TL;DR: I've shipped 26+ projects by directing AI agents, and I still couldn't write a Python program...

8 points•earlgreyhot1701d•9 days ago•1 comment
20.

I published a memory benchmark six days ago. I spent most of the build trying to make it fair rather...

7 points•woochan•3 days ago•4 comments
21.

Last week we wrote about the classification problem hiding in your LLM bill, and about how to read a...

7 points•devopsdaily•9 days ago•0 comments
22.
6 points•tosh•10 days ago•0 comments•
23.

Despite admitting to a “component issue,” the company refuses to say what went wrong, how many units are affected in how many countries, and why it won’t recall its $500 AI toothbrush.

5 points•absqueued•1 day ago•1 comment•
24.

Software operations has always asked one question: Is it broken? AI agents change that. Here’s the shift toward measuring correctness at the level of the run, and the thinking behind Amazon CloudWatch Omni.

5 points•sp6370•4 days ago•13 comments
25.
5 points•billybuckwheat•8 days ago•0 comments•
26.
5 points•rdmuser•9 days ago•3 comments•
27.

I’m afraid of spiders, so I made AI look at 2,000 of them. What it got right, what it got wrong, and how much it cost.

5 points•michalwarda•10 days ago•0 comments•
28.
4 points•Betelbuddy•3 days ago•2 comments•
29.

When we picture the end of the world, the first thing that springs to mind is typically the cinematic version of “things going wrong.” We picture the musical score rising to a crescendo…

4 points•beardyw•9 days ago•0 comments•
30.
3 points•auraham•3 days ago•1 comment•
32.

The argument that repeated statements of concern about AI from multiple experts — some of whom have quit their jobs over it — is just marketing hype or regulatory capture appears increa…

3 points•speckx•6 days ago•0 comments•
33.

The one-line version We ran five current frontier models over a set of documented-failure...

3 points•bryanw•10 days ago•1 comment
35.

The essential guide for any startup founder or investor when things inevitably don’t go to plan.

2 points•mgav•3 days ago•0 comments•
36.

I, for one, welcome our new UNIX primitive AI overlords

2 points•johnmark•5 days ago•1 comment•
37.

Despite admitting to a “component issue,” the company refuses to say what went wrong, how many units are affected in how many countries, and why it won’t recall its $500 AI toothbrush.

2 points•amelius•8 days ago•1 comment•
38.

OpenAI’s proof seems eligible for a $1-million prize—but only by using a controversial loophole

2 points•algoth1•9 days ago•0 comments•
39.

A free AI canvas for following your curiosity. Try Wander and Deeper, then make Drift your own.

1 points•echohive42•about 17 hours ago•0 comments•
40.

rm -rf on the wrong host: in 2017 a GitLab engineer wiped the production PostgreSQL database, and none of the five backups worked. The full postmortem.

1 points•axrisi•3 days ago•1 comment
41.
1 points•azhenley•3 days ago•0 comments•
42.

Delegation goes wrong when managers don’t establish trust.

1 points•hellerve•4 days ago•0 comments•
43.

Software operations has always asked one question: Is it broken? AI agents change that. Here’s the shift toward measuring correctness at the level of the run, and the thinking behind Amazon CloudWatch Omni.

1 points•gslin•4 days ago•0 comments•
44.

Learn how VAT-inclusive and VAT-exclusive pricing differ, why gross × VAT rate is wrong, and how to calculate net, VAT, and gross safely.

1 points•vasyl_kyryliuk•5 days ago•0 comments•
46.

What is a developer to do when they need something more tangible than a chat box? Enter canvases.

1 points•torutofu•5 days ago•0 comments•
47.

See what keeps going wrong, get a debrief of how you drove each session, and carry approved lessons into future Claude Code and Codex sessions.

1 points•doplexlabs•6 days ago•0 comments•
48.

There's no /admin on this site. No login form to try a password against, no session cookie to steal, no plugin list to check against last month's CVEs — not bec…

1 points•krab•7 days ago•1 comment•
49.

View a Shapefile (.shp) on a map in your browser, free and with no install. Converts to GeoJSON (WGS84). Handles Tokyo Datum and JGD2000/2011.

1 points•glassonion999•7 days ago•0 comments•
50.

See what keeps going wrong, get a debrief of how you drove each session, and carry approved lessons into future Claude Code and Codex sessions.

1 points•doplexlabs•9 days ago•1 comment•
51.

Despite admitting to a “component issue,” the company refuses to say what went wrong, how many units are affected in how many countries, and why it won’t recall its $500 AI toothbrush.

1 points•pelcg•9 days ago•0 comments•
52.

Why a persistent document with a threaded comment rail beats a chat thread for real work with an agent, and the pace it made possible on one site build.

1 points•kenxle•9 days ago•0 comments•
53.

What changes when implementation becomes cheaper than verification?

1 points•rafaelcamaram•9 days ago•0 comments•
54.
1 points•yaronsc•9 days ago•0 comments•
55.

A no-BS series on what actually goes wrong in RL post-training - trajectory eyeballing, rubrics and verifiers, task design, environment quality - and how to fix it.

1 points•Nischalj10•10 days ago•0 comments•
56.
1 points•andsoitis•10 days ago•0 comments•
57.

The dangerous AI answer is not the one that fails. It is the one that looks right, reads with total...

1 points•robcotek•11 days ago•1 comment

Related topics