Need help?
<- Back

Comments (224)

  • dudeinhawaii
    Great site, triggered memories! haha.To try to add something to this discussion -- I think that while I've seen these sort of loops less --- what I have seen is "overly helpful".Models nowadays want to double-triple-quadruple check things. I'm being silly but it verges on "I have a working solution but let me write a variation in Rust to ensure a convergent solution and prove this works".I've had to stop models nowadays mostly because they're being agonizingly pedantic in their validation. Opus is actually one of the most pedantic and "off track" here. But again, not in a bad way. I'm usually like "stop testing latency between 50 runs of this app... this is version one.. we're going to make a million more changes.. you're not buying us anything".
  • dwedge
    I got way too annoyed at this before realising it was an optional game and I could just close the tab
  • _fat_santa
    At least with Codex, this has not been my experience at all. It still screws up sure, but in every case I can ask "why did you do this" and it can trace back what made it take that particular decision. Typically it's always that I either didn't specify the problem correctly or made a really dumb mistake (executing the task on the wrong project....did this one yesterday) or it's something within a skill file that instructs it (at which point I fixup the instructions).Once in a blue moon it's actually the model making a material error in it's thinking and I have to go back and redo it.
  • JohnMakin
    > Worth naming: the Add to cart button is still black.Got an audible guffaw out of me. This really is what the experience is like sometimes if you're just giving it a result without being specific in implementation, and it comes out of nowhere, some days much worse than others.I've become patient with it, but whatever this style of output is called or doing - it is both condescending and entirely unhelpful, and it seems designed to frustrate.
  • captainbland
    This is actually what keeps people using AI: variable reward schedule. It's basically gambling.
  • johnisgood
    > Why is half the site blue now? I asked you to change one button.> Half the site is blue. I asked for ONE button.Those are my only options when the site is clearly not blue, two buttons are.There is a reason for why I am much more specific than this.
  • kstenerud
    That's so weird... This doesn't at all match my experience with Claude. I've never seen it behave this way.
  • oujiii
    Haha this is spot on how I've been feeling lately. I find it unbearable to work with this model for this reason... any trick out there you can do to steer it not to overcomplicate things? I guess Codex here I come
  • xd1936
    Laughed out loud at the overly cautious Terms of Service that it generated for "Cyanide Blue", the color it made up
  • alentred
    -= CAUTION, SPOILERS =-This got me on "cyanide blue", and I was ROLLING ON THE FLOOR LAUGHING on "Approaching usage limit". I can barely stop laughing now and my stomach hurts. I mean, Thank You!
  • heaney-555
    I haven't experienced anything like this with Codex. Why do people stick with Claude Code if it doesn't do what you ask it to?
  • techscruggs
    I never really understood what being "triggered" was like until now.
  • andremendes
    I lost it when it finally did the right thing, but then it added a never-requested gradient to the button. Very good!
  • epistasis
    One note for those still using the Claude system for chats: there's no system to get generated images, spreadsheets, etc. out of the system. They claim it's a "security concern" to provide that data to you, as if they are protecting you by refusing to follow data export laws.I'm hesitant to email their data emails, as it's common for companies to delete all data upon any request, instead of providing data as they are required to.
  • josh_p
    a funny codex anecdote:I had 5.6-Luna coordinate a code review in which it spawns 2 agents looking for different things. My prompt was "review the currently checked out branch. diff target is `next`. The jira ticket is XX-XXXXX..." My `next` branch was a few commits behind `origin/next` but it still did its review against the stale local version instead of clarifying or inferring that I meant `origin/next`. The findings were very confusing until I realized what I did.I'm noticing the need to be really specific with any instructions lately, which I don't think is a bad thing. I expect co-workers (or anyone really) to tell me what they need in specific terms so I can get it right. I can do the same for the machine, I guess.
  • arbirk
    One thing I have to be honest about, and it's mine.. The one thing I would check before... do you want to do that? Say go an and will do it without the check While checking I found 3 vulnerabilities and 2 potential optimizations of which I fixed 2 and 1. Do you want me to file the other as issue, or stop for the day? We have done <lists a weeks worth of work> this morning. I feel you need a break
  • pablopudding
    I’m laughing and crying at the same time. This is what work feels like now. Thank you, well done!
  • inerte
    To be fair I’ve worked on human programmed systems where similar “it should be a half point story” requests would be met with snark by the engineers and take 2 sprints.I guess we are all PMs now.
  • snkline
    Seems to be getting a polarized response. I quite enjoyed the it, but I do think the creator should have made it clearer that a) it is in fact a joke site and b) it does not consist of actual Claude responses.It is easy to misinterpret this site, and therefore not "get" the joke.
  • qazxcvbnmlp
    Oh dear - this explains why people have bad experience with ai. The prompts in the game are terrible.The user was providing no context, they had no ability to give the model background or context. A simple why would have prevented 95% of these side quests. “Im trying to increase the relative visibility of the add to cart action on the page. Can we please change it to blue without changing any other buttons. /effort low. Let me know if you have questions and before editing anything tell me what you are going to do”
  • cropcirclbureau
    Is my job a joke to you??
  • syntaxing
    > 23 agents total.This hit a bit too close to home. Sol has the same issue, spawns a lot of agents for no good reasons (besides burning tokens).
  • gwbas1c
    I don't get who this is making fun of:- The people who won't make any effort to learn the tools, and something as simple as reverting code (via git) needs to be done by AI?- The awful programmers who we've had to endure working with, who are so bad at simple changes that they have negative productivity?- Or Claude itself?---BTW: I don't have these problems, but I'm also not afraid to do things myself when it's easier.Edit: If I want to change a button's color, I just change it manually. If I don't know where the code for the button is, I might start with prompting, (because AI can often find the code faster than I can,) and then once the diff is proposed, start adjusting things by hand.
  • chrismorgan
    I’ve never used any of these tools. Please tell me that this is a grossly exaggerated parody, and that the tools don’t write like this, or do so many ridiculous things. For my sanity.(I am genuinely uncertain, though I presume it’s at least somewhat exaggerated.)
  • Nevermark
    Funny exercise.For a moment I thought, wow, someone put a lot of work into creating this theme park of frustration.Next: It would be so easy to create a faux-Claude like this.Then: How hilarious to watch the transcripts of unsuspecting users in real time.Finally: I began wondering if this might be relevant to all the redundant, unnecessarily preambled, sentence structure complexifying, indirect referencing, canned phrasing, ambiguity mining, analogy maxxing, over-wordy responses I have recently been getting from Fable...
  • brap
    How do you manage your frustration in these interactions? I often find myself getting pissed off
  • robinpie
    If you interact with Claude like this and ignore legitimate issues it flags, no wonder you have a bad experience.
  • jmartrican
    Wow that gave me anxiety... lol. Ok cool so I'm not the only one who gets into these situations.
  • vinc
    You should plan the task before implementing it to make sure that it will do the right thing.
  • bdelmas
    It would have been funny a year ago but now I have no issues of that sort or even for more complex tasks
  • stevenhubertron
    You can hate AI all you want, but this is because of a bad design system, not because of a bad LLM.
  • fractorial
    Brilliant. Precisely the reason I stopped using Anthropic's products.
  • anon
    undefined
  • neilellis
    Congratulations!!! You win what’s left of the internet - just ask Claude for your prize! Motrin I’ve had this week.I use codex now.
  • monooso
    Oh god, it's so painfully accurate.
  • dannypostma
    This is scary close to my interaction with Claude this week.
  • satvikpendem
    It's funny but unrealistic as Claude does a pretty good job at only changing what is required these days with the 5 tier models like Opus 5 or Fable.
  • totetsu
    I was waiting for it to .. say usage limit reached after reverting it back to how you started..
  • pohl
    Amusing, but do people actually prompt in the style of any of the options given? All this for what is ultimately a PEBCAK error.
  • 8cvor6j844qw_d6
    I find it funny how it went off with subagents and adversarial review when a simple grep or diff is sufficient.
  • dsign
    That was funny :-)I use Opus and Sonnet 5 all the time and I find their language grating. But honest, I prefer to put up with it and get the results than to put up with my own human limitations and not get the results.
  • Kim_Bruning
    1970-01-01's Kobayashi Maru solution is the only thing that gave me closure :-P but unfortunately it's [dead].
  • anon
    undefined
  • Toutouxc
    This is so perfect and depressing that I might cry. It’s like a Kafka novel about programming.
  • Culonavirus
    This is the smoking gun.Yes, and it's mine!
  • apetresc
    So this site is just a fan-fiction that thinks it's somehow dunking on Claude? I've never had a session that remotely resembles any of this. I honestly can't tell what point this site thinks it's making.
  • yomismoaqui
    I don't get the joke... maybe because I'm using Codex?
  • andai
    I was expecting it to spend 30 minutes running headless chrome instances, taking screenshots and analyzing them in python to verify the blueness of the result.
  • rcfox
    A Bot & Costello
  • jadar
    This is so good at replicating the experience of frustration, then relief when it finally does what you asked it to do in the first place!
  • azalemeth
    I've experienced this so many times over."I was wrong" and "the honest truth" are just forever phrases that are now dead to me.
  • appleappleapple
    This spiked my blood pressure. Well done
  • telesilla
    This felt very late-90s net art. Stressful but nicely done satire.
  • bbstats
    Mine immediately did it correctly?
  • swiftcoder
    This is pure genius. No notes
  • mzajc
    > "`#16b8c4`. Yes. Apply it."> WebFetch en.wikipedia.org/…/Cyan> WebFetch en.wikipedia.org/…/Prussian_blue> WebFetch www.colorhexa.com/16b8c4Brilliant.
  • homeonthemtn
    Lord this triggered my eye twitch
  • anon
    undefined
  • vant
    glad to see I'm not the only one... anthropic needs to support my anger management treatment
  • akho
    is this using my subscription
  • chuckadams
    Not really my experience with Claude, and the prompts are not how I would phrase them, but still pretty damn funny: I especially loved the slot-machine-style picker for LLM-isms.
  • K0IN
    now add a source tab and lets see, how many ppl. will fix it themselves and how many turns it needs.
  • gitowiec
    This is kind of funny but with tears (not off joy). I stopped playing because it made me angry
  • almostdeadguy
    Nails the Claude dialect. Technical nonsense like:> I'm collapsing this back to the rendered outcome:And intermixed with SaaS product page idioms from a brain-damaged marketer like:> No broader cleanup.> No further architecture work.> Just the button.Aside from the patterns everyone knows like em-dashes, "its not X, it's Y", etc. I think the key features of claude diction is it sounds like a junior engineer over their skies who is trying to make up for that with extra verbiage mixed with extremely grating SaaS marketing-ese.
  • dmd
    Was this made by someone who hasn't actually used any of these tools in over a year?
  • jezzamon
    Funny game.Do people really prompt AI like this? Multiple times the choice was either to yell at the agent, or ask it why it did something, neither of which are very fruitful lines to go down if you know what you're doing
  • NickNaraghi
    You didn't say please or thank you.
  • airstrike
    This does not match reality at all, speaking as the #1 user on agent hours per clauderank.com
  • GracefullyShot
    it gave me headache in 2 turns, just like opus 5 !
  • anon
    undefined
  • improbableinf
    Thank you for creating this. Just thank you
  • kaoD
    Am I the only one whose experience doesn't match this?My gripe with Claude is that while investigating how to do this it will report 200 other incidental findings which I overlooked and I realize those are broken too and need urgent fixing, derailing me, not it.
  • tamimio
    This is gold, thanks for the giggles! I think it was designed that way to burn tokens.
  • burnoutdv
    Just when I came back to my pc and was thinking "I hate this world were everyone talks about AI like fanatics" this made me a little bit happy, especially the unhingend all caps options towards the end
  • felixgallo
    Here come all the totally organic "wow, I guess I better switch to OpenAI" comments.
  • anon
    undefined
  • nullbio
    Thanks, I hate it.
  • dazhbog
    PTSD 9000.. I miss the old days, less load bearing BS and more in the zone coding..
  • khernandezrt
    I mean honestly if you're using an agent for something this simple you deserve this and all the token usage that comes with it.
  • r_lee
    now THAT is a load-bearing simulation
  • bennettpompi1
    this is hilarious lmao
  • anon
    undefined
  • moralestapia
    Wow, this is SO on point.It made me stop using Claude at all. Codex has almost surgical precision, and I like that a lot.(But nowadays I just use DeepSeek Flash. it does screw up but its cents so ¯\_(ツ)_/¯).
  • AIorNot
    Lol this is great
  • random_cat_8745
    rofl, brilliant
  • teekert
    [dead]
  • arrowsmith
    [dead]
  • 1970-01-01
    [dead]
  • anon
    undefined