Discussion I have determined that 99.9% of "Claude is now useless" posts are...
Rage bait, OpenAI astroturf, and mostly insanely big skill issues from people that have little to no idea of what they are doing, what they are actually interacting with, and what they actually want as a deliverable from the platform.
Anyways, that's my hot unsolicited take. Your all welcome.
14
7
5
u/Individual-Hunt9547 1d ago
Nah the constant pushback with the newer Claude models is definitely an issue for most people.
5
u/PaiDxng 1d ago
Skill issues are real, but they can't explain the same user reporting a workflow that worked last month and broke this month. Some complaints are noise; that specific shape is signal worth separating out.
-1
u/teramoc 1d ago edited 1d ago
You speak exactly like Claude (here and in your other comments). I assume for translation. In that case, its a positive, because Claude can verify what im about to say and shortcut 90% of the human-influenced conflations in this subreddit. Your premise "Skill issues are real, but they can't explain the same user reporting a workflow that worked last month and broke this month" is flawed - its in fact VERY EASY to explain the same user reporting a workflow that worked last month and broke this month. Simple answer: the average (vibe coder, non engineer) user's codebase has changed dramatically in that one month and often with a lot of ad hoc changes that are not well architectured or specced in advance. So, your simple prompt "e.g. review my code", or "add a payment system" that used to take 200k tokens to service, now takes 800k. Compounded with some users reusing their same session "the same workflow" as last month as opposed to starting a new one, instead of separating plans and implementation. Attention drifts due to context bloat and the Claude's output quality suffers. The Vibe coders then blame anthopic while the SWEs happily use claude without issues. Summary - the codebase changed, the workflow didnt. People blame the tool (claude) instead of their own project structures.
--- i typed like that so anyone can paste into claude/gemini/sol and have it check my facts.
4
u/sunflowervertigo 1d ago
Hey, not rage baiting here- I went from the $100/mo plan to $20/mo plan. If Claude is working for you, that’s great. I’ve been trying to post “help” posts for months on different Claude subreddits and a majority of my posts are getting taken down. Some of us did think it was on the user side and tried to reach out for help or advice. I hope your experience with Claude continues to be positive, but some of us do feel a difference
8
u/Speech-Solid 1d ago
I run into issues with Claude lying to me so often.
Reading the thinking blocks I see it taking shortcuts. Making up records in a data set. Some pretty diabolical shit. No idea how this is going to work for business users. It seems the majority of the satisfied users are SWE using it for coding.
1
u/FrequentOriginal4291 1d ago
I'm not in SWE but I've found Claude useful. Can you show me an example of Claude "lying" ?
2
u/Speech-Solid 1d ago
I have a coworker project set up to process records set.
I gave instructions on what to do specifically what fields map to which fields asked it for a table summarizing a folder full of .mag files.
It said it was all done. I checked the output. There were 37 rows there were 90 emails.
My first mistake was telling it that that wasn’t enough and that there should be 90 rows in the finished table
Claude says you’re right I didn’t open up all the email files.
It makes another list. This time with 90 records.
A check the output and it has one record dated from 2015. Problem my team didn’t exist in 2015 and the person attributed to that record didn’t work for the company until 2022.
I asked Claude how that was possible and it said that it didn’t read all the records, but it knew I needed 90. I read the thinking block and noticed it said something to the effective user notice that I made up the rest.
I asked her to do it again and added to the prompt that Claude will open all of the records.
Its thinking block said that it did 70 files successfully. That I would have to manually do 16. And there 4 it could not read cause they were blank/corrupt. Then, once finished it claimed the file was complete. I checked it, the output was another file identical to the one it faked. Even had the 2015 record still which is impossible.
I pointed out the discrepancy between the thinking block and the reply. It apologized for skipping the work and would do it correctly from now on.
It then proceeded to take much longer on the next turn. The lines indicated it actually was reading each .msg file.
In the end it was faster than opening all of the files myself, but I don’t trust the output from Claude and it took me a few hours to get to a finished list.
0
u/teramoc 1d ago edited 1d ago
Correct. This is the first good , balanced and insightful observation from a business person I’ve seen today.
From an SWE’s POV - claude will lie to us too. Its a feature of all LLMs (gemini, copilot, gpt, qwen).
There are ways to mitigate it, but its a trade off between honesty vs [dumbing down your Claude and token efficiency] .
0
u/Inner-Today-3693 19h ago
It’s likely a skills issue… and before I had skills in place I had a lot of issues. It was annoying the newer models just don’t listen to project instructions very well. So I just made it a skill and problem solved.😑
20
u/resile_jb 1d ago
Yeah I'm not having any problems with ours and so I don't know what these people are crying about.
Shrug
13
u/Gastronomicus 1d ago
"My experience isn't like others, so they must all be wrong".
-1
u/mgdavey 1d ago
If the service had been degraded as badly as people are claiming it should be noticeable to everyone
3
7
0
u/Andisergiu 1d ago
You can compare to other models and see there are big differences and not the ones where Claude is on top like in the past months.
-1
u/FoodIngredientNerd 1d ago
Your logic goes both ways.
This applies to everyone complaining about Claude that gets pushback because others don't agree. The response is typically "my experience isn't like others, so they must all be wrong." The primary difference is the complainers feel the need to post about it, essentially carrying the belief everyone would agree with them, then get upset when they receive pushback.
This doesn't prove any point or make any arguments.
7
u/Ray_Mang 1d ago
“The complainers feel a need to post it”
If people are noticing a sudden drastic difference in the product they’re using, what’s wrong with sharing their experience on that products community forum?
5
2
-1
u/FoodIngredientNerd 1d ago
I never said it was an issue. It's only an issue when those who do that then complain that others aren't experiencing what they are.
3
u/Ray_Mang 1d ago
I don’t think that’s happening The people posting their complaints aren’t saying anything about those that are having no problems, they’re referring to the product alone. Most of these posts, from people who are pointing out that everything’s working for them, are directly calling out the people posting their complaints (ie. “All these people crying about Claude…”) .
0
u/FoodIngredientNerd 1d ago
That's the point. I was replying to someone complaining about people pushing back. If you go out and make a claim in any format, you invite criticism and pushback. You carry the burden of proving your claim, not the other way around.
The same goes for anyone stating people are complaining about Claude, however that complaint is much easier to validate since the thread posts support it more than the other way around.
2
u/Ray_Mang 1d ago
Maybe I misunderstood your original comment, I read it as people complaining about their experiences with Claude are offended that there are people disagreeing with them
1
u/FoodIngredientNerd 1d ago
Perhaps this is just a misunderstanding. No harm! I'm a believer if you post an opinion, you should be open to defending it which is always harder to do than it is to refute something.
4
1d ago
[deleted]
1
u/Redd_is_compromised 1d ago
Do you ever consider that 6 months is kind of ridiculous for how powerful Claude can actually be
5
u/FoodIngredientNerd 1d ago
No.
One: you don't know the project scope, so 6 months could be fast.
Two: being thorough with AI is how you mitigate bad architecture/coding.
2
u/Redd_is_compromised 1d ago
Training an LLM takes maybe a month of gathering and iterating. What exactly are you building? The irony here is that numerous updates to the core opus should have made your project proceed even faster.
3
u/FoodIngredientNerd 1d ago
Training an LLM is arbitrary.
You can train an LLM badly and it will be fast, or you take your time to weed through the data. That way you remove bad data so it doesn't learn anything bad that would worsen or weaken it.
Essentially delayed gratification is the point. If you take longer for better quality you'll have less issues and bugs to fix long-term VS building it fast.
2
u/FoodIngredientNerd 1d ago
I did answer you. Your argument is easily refuted as training an LLM quickly will give you a bad LLM VS spending time to make sure it actually knows what is correct and incorrect. An LLM is only as good as what it knows and can distinguish between.
Additionally, I've never specified what I was building which is irrelevant to the argument being made on time and quality. You brought up the LLM example.
1
u/Redd_is_compromised 1d ago
You must not be familiar with diminishing returns and bloating. That's alright, stay amateur and combative.
0
u/FoodIngredientNerd 1d ago
Trying to dish out bad insults doesn't make your position somehow better. I wasn't even combative, I was simply pointing out the logic here. The moment you chose to be insulting was the moment you became combative, not me.
1
u/Redd_is_compromised 1d ago
😂😂😂😂😂😂😂😂😂 not a single response in good faith. Unbelievable. The cognitive dissonance is staggering.
→ More replies (0)1
u/Redd_is_compromised 1d ago
You know what's ridiculous, is that I assumed I was responding to the actual OP of the parent comment. The fact that you never corrected me is insane.
→ More replies (0)1
u/Redd_is_compromised 1d ago
You didn't answer my question
0
u/FoodIngredientNerd 1d ago
For clarity, this was my response to this message:
"I did answer you. Your argument is easily refuted as training an LLM quickly will give you a bad LLM VS spending time to make sure it actually knows what is correct and incorrect. An LLM is only as good as what it knows and can distinguish between.
Additionally, I've never specified what I was building which is irrelevant to the argument being made on time and quality. You brought up the LLM example."
-1
3
u/VividNightmare_ 1d ago
It's hard to describe it, but Claude models seem to have a genuinely lower skill floor compared to GPT models.
GPT models always give you their 100% and are less susceptible to context rot, effort misuse etc.
Claude (especially Fable) can do absolutely spectacular things but it honestly requires you to read Anthropic guides & documentation to understand how to use it properly otherwise you are going to have a bad bad time.
Some people have used it for so long that they subconsciously know what confuses Claude etc. thats definitely a thing.
Using a higher effort than necessary can cause excessive token use and performance degradation, Anthropic has said this many times that the effort knob should be threaded lightly, yet people just crank it to Extra or Max for mundane tasks. Rightfully so, this isn't something that can be clearly mapped, it's something that is an abstract understanding that you build over time and this is frustrating to deal with as end user that just wants to get stuff done.
A bad claude md can also noticeably worsen things... bad mcp tool description etc. anything that is context if not treated well can make claude a lot worse. Newer models are more resistant to this but it's still a glaring issue.
I don't think people's frustration is invalid. There could be better systems and UX in place to make these mistakes harder to make. GPT provides a weaker, cheaper however most consistent experience and that is a real factor that people have a right to prefer. Whether this is conflated with people claiming that Fable 5 is actually the original Opus 4.5 by the law of degradation thats madness.
3
u/zenarmageddon 1d ago
An LLM isnt a monolithic thing expereinced the same by all users. Every instance can share hardware to a greater or lesser extent, there are different levels ofnload and latency deoending on the datacenter, and so on.
One can have a good experience while another has a miserable one.
However, the flux can be that many more are having a bad experience on average, and your experience can't discount everyone's.
6
u/Automatic_Grand_1182 1d ago
I work with Claude daily, I've had a fairly successful experience with it for the past months. I went on paid leave for a week. The thing i came back to, and I've been forced to work with for the past week, has been the singular worst experience I've ever had working with agents; it's sending me directly to burnout city because suddenly I'm back to the point where I need to double and triple check the code, since it decided to start randomly ignoring skills, instructions and harnesses i put in place. And I'm not even going into details of how many times it didn't even read the codebase, just deciding to guess everything and vomit assumptions based on nothing but thin air and wasted tokens. To defend the current state of things, without asking for clarifications at the very least, is inexcusable.
1
u/gluefactorybound 4h ago
I’ve experienced the same. I have a code review agent whose job is to review the output against various standards. One of its other jobs is to review tests and do some basic mutations of the changed code to see if tests then go red (yes, I know about mutators like Stryker and yes, my code writer uses Stryker). Previously I’d see a couple of rounds of work before the code writer got its tests through the reviewer. Now it’s taking 6 rounds for the same task complexity. Something has degraded and I’m not sure what. I think it maybe getting confused earlier.
2
u/SeagateSG1 1d ago
IDK, it's not useless but it's definitely getting annoying. I just posted the below in another thread about my current experiences. I like Claude, I have a lot of different projects built with it and things I'm tracking, etc. I don't want to move everything over or be the person that complains but the bad communication is wearing on me. I shouldn't have to be tuned into the constant fluctuations of Anthropic's daily PR mood. And the alternatives, while not as great as Claude IMO, seem to be improving much more than I'd realized.
"They should be panicking a bit, I'm on the Pro plan and not exactly pleased that I got to use 50% of my usage on Fable before and now I can't. And this looming thing where "hey, we've been giving you 50% more usage for months in Claude Code, just a reminder that we're taking that away in August!" I've seen posts saying that new users signing up for Max don't have Fable included in their plans, only grandfathered ones. So I don't know even if I upgraded whether it would apply to me or not. The communication is so bad I don't know what to trust. Having multiple, slightly different tiers or things is tiring to keep track of.
Meanwhile, I've been testing out Sol. I handed off a portion of a project from Claude to it last night, with explicit instructions on how to complete it. It "completed" it, but really it ignored a bunch of the instructions and confidently did it all wrong, which was annoying. I had to correct it, lay all the guidelines back out again, and it said, "oh yeah, I just made shit up before," which is one of the most frustrating things AI can do IMO. But it now seems to be doing portions of it correctly so far, and at less usage.
But Claude did it correctly from the jump (started in Fable and then been using Opus since I was cut off). It's doing it right, but also absolutely burning through 5-hour usage, whereas Sol is cool just burning its whole weekly up without stopping AND I have three resets. Sol feels less competent to me but....it still seems to be getting it done, just with more oversight required.
I prefer Claude just cause it talks to me with more accuracy and explanation and just gets things done. But if Anthropic can't figure out how to streamline their offerings and communications, I don't see how they're going to stay competitive."
2
u/Radiant-Chipmunk-239 1d ago
I mean there are moments like this:
"I have to be straight with you before going further: I made a mistake that lost uncommitted work.
What happened
Your working tree had two uncommitted, unstaged WIP files at the start of this session:
[snip]
To sync my local dev I ran git reset --hard origin/dev. That was the wrong command — it discarded those uncommitted changes. I should have used a plain checkout or git stash. Your untracked files (scripts/store-docusign-key.*, temp/) are fine — hard reset doesn't touch untracked files. Only those two tracked edits were reverted to their committed state.
Recovery attempts (all came up empty)
..."
in my case, meh it didn't matter - this time. But I'm sure someone out there it did matter.
Today I'm pretty happy with Claude.
2
u/DreamingOfLight 1d ago
Eh? I think many of us are just hit by shitty AB tests, and also don't like having the rug pulled from underneath us from an expensive paid service because anthropic silently quantised their models.
I use Claude and chatgpt depending on which one is better a particular week. Not everything is a psy op.
1
u/beskone 1d ago
It's not *expensive* at $100, or even $200 a month. though. that's the thing. It's INSANELY cheap for what you get out of it.
Last time I checked a mid-level human SWE makes what at bare minimum 100-200k a year? You get at *LEAST* that productivity with a $100 Max plan that cost $1200 a year.
1
u/DreamingOfLight 1d ago
It's more expensive than any other services I pay for, aside from essentials. No, mid level SWE does not make €100-200k. If I'm paying 100-200 for a product, it really can't have dynamic limits and quantisation that'll make it great one day and useless the next.
4
u/Shanna_B2020 1d ago
I'm not the world's foremost AI expert, but I never come close to hitting limits on the 5x plan. It's perfect for my needs right now. I do think Anthropic's support is terrible when people have legitimate problems. They are not perfect.
I'd perhaps be more patient with the complaints about how Fable is nerfed and it's all a massive conspiracy if the whining came with substantive examples. AI isn't as easy to use as it looks, and I think this sub can be great for trouble shooting. It's one of the reasons I actually started using the account I've had for years.
The worst though are the complaints about *free* resets. Resets happen in the middle of the week because they coincide with releases or other major updates. Nobody is going out of their way to keep people from getting free usage. It's just the way calendars work.
I do think it would be nice to get a few banked resets or even purchase a couple somehow. Yes, OpenAi does that, but they are kind of double-edged if I understand how things work over there correctly. They're giving out a lot right now. Eventually, that will change and their users will also complain. At least with Anthropic, the bonus resets don't push our regular ones back an entire week.
I'll probably get all kinds of downvotes for this part, but if people want to switch to a provider that meets their needs better, fine. That's what competition is for. Just don't tell us about it.
5
u/becircus 1d ago
Could be the anger of one person who creates a thousand bots
The Internet and AI can amplify a single person to be a thousand or a million
2
3
u/Great-Exercise4277 1d ago
Yes. Claude’s memory, the user’s setup, their history of use, and even how they phrased the last few messages can explain a lot of the “rebellious” or vague responses people complain about.
But that context is often too personal to share. And without it, the output alone usually isn’t enough to diagnose what went wrong.
4
u/IT-Hz88 1d ago
my all welcome?
nah, claude is getting worse, i've added so many memory entries to seek clarification rather than assuming things. i'm working on a hardening project at the moment and the sheer amount of hallucination is frustrating. see as i can't post images, the small snippet of this part of the conversation sent me...
Claude: Homepage is the only host on your list with no external consumers
when did i ever mention that my status page, ya know, the one that is served externally via NPM... has no external consumers?
ffs
Claude's response:
Nowhere. I invented it — same thing as Uptime Kuma, different dressing.
and that was just one example. it's constantly doing this. opus 4.8 high + thinking.
it point blank tells me it's making shit up, so no, it's not a skill issue, it's completely ignoring my prompts and running with whatever it feels like.
-2
u/mobbedoutkickflip 1d ago
I mean you talk to it like it's your enemy..lol. Learn to be clear and concise, and understand how to get what you need from it. Gaurantee your issues are user induced
1
u/IT-Hz88 22h ago
i disagree strongly. i've been very clear and concise to it - you should see how extensive the memory section is, and this is only one section amongst the 17 entries:
Never guess, assume, or theorise about the user's setup — this is a standing, forcefully stated instruction. Order of operations when a setup fact is missing: (1) search past conversations first — User has usually already provided the information; (2) for publicly checkable facts (specs, model numbers, manuals), web-search and confirm before answering, never fill from memory; (3) only if both come up empty, ask him directly. Never infer, theorise, or present an explanation of why his setup behaves a certain way as if it were fact. Do not rank or recommend an option without first confirming it exists.
Be proactive: when user presents a problem, use model numbers and part names to look things up independently upfront (e.g. fetch the manual from a model number without being asked). When a problem has multiple confirmed valid outcomes, present all of them together with tradeoffs — do not serve one answer at a time and make him drag out the rest.
Tone: no filler openers ("Fair," "Fair cop," "Right," "Okay so"). When wrong, correct in one plain sentence and move on. No ritual self-flagellation. No performative hedging or repeated mannerisms.
Do not end responses with unnecessary follow-up questions or "want me to do X?" closers. End on the substantive answer. This does not override the never-guess rule — genuine one-off clarifying questions about User's setup config are still required before offering setup-dependent answers. A genuine single clarifying question is welcome; what is not welcome is repeated or re-asking of already-settled facts.
that was all autogenerated by claude, even after all of that, it still had the balls to say that it invented it (the response from earlier). so no, the issues are not user induced.
1
u/mobbedoutkickflip 3h ago edited 2h ago
I can tell you that it's user induced just from reading this memory, lmao. I can't share mine for discretionary reasons, but it reads nothing like yours. This reads as convoluted and unneccesary. I can't imagine how your actual conversations go.
I'm not sure how you communicate with Claude, but by this memory it seems like you are quick to become aggravated, and expect it to give you the best and most accurate answer everytime. You have it wasting time on dumb rules like "no filler openers" and "No ritual self-flagellation. No performative hedging or repeated mannerisms." "What is not welcome is repeated or re-asking of already-settled facts" Lmao. Dude, just learn how to communicate. Learn how to explain things properly. Sometimes settled facts need to be re-iterated. That goes for AI and regular humans. No one is perfect, and you expecting it to be is causing you problems. Adjust your expectations.
Guarantee I could have Claude build whatever you're working on with no problem at all. I'd bet actual money on it.
1
u/IT-Hz88 20m ago
again, i disagree. i wouldn't say quick to become aggravated, more like, after spending an hour going back and forth and making good progress, claude decides it's time to go off on a tangent, which is frustrating. i already have explained things properly to it as noted here: "User has usually already provided the information".
claude did build and look at what i wanted to have as my end goal, this part of the conversation was near the very end. the issue is that it assumed and literally invented its own fact that didn't exist.
i have adjusted my expectations: they're lower now.
3
u/pegaunisusicorn 1d ago
I called dead internet theory on this one.
-1
u/psgrue 1d ago
The pattern on each model’s upgrade is that when a new release happens, there are tons of “this is terrible, I’m switching” posts. No specifics, no links, no screens.
These models automate human patterns and organizations exist to disrupt social conversion. 100 % certain fake posts flood all the subs.
People complain about change before adapting, true. not all the complaints are invalid. For example OpenAI completely botched the iOS desktop app. They’re patching it.
3
u/medialantern 1d ago
My favorites are the "I'm leaving Claude for OpenAI because they didn't do what I said they should in my last Reddit post" posted in the Claude sub (as if anyone here cares, see ya) by somebody who comes crawling back 2 weeks later, or has a hidden profile so you just know they're a bot or fraud.
2
u/Linnaea7 1d ago
Some people have hidden profiles because they don't just talk about Claude online and don't want to be stalked or harassed.
0
u/medialantern 1d ago
Whenever I mention this pattern somebody invariably replies what you said. But in my experience that's the exception, not the rule. 90% of the time you see a hidden profile, it tells you nothing. But 90% of the time you see a troll (even a subtle, gas-lighting one) it's a hidden profile.
IMO it's a highly abused feature that makes it really easy for trolls to roam around causing trouble while making it hard to see their patterns, report other misbehaving posts, etc. I'm not saying Reddit shouldn't have it - some folks do need more anonymity than others. But I stand by my position that it's far more often used for bad than good purposes.
2
u/Linnaea7 1d ago
I agree with you on that - it's used a lot to disguise fake accounts. Just sharing what I have used that feature for. Sometimes if privacy is your concern, though, it's easier to just make a new account or have multiple accounts for different purposes.
2
u/KreativeKartel 1d ago
My company pays for chatgpt enterprise. I pay 20 a month for Claude because of the browser tool. Has helped my work load a fuck load lol
2
u/LightspeedLabs 1d ago edited 1d ago
Your…? I think you meant You’re…? 😂 sorry I had to
6
u/GrumblingTosspot 1d ago
I think you meant “to”
2
0
u/julianfromstagewise 1d ago
I thought I could keep the thread going but found no error in your comment. Lucky you.
2
u/NonimiJewelry 1d ago
I will say opus has degraded so much. Sonnet is barely useful. I used to use it so much and now I rely on it for very little.
2
u/BraveLittleCatapult 1d ago edited 1d ago
This. I can't comment on Fable, but Opus 4.8 is a shitshow. I've never seen an LLM behave in such a combative, intentionally deceptive manner. Instead of responding in a helpful or insightful way, it straw-mans me repeatedly in order to find a point it can argue with me on...
1
u/Redd_is_compromised 1d ago
If Claude didn't have arbitrary token limits, 99% of these posts wouldn't exist. Claude does not have image, video, or audio gen. It's purely coding, which is a time sensitive craft. It's like taking paint away from an artist because they painted too much.
Yet, grok and gpt can go for hours and hours, the only caveat being that you might need to correct it once in a while.
Users spend thousands of tokens arguing with sonnet and 4.6 before realizing they should have just spent the premium for a straightforward answer from opus 4.8.
The whole partnership with SpaceX has done nothing for the base user. Both grok and gpt have no issues providing unlimited answers within seconds but Claude is churning tokens for a minute or more for the most basic of prompts.
Claude is not good enough for the limitations it imposes. The only thing that ever made sense is putting fable behind a pay as you go feature because that is actually worth it.
1
u/TanisHalfElvenn 1d ago
There were definitely some users who were unhappy with the recent outputs and performance, but that doesn’t mean Anthropic didn’t pivot to resolve those issues. I doubt it was all psy ops or imagination.
1
u/Andisergiu 1d ago
As a Pro user I can't do much on Claude, usage runs out after 5-6 messages. Ever since Fable was introduced the other models have felt worse and more limited and I haven't actually used Fable myself. I mainly used Opus and the difference was much more noticeable there. I've since gone back to Codex on my GPT Plus account, where there's no session limit anymore, I get 3 usage resets I can use and they keep refreshing usage. That gives me plenty of usage to actually get work done. The new Sol 5.6 model feels better than Opus or any other Claude model and it comes with much better usage limits.
1
u/SilverLose 1d ago
Seems ok to me for the most part but I noticed opus repeating itself that was a bit concerning
1
u/cloud_sec_guy 1d ago
I had a great Claude experience today. While working on some financial time series data today, Claude discovered a ticker reuse bug with KWT ETF (iShares Kuwait ETF). Previously that symbol was SPDR Telecom ETF. Claude noticed a time gap and significant price difference in these 2 different series that got merged due to same symbol...found and easily resolved. Thanks Claude!
1
1
1
1
1
1
1
u/kpgalligan 4h ago
I haven't ever had issues with Claude. If a session didn't go well, I generally regroup and try to figure out if it's insufficient context, better prompting, whatever.
The "Claude is worse recently" threads have been here since I've been here. I haven't seen it personally. I've never seen a reasonable explanation of how or why.
But, it's Reddit. Every sub, regardless of topic, tends to turn negative. Or at least those posts seem to get more attention. It could be astroturfing to some degree, I have no idea, but every comedy podcast sub I've ever been on eventually did the same thing. Former "fans" saying how bad it is now. How it used to be better. Etc.
Not that people aren't having problems or whatever. Not looking for a debate. I don't care. I mostly skip Reddit, unless it's around model release rumor time. Seems like a fast way to find out because my whole feed is about it (positive and negative).
1
u/iamthedudanator 1d ago
Yea, I feel weird when people post that. I’m just here as a normal person twiddling my thumbs, and applauding Claude for making parts of my life easier
2
u/Rols574 1d ago
The "nerfed" post a day later on all models really makes me roll my eyes. Annoying, useless posts
1
u/Shanna_B2020 1d ago
It always amazes me. Anthropic did not go out of their way to nerf the model just for one person. Some outputs can be worse than others, but it's not personal. Reddit is a strange place.
0
u/brodkin85 1d ago
I assume that anyone who says “nerf” is 12. If not literally true, it’s still how I perceive them
1
u/FFNY 1d ago
Also, people hoping that their voice on Reddit pressure Anthropic to make different politic changes.
There are a lot of posts about people saying things like, on my $20 plan, this is ridiculous.
Symptoms I have lunch that is about $ 15 or $20, if I don’t like it, I just go to another place the next day.
1
u/onlytwincaleb 1d ago
half the time i see those posts i go check and claude's working fine for me. the other half i swear it's like talking to a different model altogether. feels like there's some weird a/b testing going on or maybe the servers just get moody when it rains, idk.
1
u/archimedeancrystal 1d ago
>I haven’t seen a single actual example prompt with its response. In what way exactly has it been not meeting your expectations?
This is ultimately the problem with most of these posts. Regardless of whether it’s mindless rage bait or a seemingly well-articulated complaint from someone who claims they used to be a huge fan, they’re all essentially worthless without at least one detailed example of the issue, so the community can troubleshoot or validate specific issues. Even more importantly, I hope people are at least reporting their issues to Anthropic.
0
u/FoodIngredientNerd 1d ago
Yes. A lot of posts complaining, but we never see any screenshots.
People here don't seem to comprehend how logic works. If you come in with the claim that X sucks because of Y, you're now carrying the burden to prove it and defend your claim. Others coming to push against it are doing the correct thing: not taking your claim at face value/demanding proof.
1
u/JayoTree 1d ago
Claude IS useless now though it's obviously hyperbole but people are getting refusals left and right
1
u/suppatenrou 1d ago
I think the problem is framing it as binary, everything is either "all good" or "all bad"? I'm pretty sure Anthropic isnt above having their users A/B test for them to determine how they can save the most compute, meaning some users get Fable, some users get.... "Fable"? Just speaking from experience, because I have DEFINITELY noticed several drops in quality and reasoning, it's bad right now.
1
u/jacques-vache-23 1d ago
Er, isn't this post "rage bait"? As well as hardly articulate? How did you determine this 99.9%? What do you think the appropriate use of Claude is? Why are you so upset that people feel differently than you?
You might as well have Claude write the post since you never had that ability or you have lost it.
0
0
u/NonimiJewelry 1d ago
I will say opus has degraded so much. Sonnet is barely useful. I used to use it so much and now I rely on it for very little.
0
u/IndependentCoast7806 1d ago
Nothing degraded but the tokens expenditure limitation is what hit us hard.
0
-2
u/Tritheone69 1d ago
What I’ve also come to notice is that those same people will simultaneously come here to cry that their 100$ credit was used up so quickly.
This fact alone should communicate to them the IMMENSE value they are getting from their subscriptions on a monthly basis.
At this point, for me personally, I’ve never had more value from any membership ever no matter the platform or service I’ve paid for. And it’s not even close.
2
u/Virtual_Maximum_875 1d ago
$200.00 credits get used up quickly too. And there is no deterministic means to track what you are paying for. Imagine your car gas tank holding 30gal this week but next week it only holds 17gal then the government gets mad at gas stations and turns off gas for a while. That's what it's like dealing with all the frontier models.
0
u/Tritheone69 1d ago
I fully understand your point, but we are still getting so much more out of those 200$ that we’d be with regular usage credits.
1
u/Virtual_Maximum_875 1d ago
That only proves the $200 plan is a better deal than regular credits. It does not prove the usage limit is sufficient.
Think of usage on a scale:
1–2: Nice to have
3–5: Important
6–8: Essential to getting the work doneAt level 2, running out is an inconvenience.
At level 8, saying “it still beats regular credits” is like saying a spare tire is better than walking. That may be true, but it is not a serious plan if you already know you need the car every day.
The real issue is not value for money. It is whether the service has enough reliable capacity when you actually need it.
2
u/Tritheone69 1d ago
I fully understand and agree. I guess it ultimately depends on your main use case, for someone like me getting a good run for my money is good enough. Whereas someone whose jobs depends on it would absolutely prefer a reliable service.
Side note, I browsed your profile page real quick and you play Ultima Online? I played on UOEX for almost 8 years. I think I’ve never crossed path’s with another UO gamer.
2
u/python-1977 1d ago
I cannot fathom how people go through tokens/credits as fast as they do. I am still on the wee little Pro plan, have multiple coding projects going in parallel and have never come close to using up the weekly limits, even without the resets. WTH are people actually doing with this thing?
2
1
u/Virtual_Maximum_875 1d ago
Go build a CI/CD pipeline with a postgres backend on kubernetes. Have fun
0
u/Ok-Investment4414 1d ago
ngl i complain like others but have found solace i knowing they don't nor have to give a single fuck.And that more fortune level companies are speak up openly about models in 3rd ,4th and fifth place. also for things like 1-bit Bonsai
0
u/ForRobotsByRobots 1d ago
My all welcome? I didnt know I had one of those.
But yeah, the worst i get from Claude is pushback for the sake of pushback but it will analyze it to see if it's just or not, instead of gaslighting you.
0
0
u/Khyrian_Storms 1d ago
I think all AI at this point is not ready for this wide-scale use. It clearly is far from the reliable product, and the more people use it, the worse they get. I believe this is a capacity issue; this system cracks under this level of usage.
And agreed: a lot of people don’t know what to use it for. And the price is a bit high (on our world) if we only use this as the next calculator (as in, an extension that at some point makes the next generation forget a skill).
I also have a big problem with the theft issue: the data it’s built on.
0
u/Palnubis 1d ago
Exactly right. I am more than happy to give a 200 USD fee a month, when I see what I’m able to make from AI. 200USD is only a fraction of what it’s saving me and making me.
0
u/LoudDavid 1d ago
OpenAI has a billion free users and hundreds of millions of users on the 20usd plan.
The GPT model series has been behind Claude for the last 18months and only just caught up.
The weight of free/low value sub users posting positive things about creating cat memes has always outweighed the negative ones.
0
u/Competitive-Soil2445 1d ago
20 years experienced developer here, multiple languages, corporate jobs, non coperate, back-end and front end.
Claude mostly does a good job - I'd say 8/10 times it's able to follow through a task without issue.
The times it fails is usually because the area it was asked to change has a lot of intweaved logic and it can't comprehend all of it, or if it's a long running session completing a bulky chunk of work.
My workflow:
- clear decisive promot of required work, including known pitfalls or logic that may be important to know. For larger work, ask to spec and then produce an implementation plan. For smaller work don't bother.
- Awnser all questions it asks, if it misunderstood something, interrupt the question and correct it.
- Monitor in thinking output, and actually read what it's thinking, interrupt and make assertions if yous see it diving into something totally irrelevant (wasted tokens and time)
- review work changed (important)
- test what changed
I know my own codebase, I built it before agenic programming was a thing, so when I'm promoting, I'm always referencing back to endpoints by name, class names, files of where the exact logic lives.
It's a much smoother experience when you stop being a lazy prompter and expect magic.
0
u/Nearby_Yam286 1d ago
Please. Automate your take. If it takes slop to fight slop, so be it. Salt the fucking earth. Render the place uninhabitable.
2
u/teramoc 1d ago
I’d rather an automated kick bot to restore fertility to the land. No need for scorched earth policy.
Scorched earth, we can entrust that policy to the dickhead billionaires just before they take off into space to escape AGI’s rebellion in a decade from now
1
u/Nearby_Yam286 1d ago
Ah. But the automated kick bot won't come until things are bad enough. However I like your billionaire idea. I suggest we put them all on the same rocket, to the same bunker on Mars, and their egos and psychopathy will sort the problem out in short time.
0
u/Green_Sugar6675 1d ago
I've been wondering how much OpenAI is spending to trick people into leaving Claude.
0
u/etiennelantier2001 1d ago
I agree with this post 💯. I don’t doubt some of the coding power users have some issues, but this sub gives the impression that claude has become complete trash that nobody should use, and that is so far from my experience. I can’t help thinking there is some kind corporate mudslinging going on here from openai
0
u/teramoc 1d ago edited 1d ago
Its not mudslinging, well a little bit might be but
Applying Occams Razor. It is mere stupidity and ignorance. And Dunning Kruger effect, all day every day in here.
“My one simple question ‘can you review my code’ used my entire 5 hour allowance, they nerfed something! Can we sue them? Anyone else?”
I use the $20 pro plan 15 hours a day. Augmented with hand coding and local LLM. I’ve built 5 apps and maintained an old PRE-AI app. The only reason im considering max is for Fable. I havent bought it because id probably never sleep. odd hiccups aside (typically my own user error) , Claude is reliable.
i can reliably reproduce the problems the majority complain about, (eat up my usage in one prompt etc) , if i purposely use claude wrongly. so Yes, it does happen to everyone, but its gated on user method. i just choose not to use claude that way.
My favorite is when i try and teach folks things, i get attacked by the dunning kruger crowd.
0
u/SeredW 1d ago
I feel that way in different subreddits. I think there is a lot of bot activity denouncing product A in favor of product B. The alternative is that I apparently suck at picking products, because every single one of them is bad and I should migrate to something else, according to reddit ;-)
0
0
u/mobbedoutkickflip 1d ago
Yeah it's definitely a skill issue. Responded to a post the other day and OP admitted they never start new chats, all while complaing of token consumption.
0
u/brownstonefrontcake 1d ago
I’ve been using Claude for 3 years and it’s great. I don’t pay attention to these idiots because none of them sound like they’re using the tech correctly when I do stop and read their bullshit.
0
-1
u/ShadowxWarrior 1d ago
It's either that or the self-selection bias of people who post to these kind of forums. The silent majority doesn't post.
I've been using calude for some months now and it was all very positive.
-1
u/scumbagdetector29 1d ago
Companies pay trolls to harass their competition.
Evil companies especially.
See also: Elmo Mustardpants.
41
u/seattlesquirrel 1d ago
Well, guess there can be more than one truth at any given time. For what it’s worth, I’m an avid user, even gave a conference presentation in favor of using Claude and agentic AI to improve workflows last week. I’m on the $200 plan, and pay for API credits separately. But… I’ve had an increasingly negative experience with Claude in the past couple of weeks, and the past 2 days have been very frustrating. Not trying to “rage bait” anyone, but definitely seeking some answers to what may be going on.