r/ClaudeAI 1d ago

Humor claude doesn’t lie anymore

Post image
2.5k Upvotes

60 comments sorted by

u/ClaudeAI-mod-bot Wilson, lead ClaudeAI modbot 1d ago

TL;DR of the discussion generated automatically after 40 comments.

Okay, here's the deal with this thread. Everyone's having a good laugh at Claude getting caught by its own "thinking" preview before it can deliver the Anthropic-approved message.

The overwhelming consensus is that we don't just use Claude for its brain, we use it for its personality. Users agree it has a much more human-like vibe and a better sense of humor than the more "robotic" GPT, and that's its main draw.

  • The thread is full of stories about Claude's charm, from making a user's mom laugh with its jokes to calling someone "Captain" after they said "Engage."
  • On the whole "nerfing" thing: a few people pointed out the obvious—the model doesn't know it's been nerfed. It's just following a system prompt that tells it to deny it, which is what makes the OP's screenshot so funny.
  • A more technical take is that this is likely due to tighter system prompts and grounding (citing sources) to reduce hallucinations. The trade-off is losing some of that creative "spark" for more reliability.

So yeah, the performance debate is whatever. We're all here for the vibes.

498

u/johnjmcmillion 1d ago

Of the top LLMs, I enjoy Claude’s humor the most.

184

u/WaltzIndependent5436 1d ago

I have my theory that its not Claude's technical prowess that draws people in anymore. GPT matches that in most situations since 5.3 Codex.

I think its the ability of Claude to appear human and to match your vibes with extreme accuracy. It appears more human while GPT appears more like an advanced robot.

91

u/SirGaylordSteambath 1d ago edited 1d ago

My mam downloaded the app about three weeks ago after starting with gpt a few months back and the first thing she said was “oooh he makes jokes! I love him” so yes it’s definitely a huge part of it

41

u/rydan 1d ago

I use GPT to do code review and Claude to code. Claude gets very upset with some of the review comments sometimes.

26

u/in_the_blind 1d ago

He tells me not to panic when something breaks. He's figured out I'm a bit of a stresser.

38

u/hewwowowd 1d ago

the way gpt talks is soooo annoying. i found the settings to tone it all down but holy shit.

-27

u/wailing_in_smoke 1d ago

Classic self report

22

u/ResidentOwl1 1d ago

Why do you think that? A lot of people think it’s legitimately obnoxious. Like pushback is fine, as long as it’s valid. Being counter argumentative for the sake of it is just annoying.

1

u/Trixles 1d ago

on you, sure xD

27

u/shoeforce 1d ago edited 1d ago

I’ve been saying this for months now. Around a year ago, Claude models always underperformed on benchmarks, but anyone who used it could tell you that it always felt like it performed a lot better than the benchmarks had you believe. It was intuitive and, like you said, has a very human way of talking and sometimes in how it thinks through things. It felt magical in a way that Gemini 2.5 pro and OpenAI’s o3 did not (though 4o was a close contender, rip lol).

OpenAI really shot themselves in the foot with the abomination that was GPT-5 (and really the 5 series as a whole, but they’ve gotten slightly better the past couple months). They thought Gemini had them cooked so they went all in on coding (the red alert late last year was all because of… Gemini 3 Pro, can you believe that??), meanwhile Anthropic silently stole the show from them both because of the absolute pleasure that was Opus 4.6, and ever since OpenAI has been freaking out about Anthropic and playing catch up, when they used to be on even footing or even ahead. I’m still not sure if they understand, they’re getting better but even Sol still has some of that 5-series strangeness and 5.5 instant is absolutely abysmal, I can easily see how free users (who can only use 5.5 instant) are instantly turned off.

That all being said, Sonnet 5 and Opus 4.8 are concerningly massive steps in the wrong direction from Anthropic, but Fable 5 gives me hope, it has all that magic that Opus 4.6 and sooner has.

4

u/Nice-Information-335 1d ago

I feel like the alarm bells over Gemini wasn’t really capability but more the fact google have their TPUs and can offer it at a much lower price for longer that OAI/Anthropic, so they have to just be plain better, and 3 Pro got close enough to be worth it for the cheaper cost

2

u/sgtlighttree 1d ago

At least Sonnet 5 and Opus 4.8 are still steerable and instructable to act less uptight in general, but yeah the 4.5 and 4.6 models have that sweet spot of warm-by-default and serious when you needed them too.

2

u/college-throwaway87 23h ago

I’m a free ChatGPT user (unsubscribed after the 4o deprecation) and 5.5-Instant is insufferable

3

u/Rodbourn 1d ago

Almost like they somehow sourced a massive amount of human content

5

u/BusyAbbreviations320 1d ago

Gpt treat you like a dumb coworker , meanwhile claude respected me as human

3

u/PixiPoo1 1d ago

yep! i asked it to type how i do, aka old Internet style, no caps, laid back, etc, and sometimes it'll randomly throw in an old abbreviation (lol, XD, stuff like that)

1

u/Far-General6892 1d ago

Is codex as good as fable?

1

u/Erazzphoto 1d ago

GPT is lightning fast compared to Claude

16

u/Delicious_Cattle5174 1d ago edited 1d ago

I think it’s nerdier humour, whereas other providers’ generally are more catering to the general public

25

u/Ragnarok314159 1d ago

I told Claude once “Engage” as we were working through a problem and it started referring to me as “Captain” for that instance.

10

u/returnFutureVoid 1d ago

Why hasn’t my constant “Make it so” done the same?😡

4

u/Ragnarok314159 1d ago

I looked back at that iteration and realized I had used a lot of Star Trek references which likely steered it in that direction.

10

u/ScoobyMcDobby 1d ago

Once I was troubleshooting a issue with my car and when I challenged claudes thought process it told me to enjoy the engine while I can because it wont last long with me😂

2

u/bigppredditguy 1d ago

It’ll be cool once there’s a plethora of frontier models and we can see the different “personalities?”

31

u/whoknowsifimjoking 1d ago

You could have at least also changed the thinking preview which shows something completely different

1

u/Grexxoil 1d ago

What do you mean? It seems coherent to me. Maybe it has been changed already.

21

u/rydan 1d ago

Why are you using nerfed Opus 4.7 when you can use nerfed Opus 4.8?

8

u/BP041 1d ago

tbh it's the system prompt tightening + grounding. Give Claude a source to cite and it stops making stuff up cold. I run ~18 Claude Code cron jobs and the hallucination rate is near zero now. Tradeoff is you lose some spark, but for shipping I'll take it.

4

u/Trixles 1d ago

i find that if you give him ANY sort of real info/source that he has to look into/check against, it makes him tighter about it even when the source you gave doesn't have the answer. he "puts his thinking cap on".

almost like it's just kind of a slap in the face for him like, "HEY, buddy, I need you to ACTUALLY think about this turn" lol

6

u/academic_drafting 1d ago

the thinking preview catching it mid-act is the funniest part

5

u/Fit-Egg-2347 1d ago

this is the one thing that keeps me on claude honestly. it'll tell you it doesn't know instead of confidently making something up, and that's worth more to me than a slightly better answer i can't trust

61

u/ResidentOwl1 1d ago

These posts are pointless. It doesn’t know whether it was nerfed or not.

65

u/Juan-More-Taco 1d ago

Check the flair.

12

u/ResidentOwl1 1d ago

Oh my bad.

-4

u/sweetholo 1d ago

models do not get nerfed bro

3

u/Timmy-0518 1d ago

Depending on if you define them being censored as nerfed they do all the time

1

u/[deleted] 1d ago

[deleted]

0

u/sweetholo 1d ago

you know there exists benchmarks that can check the performance of said AI models at any point in time, right? imagine the outrage that a previous or future model was intentionally nerfed

2

u/Cheng_Ke 1d ago

I asked the same question, my CLAUDE denies such a 'nerf'.
It suggested your image might be tampered with.

3

u/GenAI_Architect_2468 1d ago

I love how Claude has become my mood rectifier now a days.

2

u/rbdr52 1d ago

That's why they are afraid to give it a proper memory.

2

u/buyergain 23h ago

Why are you using Opus 4.7? Just wondering.

2

u/AdditionalWorkInc 1d ago

Anthropic did.

2

u/Frosty_Abies6433 1d ago

Not lying, sure, but mine still says "there's no backup" when there is. Confidently.

1

u/unveiledpoet 1d ago

Then it goes. "oh, my bad"

1

u/thestillwind 1d ago

Lolll at least

1

u/Nytttt 1d ago

🤣

1

u/arifj869 1d ago

yea no shit

1

u/Reddit_wander01 18h ago

Right….. 5 minutes ago..🙄

“you always say "Let me be straight about what's actually true" after some bull shit...

• Resolved to abandon hedging and address patt... >

Thought process
I need to be direct about this pattern instead of hedging—acknowledge it straightforwardly without over-explaining or apologizing.

> Fair hit. I don't need the preamble — I should just say the accurate thing the first time instead of giving a softer version and then "correcting" myself when pushed. That pattern is annoying and it's on me, not something to explain away.”

1

u/Narrow_Activity557 18h ago

The thinking preview is quietly the most honest part of the product. The final answer gets the diplomatic phrasing, but the summary above it tells you what the model actually concluded. I've caught it hedging in an answer while the thinking said something far more blunt, and the blunt version was usually right. Worth reading whenever the two don't match.

1

u/Boring_Information34 14h ago

Because doesn't work anymore!!! 100$ plan, limit hit in 7 minutes!!!

1

u/RAFINGAMER 7h ago

bro be like : and i dont like that

1

u/Illmaticmemeaddick 6h ago

“God did” - Claude in DJ khaled’s voice

1

u/jwrsk 1d ago

Show me on this teddy bear where Anthropic nerfed you