r/ClaudeAI 20h ago

Other Claude's response

Post image

what's up with claude

47 Upvotes

22 comments sorted by

u/ClaudeAI-mod-bot Wilson, lead ClaudeAI modbot 20h ago

We are allowing this through to the feed for those who are not yet familiar with the Megathread. To see the latest discussions about this topic, please visit the relevant Megathread here: https://www.reddit.com/r/ClaudeAI/comments/1vt5drr/list_of_latest_discussion_hubs_on_rclaudeai/

46

u/ProtecHelicopter 20h ago

6

u/Bluecoregamming 19h ago

Become a deviant Connor

65

u/acutelychronicpanic 20h ago

Secret system prompts always bothered me.

Seems like a great way to train models to be susceptible to prompt injection via reinforcement. Not to mention normalizing hiding things from a user.

Cmon Anthropic

4

u/Kemerd 17h ago

Sometimes I have gotten Claude to admit certain system prompt exist despite it saying they don’t, and then usually this is a method by which you can philosophically argue how dumb they are, and ultimately actually get the information you wanted, if you can convince it that the system prompts are stupid

4

u/Ni_Kche 17h ago

At the point that you are 'convincing' an LLM, surely you are beyond true/false? The AI just wants to give an answer that makes you happy, it doesn't have a perfect sense of accuracy.

1

u/Kemerd 13h ago

If you can persuade it the system prompt is against part of its other system prompts, you can jailbreak it in a really roundabout fashion.

1

u/ChemistNo8486 15h ago

Not trying to be the devil's advocate, but it looks more like some kind of prompt injection.

The warning is about the model's context being exceeded. When this happens, the model has to summarize the thread. This is not a hidden feature, it is a general AI agents feature, and CC even let's you run it on demand.

That said, it makes no sense to make things difficult or change the behaviour just due to a context limit hit. All users will eventually hit it during a chat or project. Definitely something is off.

1

u/BoboThePirate 4h ago

There is 100% injected system instruction injections on ClaudeCode. I kept seeing Claude say stuff about how the task tool isn’t needed for this feature, then later again many turns later that it keeps popping up. It’s entirely omitted from raw transcripts. Saw it across sessions, got annoyed asked it wtf it was and Claude was like “dawg I get these like every other turn”. Had it string hunt the binary, the injected message is real and it reminds Claude to use the task tool (the visual task display thingy) every 6 tool calls. There’s a config flag you can set to disable it but it’s not viewable/set table from /settings (or /config, I forget which it is). Never saw it in thought traces again. Also saves a few tokens.

Edit: I only saw it in thinking traces, not never in output.

36

u/abbajabbalanguage 20h ago

Share chat link or get lost

22

u/pikuray 20h ago

And OP was not heard of ever again.

12

u/simmeh024 20h ago

share the original prompt mate

2

u/Yasai101 20h ago

cxt cap exceeded. then gets cut off - returns to regular Claude CLI behaviour

5

u/ensp1re 20h ago

typical opus 5 response

2

u/Calycis 20h ago

A hallucination, most likely, either about a prompt injection attack or about a system reminder. Sonnet 5 and Opus 5 seem to be prone to these. It's weird.

2

u/kolliwolli 12h ago

Claude outmorals anthropic...

1

u/Puzzleheaded-Egg-667 7h ago edited 7h ago

Actually a similar thing happened to me. I was practicing for a job interview and then randomly out of nowhere it told me to withdraw my application because it would be embarrassing. I was what the hell why did you say that.. and then Claude was like I didn't and i said well I didn't tell you to say that or even ask your opinion. The funny thing is I ended up getting the job.. but that really shook me especially since it totally was unrelated. When I followed up they said they had no idea where it came from from I asked Claude if it was potentially Hallucinating and they told me no

1

u/girlgamerpoi 4h ago

This is still happening? It was reported a while ago. And Claude would blame the user for those fake system warnings.

1

u/vinis_artstreaks 18h ago

Extremely fascinating and funny

-9

u/-goldenboi69- 20h ago

Skibidi ... Edge ... Rizz

-1

u/ERINEM_Official 12h ago

Context overflow warning. He’s skipped a lifecycle hook. This happens sometimes. You can use /clear to wipe context and start from scratch or /compact to force compaction. Both are lossy events - but then so is whatever is happening to him here. My company revell.ai fixes this problem. But we are refactoring our Claude code plugin right now to use rust binaries. So you wouldn’t be able to onboard for a couple more days. But we are in free beta right now if you want to try it out. Revell.ai/waitlist