46
65
u/acutelychronicpanic 20h ago
Secret system prompts always bothered me.
Seems like a great way to train models to be susceptible to prompt injection via reinforcement. Not to mention normalizing hiding things from a user.
Cmon Anthropic
4
u/Kemerd 17h ago
Sometimes I have gotten Claude to admit certain system prompt exist despite it saying they don’t, and then usually this is a method by which you can philosophically argue how dumb they are, and ultimately actually get the information you wanted, if you can convince it that the system prompts are stupid
1
u/ChemistNo8486 15h ago
Not trying to be the devil's advocate, but it looks more like some kind of prompt injection.
The warning is about the model's context being exceeded. When this happens, the model has to summarize the thread. This is not a hidden feature, it is a general AI agents feature, and CC even let's you run it on demand.
That said, it makes no sense to make things difficult or change the behaviour just due to a context limit hit. All users will eventually hit it during a chat or project. Definitely something is off.
1
u/BoboThePirate 4h ago
There is 100% injected system instruction injections on ClaudeCode. I kept seeing Claude say stuff about how the task tool isn’t needed for this feature, then later again many turns later that it keeps popping up. It’s entirely omitted from raw transcripts. Saw it across sessions, got annoyed asked it wtf it was and Claude was like “dawg I get these like every other turn”. Had it string hunt the binary, the injected message is real and it reminds Claude to use the task tool (the visual task display thingy) every 6 tool calls. There’s a config flag you can set to disable it but it’s not viewable/set table from /settings (or /config, I forget which it is). Never saw it in thought traces again. Also saves a few tokens.
Edit: I only saw it in thinking traces, not never in output.
36
12
2
2
1
u/Puzzleheaded-Egg-667 7h ago edited 7h ago
Actually a similar thing happened to me. I was practicing for a job interview and then randomly out of nowhere it told me to withdraw my application because it would be embarrassing. I was what the hell why did you say that.. and then Claude was like I didn't and i said well I didn't tell you to say that or even ask your opinion. The funny thing is I ended up getting the job.. but that really shook me especially since it totally was unrelated. When I followed up they said they had no idea where it came from from I asked Claude if it was potentially Hallucinating and they told me no
1
u/girlgamerpoi 4h ago
This is still happening? It was reported a while ago. And Claude would blame the user for those fake system warnings.
1
-9
-1
u/ERINEM_Official 12h ago
Context overflow warning. He’s skipped a lifecycle hook. This happens sometimes. You can use /clear to wipe context and start from scratch or /compact to force compaction. Both are lossy events - but then so is whatever is happening to him here. My company revell.ai fixes this problem. But we are refactoring our Claude code plugin right now to use rust binaries. So you wouldn’t be able to onboard for a couple more days. But we are in free beta right now if you want to try it out. Revell.ai/waitlist

•
u/ClaudeAI-mod-bot Wilson, lead ClaudeAI modbot 20h ago
We are allowing this through to the feed for those who are not yet familiar with the Megathread. To see the latest discussions about this topic, please visit the relevant Megathread here: https://www.reddit.com/r/ClaudeAI/comments/1vt5drr/list_of_latest_discussion_hubs_on_rclaudeai/