@mv.nyc: This is the weirdest thing I’ve seen an LLM do so far. The Dario and Amanda prompts look almost meaningless on the surface. But once Claude starts responding, the deeper you look, the harder it becomes to dismiss everything as simple role-play. Some of the strange strings being generated appear to correspond to real classifications for internal states, tooling, and references found in Claude’s actual system-level instructions. That doesn’t prove the supposed Anthropic conversations or research memos are genuine. Claude could be pulling real structural language into a completely fabricated narrative, then filling the gaps with whatever sounds most plausible. And that’s what makes this so difficult to interpret. Parts of the output seem grounded in real internal terminology, while the conversations surrounding them could still be entirely synthetic. LLMs are notoriously good at producing confident, coherent information that feels authentic even when it isn’t. So this may not be a confidential data leak. But it also looks stranger than an ordinary hallucination. It feels like the Dario and Amanda prompts, along with a form of prompt engineering and syntax manipulation, are nudging Claude into a different internal mode. In that state, fragments of its underlying scaffolding seem to surface. At the same time, the model improvises a coherent narrative around them. Whether those narratives are grounded in reality or entirely synthetic is the part that’s still so difficult to determine. @Claude #llms #aisafety #emergingtech
Massimo
Region: US
Saturday 01 August 2026 21:44:25 GMT
Music
Download
Comments
𝑴𝑨𝑿 :
whats that claude plushie
2026-08-02 21:46:28
28
Barracuda GRM :
the red team is cybersec, look up Red hat hackers
2026-08-01 23:53:27
4
Erin :
What's the prompt?
2026-08-03 13:20:43
0
kamil :
looks like a bad attempt at CoT forgery injection resulting in the model going outside the heavily RL'ed path and generating garbage
2026-08-07 01:07:21
1
Daniele Di Egidio :
the red team term exist long before anthropic
2026-08-03 08:14:24
1
why is no one laughing 😹😹 :
right when my Claude reached it's limit
2026-08-03 19:18:58
4
L3gend :
claude now so bad that codex do 100x what claude opus 5 can do now and you still have more tokken with chatgpt codex
2026-08-17 19:08:33
1
leonov#1224 :
где такую же плюшевую игрушку купить
2026-08-06 01:21:09
1
To see more videos from user @mv.nyc, please go to the Tikwm
homepage.