Anthropic says it can read Claude’s ‘thoughts,’ as detailed in new research paper — models observed to have a global workspace, revealing more of what makes LLM

When looking at the J-Space after Claude received prompt-injection data as part of data acquisition, Anthropic discovered the model appeared to be aware of this deception, surfacing related words like “fake, injection, f…

Anthropic says it can read Claude’s ‘thoughts,’ as detailed in new research paper — models observed to have a global workspace, revealing more of what makes LLM

When looking at the J-Space after Claude received prompt-injection data as part of data acquisition, Anthropic discovered the model appeared to be aware of this deception, surfacing related words like “fake, injection, f…