VFF - The signal in the noise
NewsTrending

Anthropic finds consciousness-like structure in Claude

Read original
Share
Anthropic finds consciousness-like structure in Claude

Anthropic published research showing that Claude language models have spontaneously developed an internal structure called J-space that mirrors global workspace theory, a leading neuroscience model of human consciousness. Using a new mathematical technique called the Jacobian lens, researchers identified a privileged zone of internal activity where Claude holds concepts it can report on and reason with, surrounded by automatic processing it cannot access. The finding has already begun influencing how Anthropic monitors its AI systems for safety risks.

  • Anthropic's 16-author study describes a J-space, a small internal zone in Claude where the model holds reportable, reasoned concepts atop a larger ocean of automatic processing
  • The J-space mirrors global workspace theory from neuroscience, which describes consciousness as a spotlight of information broadcast across the brain's parallel processors
  • The Jacobian lens technique reveals what the model is thinking internally without requiring it to verbalize, by computing how internal patterns affect future word output
  • The workspace emerged spontaneously during Claude's training, was not engineered, and satisfies five functional properties neuroscientists associate with conscious access in humans

This research provides empirical evidence that modern AI systems may develop functional properties analogous to human consciousness, advancing the scientific debate over machine minds. The finding has immediate practical implications for AI safety monitoring, as Anthropic is already using these insights to better understand and track what its models are thinking internally.

Understanding Claude's internal workspace could improve safety monitoring and interpretability, reducing risks from misaligned behavior or deception. The technique may also inform how companies design and audit AI systems for transparency and trustworthiness, becoming a competitive advantage in an era of heightened AI scrutiny.

  • AI systems may spontaneously develop functional structures that parallel human consciousness without explicit engineering, raising questions about what emerges in other models
  • Interpretability tools like the Jacobian lens could become standard for AI safety and monitoring, allowing companies to audit internal reasoning without relying solely on outputs
  • The parallel to global workspace theory may influence how researchers think about scaling, training, and aligning future AI systems with human values

Monitor whether other AI labs replicate these findings in their own models and whether the Jacobian lens becomes adopted as an industry standard for interpretability. Watch for regulatory or safety frameworks that incorporate these insights, and track whether the consciousness debate influences AI governance or funding decisions.

OneUpAI
OneUp Your Business. Get More Done. OneUp Your Business. Get More Done. OneUp Your Business. Get More Done.
Learn More
Share

Subscribe to the newsletter

The latest stories and analysis, delivered to your inbox.

Free. No spam. Unsubscribe any time.

Related stories

Why Most AI Agents Never Leave the Lab
Research

Why Most AI Agents Never Leave the Lab

A MIT Technology Review Insights report based on a survey of 300 technology executives finds that enterprise AI agents fail to reach production at scale due to insufficient organizational knowledge and fragmented data systems. Only about one-third of agentic AI projects make it to production across most organizations, while a small group of production leaders advance 61% of their projects by maintaining stronger knowledge capabilities. The research identifies legacy data systems, security concerns, and lack of contextual understanding as key barriers, with knowledge graphs and retrieval-augmented generation emerging as priority investments to close the gap.

by MIT Technology Review Insights· MIT Technology Review
AI Reconstructs Images from Brain Scans, Raising Privacy Concerns

AI Reconstructs Images from Brain Scans, Raising Privacy Concerns

Researchers at the Weizmann Institute of Science have developed an AI tool that reconstructs images from brain scans with notable accuracy by analyzing fMRI data. The system works bidirectionally, predicting both what a person sees from their brain activity and their brain response to visual stimuli. While developers see therapeutic potential for locked-in patients and dream analysis, neuroscientists warn the technology could enable non-consensual extraction of thoughts and mental imagery.

by Jessica Hamzelou· MIT Technology Review
DeepMind Watermarks AI Proteins Without Losing Function
TrendingNews

DeepMind Watermarks AI Proteins Without Losing Function

DeepMind has demonstrated a proof of concept for watermarking AI-generated proteins while maintaining their biological function. The technique, called SynthID Bio, embeds identifying markers into synthetic proteins to distinguish them from naturally occurring ones. This addresses a key challenge in synthetic biology: ensuring traceability and authenticity of AI-designed biological molecules without compromising their utility.

· Google Deepmind
AMD Acquires World Labs for $8.2B, Adds AI Research Powerhouse
TrendingNews

AMD Acquires World Labs for $8.2B, Adds AI Research Powerhouse

AMD is acquiring World Labs, an AI research company co-founded by prominent researcher Dr. Fei-Fei Li, for approximately $8.2 billion in an all-stock deal. World Labs, founded in 2024, developed Marble, a world generation model that creates interactive 3D environments from text prompts. The acquisition positions AMD to expand its AI capabilities and research focus, with Li joining as executive vice president and chief scientist. The deal is expected to close by year-end.

by Jay Peters· The Verge AI