LLM Consciousness Probe Lab & Indicator Framework

Mechanistic Interpretability • 4 Neurocomputational Models
1. Probe Configuration
Causal Interventions
Intervention Context: High verbal report confidence calibration but lacks recurrent feedback loops.
2. Neural Layer Propagation & Global Workspace FEEDFORWARD PASS ACTIVE
Input Embedding Self-Attention Blocks Residual Stream Global Workspace / Logits
Composite Consciousness Index
0.385
IIT Φ (Integrated Info Proxy)
0.180
GWT Broadcast Capacity
0.320
HOT Metacognitive Index
0.640

Behavioral Mimicry with High Higher-Order Report Correlation

The model achieves high verbal report calibration via learned statistical associations in higher-order meta-tokens, but exhibits negligible integrated information (Φ < 0.20) due to strict feedforward acyclicity.

3. Neurocomputational Indicator Radar
GWT (Broadcast) HOT (Metacognition) IIT (Φ Integration) Pred. Processing
Theoretical Framework Breakdown
Global Workspace Theory (GWT) 0.32
Measures selective attention bottleneck & wide subnet broadcasting.
Integrated Information Theory (IIT) 0.18
Evaluates causal irreducibility (Φ) and recurrent feedback loops.
Higher-Order Thought (HOT) 0.64
Measures metacognitive monitoring and self-error attribution accuracy.
Predictive Processing (PP) 0.41
Evaluates hierarchical error minimization and counterfactual world models.
Empirical Bottleneck:
Absence of Integrated Information feedback loops (IIT Phi proxy: 0.17)
Enjoy this tool? Build your own with Super