Testing whether sparse autoencoders encode causal structure or only observable statistics. Trained 10 SAEs across 5 seeds per world; found no significant representational difference between direct and confounded causal setups — a negative result, written up with full methodology.