spb/modelmap Public License
Internal cartography of local LLMs on Apple Silicon — registered, gated, negative-first. Public atlas at modelmap.io.
Python 66.3%
JavaScript 24.5%
CSS 8.1%
Shell 0.7%
1---2project: modelmap3document: qwen3-0.6b-4bit/probes/v2 — confidence4author: Simon-Pierre Boucher5contact: contact@spboucher.ai6website: https://modelmap.io7created: 2026-08-128status: reviewed9---1011# Confidence — qwen3-0.6b-4bit / probes / v21213```text14Level : 115Seeds : 516Prompt sets: 6 (token-balanced, structure-borne; overlap certificates in manifest)17Methods in agreement : 1 (linear probes only — Level 2 requires a second method)18Causal verification : ATTEMPTED AND FAILED (expC run #1 layer-skip: top-519 differential layers not confirmed; survival 0/1)20```2122Per-property verdicts (differential real−twin, mean pooling):23- word_order: null-dominated (twin acc 0.96; surface statistics explain the map); maxAcc A (mean-pool) 0.996, twin acc 0.958, signal layers 2/2824- agreement: trained-model signal (real−twin sel > 0.10 on 25/28 layers, max +0.38); maxAcc A (mean-pool) 0.967, twin acc 0.729, signal layers 25/2825- arith_valid: trained-model signal (real acc 0.86–0.90 vs twin 0.56–0.58); maxAcc A (mean-pool) 0.858, twin acc 0.558, signal layers 5/282627Published claims are DIFFERENTIAL only (real minus random-init twin), per the28doctrine adopted after v1's validity-gate failure. The strict twin gate29(selectivity < 0.05) still fails on word_order and agreement — the twin30extracts real surface signal from tokenization statistics — so raw probe31accuracies are never cited as evidence of learned structure. What survives:32agreement and arith_valid show layer-resolved trained-model signal33(Level 1, correlational; 5 seeds × 2 sets, controls listed above).34