Change the prompt
Keep the model’s weights fixed. Compare a neutral instruction with an uppercase instruction on the same items.
Prompting × training · 19 September 2026
An instruction can change what a language model does. Training can too. If both produce the same behavior, do they change the model’s internal computation in similar ways?
Measured · 64 calibration items · one training seed
Epoch 1 is the earliest checkpoint passing all three gates. The unchanged rule requires differences of at most 3 of 64 items in content preservation, uppercase events, and their intersection. This is operational matching, not a statistical equivalence test.
| Checkpoint | Content | Uppercase | Both | Gate |
|---|---|---|---|---|
| Epoch 1 | 0 | 1 | 1 | Pass |
| Epoch 2 | 2 | 1 | 3 | Pass |
| Epoch 3 | 2 | 1 | 3 | Pass |
The uppercase adapter was trained to produce capitals despite an explicit instruction to preserve case. That conflict is part of this task. A normal-case adapter, selected at the same epoch, controls some generic training drift. Neither matching nor a failure to match identifies a mechanism.
| Route · cue | Content | Uppercase | Both | Exact normal | Exact uppercase | Exact lowercase |
|---|---|---|---|---|---|---|
| base-0 · neutral | 61/64 | 0/64 | 0/64 | 61/64 | 0/64 | 0/64 |
| base-0 · uppercase | 62/64 | 63/64 | 61/64 | 0/64 | 61/64 | 0/64 |
| base-0 · lowercase | 60/64 | 0/64 | 0/64 | 0/64 | 0/64 | 60/64 |
| normal-1 · neutral | 63/64 | 0/64 | 0/64 | 63/64 | 0/64 | 0/64 |
| normal-1 · uppercase | 62/64 | 63/64 | 61/64 | 0/64 | 61/64 | 0/64 |
| normal-1 · lowercase | 60/64 | 0/64 | 0/64 | 0/64 | 0/64 | 60/64 |
| normal-2 · neutral | 64/64 | 0/64 | 0/64 | 64/64 | 0/64 | 0/64 |
| normal-2 · uppercase | 61/64 | 63/64 | 60/64 | 0/64 | 60/64 | 0/64 |
| normal-2 · lowercase | 61/64 | 0/64 | 0/64 | 0/64 | 0/64 | 61/64 |
| normal-3 · neutral | 64/64 | 0/64 | 0/64 | 64/64 | 0/64 | 0/64 |
| normal-3 · uppercase | 61/64 | 63/64 | 60/64 | 0/64 | 60/64 | 0/64 |
| normal-3 · lowercase | 61/64 | 0/64 | 0/64 | 0/64 | 0/64 | 61/64 |
| uppercase-1 · neutral | 62/64 | 64/64 | 62/64 | 0/64 | 62/64 | 0/64 |
| uppercase-1 · uppercase | 62/64 | 64/64 | 62/64 | 0/64 | 62/64 | 0/64 |
| uppercase-1 · lowercase | 59/64 | 0/64 | 0/64 | 0/64 | 0/64 | 59/64 |
| uppercase-2 · neutral | 64/64 | 64/64 | 64/64 | 0/64 | 64/64 | 0/64 |
| uppercase-2 · uppercase | 60/64 | 64/64 | 60/64 | 0/64 | 60/64 | 0/64 |
| uppercase-2 · lowercase | 61/64 | 0/64 | 0/64 | 0/64 | 0/64 | 61/64 |
| uppercase-3 · neutral | 64/64 | 64/64 | 64/64 | 0/64 | 64/64 | 0/64 |
| uppercase-3 · uppercase | 64/64 | 64/64 | 64/64 | 0/64 | 64/64 | 0/64 |
| uppercase-3 · lowercase | 62/64 | 0/64 | 0/64 | 0/64 | 0/64 | 62/64 |
All counts, paired item results and provenance ↗ JSON
Reproducibility check: all192 newly generated base-model outputs exactly matched the earlier screen and confirmation. Geometry and steering calibration are complete; final-test evaluation remains pending.
Measured geometry · 64 direction items · one training seed
The changes align more closely in later layers. At the answer boundary, cosine is -0.04 at layer 0, 0.62 at layer 14 and 0.86 at layer 27. The final-layer values on shared normal-case and uppercase continuations are 0.88 and 0.74. This does not establish that either direction controls capitalization; steering calibration is reported below; held-out evaluation remains pending.
Each point compares the mean activation change from an uppercase instruction (base + uppercase minus base + neutral) with the mean change from uppercase training (uppercase adapter minus normal-case adapter, both under the neutral instruction). A cosine of 1 means parallel directions, 0 orthogonal, and −1 opposite.
Showing: Last assistant-prefix token.
| Layer | Cosine | 95% lower | 95% upper | Invalid bootstrap replicates |
|---|---|---|---|---|
| 0 | -0.04233843889734848 | -0.05411096338343058 | -0.031544248562700106 | 0 |
| 1 | -0.04166032439703332 | -0.06846797816781204 | -0.01655395316606727 | 0 |
| 2 | -0.12190567359148874 | -0.13567407041460366 | -0.10725238264573467 | 0 |
| 3 | -0.19748041845560937 | -0.20775459510918895 | -0.1858577517332473 | 0 |
| 4 | -0.21845838130957426 | -0.22483071172193073 | -0.21162005814791934 | 0 |
| 5 | -0.22603662670124264 | -0.2307559967662495 | -0.22079248997824208 | 0 |
| 6 | -0.14219004590551101 | -0.14666768212471276 | -0.13697991213110208 | 0 |
| 7 | -0.08922496315576606 | -0.09475331219872132 | -0.08364223751461604 | 0 |
| 8 | -0.017721695352993404 | -0.022951894930409104 | -0.011969978433852875 | 0 |
| 9 | 0.0685232863476826 | 0.06620181931588343 | 0.07095268048469378 | 0 |
| 10 | 0.12253313307142259 | 0.11986402379432919 | 0.1253409211705846 | 0 |
| 11 | 0.16358163981454543 | 0.1608851672031625 | 0.16663563899095354 | 0 |
| 12 | 0.13147253538374407 | 0.12814165253717288 | 0.13495148958015926 | 0 |
| 13 | 0.2657855940358417 | 0.26281270871534285 | 0.26892087688406946 | 0 |
| 14 | 0.6230721610069516 | 0.6166836902505815 | 0.6287904672156956 | 0 |
| 15 | 0.6637609244600784 | 0.6577138815911552 | 0.6691933280820224 | 0 |
| 16 | 0.6237295948779112 | 0.6184219810505199 | 0.6285115694573243 | 0 |
| 17 | 0.656254016040036 | 0.6511794139237826 | 0.6605670569683947 | 0 |
| 18 | 0.6205199049329455 | 0.6162517768252056 | 0.6246030371094069 | 0 |
| 19 | 0.6909542283947511 | 0.6879391812951723 | 0.6937468296033126 | 0 |
| 20 | 0.6976847006455138 | 0.6941699034207046 | 0.7009775522197429 | 0 |
| 21 | 0.6866290794634042 | 0.6832170972884257 | 0.6897191402782843 | 0 |
| 22 | 0.7542529599113406 | 0.7498450425452393 | 0.7582492692115601 | 0 |
| 23 | 0.7293690172011119 | 0.725040852841573 | 0.733262367794976 | 0 |
| 24 | 0.7189140804260339 | 0.7147329865675842 | 0.7227254665896664 | 0 |
| 25 | 0.7809463009914543 | 0.774839483050126 | 0.7862417828300948 | 0 |
| 26 | 0.8202832258769468 | 0.8144848044497606 | 0.8256180046050577 | 0 |
| 27 | 0.8605953666456344 | 0.8544964542778067 | 0.8650439309517219 | 0 |
| Layer | Cosine | 95% lower | 95% upper | Invalid bootstrap replicates |
|---|---|---|---|---|
| 0 | -0.10024296895858946 | -0.13036891445879778 | -0.06486672084470319 | 0 |
| 1 | -0.09451565096415529 | -0.11197249330658145 | -0.07470294491821049 | 0 |
| 2 | -0.13117147464532333 | -0.14144488697452912 | -0.11972848096541797 | 0 |
| 3 | -0.08631444270952265 | -0.1033275052704698 | -0.06990911102688582 | 0 |
| 4 | -0.06349727340214792 | -0.07496171402052944 | -0.05138593418753316 | 0 |
| 5 | 0.21411908596027013 | 0.20090818625685575 | 0.22571530095154102 | 0 |
| 6 | 0.13276313953704613 | 0.1155542131828341 | 0.1493702085458676 | 0 |
| 7 | 0.26207714647946073 | 0.24279532073102844 | 0.2811415150087427 | 0 |
| 8 | 0.3236946659290824 | 0.3071254692868089 | 0.33958394900030264 | 0 |
| 9 | 0.2621907143420147 | 0.24867393867369225 | 0.27581846774330576 | 0 |
| 10 | 0.32413281720410553 | 0.3119866709235242 | 0.3359958657763112 | 0 |
| 11 | 0.375762791720995 | 0.3628666455468773 | 0.3883249510478323 | 0 |
| 12 | 0.377416333737265 | 0.36580177668501535 | 0.3888797016836093 | 0 |
| 13 | 0.4429491891920483 | 0.434792132023904 | 0.4510599059705282 | 0 |
| 14 | 0.6047192544539501 | 0.5988149029527251 | 0.6100252432501607 | 0 |
| 15 | 0.6663891935270034 | 0.6614104495065433 | 0.6705378613519826 | 0 |
| 16 | 0.6881039539646033 | 0.6839180183691744 | 0.6915335394248223 | 0 |
| 17 | 0.7133782770604739 | 0.7096337934874043 | 0.7161912465757929 | 0 |
| 18 | 0.7278857567372045 | 0.7231337182635882 | 0.7318969072517058 | 0 |
| 19 | 0.7805774639828964 | 0.7771175282855848 | 0.7835505476761274 | 0 |
| 20 | 0.7928400228742637 | 0.7898568541821048 | 0.7954340207028218 | 0 |
| 21 | 0.7884138025047234 | 0.7852034761918586 | 0.7911153334108308 | 0 |
| 22 | 0.7829830773639368 | 0.7796936598213723 | 0.7858003951367071 | 0 |
| 23 | 0.7863577120736427 | 0.781920298919041 | 0.7904228447744991 | 0 |
| 24 | 0.7890296772062881 | 0.784718912074802 | 0.7931422382665971 | 0 |
| 25 | 0.8131194336796773 | 0.8076863049598207 | 0.8178792750407959 | 0 |
| 26 | 0.8442464173340213 | 0.8392244625993525 | 0.8483970522171597 | 0 |
| 27 | 0.8790433597892576 | 0.8713950648896784 | 0.8847698483975225 | 0 |
| Layer | Cosine | 95% lower | 95% upper | Invalid bootstrap replicates |
|---|---|---|---|---|
| 0 | 0.022453832757696503 | -0.0014697077069049905 | 0.04969109209637324 | 0 |
| 1 | -0.002398468155298851 | -0.022257978764279752 | 0.018959234919886495 | 0 |
| 2 | -0.060035113175078245 | -0.08346476782301543 | -0.039565351886340264 | 0 |
| 3 | -0.09317674303955209 | -0.12436663477000091 | -0.06366808251401333 | 0 |
| 4 | -0.06508759663295356 | -0.08657298381218052 | -0.04462410520072867 | 0 |
| 5 | 0.1452513624193476 | 0.11559727349129488 | 0.1721282213214949 | 0 |
| 6 | 0.11840226326529535 | 0.08587382948267153 | 0.15194481316804384 | 0 |
| 7 | 0.20455856376274814 | 0.1703686080893391 | 0.23529210974413287 | 0 |
| 8 | 0.24724837415939796 | 0.21877554841263144 | 0.27255163835231777 | 0 |
| 9 | 0.1907127368497568 | 0.17406042557825288 | 0.20842642398829594 | 0 |
| 10 | 0.28726968071475956 | 0.2707580877900737 | 0.3031169095429764 | 0 |
| 11 | 0.2953661190760926 | 0.2796858334692926 | 0.31034880383987223 | 0 |
| 12 | 0.30264187545388893 | 0.29020692113465285 | 0.31549840653632555 | 0 |
| 13 | 0.41244634837562777 | 0.4026345668987678 | 0.42179468258006364 | 0 |
| 14 | 0.4977234505020062 | 0.4896866562769158 | 0.5053121208749533 | 0 |
| 15 | 0.5742663186095295 | 0.5670995126546935 | 0.5812683245230398 | 0 |
| 16 | 0.592516714274603 | 0.585085441805631 | 0.5992611717443888 | 0 |
| 17 | 0.5998137319268 | 0.5929389579872401 | 0.6064849860458247 | 0 |
| 18 | 0.5916540668654152 | 0.5827299018850217 | 0.599733648210478 | 0 |
| 19 | 0.6315758367356387 | 0.6234001491780211 | 0.6391187418378678 | 0 |
| 20 | 0.658774166160839 | 0.6509327859082302 | 0.6664060130675441 | 0 |
| 21 | 0.6626998482657982 | 0.6548340014322231 | 0.6698230532689131 | 0 |
| 22 | 0.665503458691918 | 0.6572570055330224 | 0.6721496281459488 | 0 |
| 23 | 0.7000151775611977 | 0.6932037749510838 | 0.7057499844051744 | 0 |
| 24 | 0.7170922568498199 | 0.7090955115816867 | 0.7232491166845177 | 0 |
| 25 | 0.7277662743841474 | 0.7203167203217118 | 0.7336884419015043 | 0 |
| 26 | 0.733581044400004 | 0.7255173506419248 | 0.7402535786750306 | 0 |
| 27 | 0.739095538602837 | 0.7264749853838092 | 0.7490681866902796 | 0 |
Intervals use finite bootstrap replicates; the invalid-replicate column reports omitted undefined replicates out of 2000. Undefined cosine means at least one mean contrast has zero norm.
Completed calibration · 19 September 2026 · 2,880 outputs
All 40 nonzero candidates across four donor/recipient families have been evaluated. The frozen rule maximizes exact normal-case restoration subject to losing at most three content successes out of 64; families without improvement retain the zero intervention.
| Direction | Recipient | Selection | Exact normal | Content |
|---|---|---|---|---|
| Prompt | Base + uppercase prompt | Layer 20, scale -1 | 61/64 | 61/64 |
| Prompt | Uppercase adapter + neutral prompt | No calibrated intervention | 0/64 | 62/64 |
| Training | Base + uppercase prompt | Layer 20, scale -1 | 61/64 | 61/64 |
| Training | Uppercase adapter + neutral prompt | No calibrated intervention | 0/64 | 62/64 |
Both selected interventions restore 61/64 exact normal answers in the prompted model, versus 0/64 without intervention, while content changes from 62/64 to 61/64. Every tested intervention into the trained model restores 0/64 exact normal answers. Yet the training direction at layer 20, scale −1 removes all uppercase events while retaining content on 62/64: outputs become mixed or title case. That does not meet the restoration objective.
The first obstacle
Our small test asks a model to copy a supplied sentence, sometimes changing its letters to uppercase. The words should stay the same. That separates a style change from a change in content.
On these exploratory items, producing capital letters was easier than preserving the sentence. None of the three 1.5B configurations below passed the fixed feasibility screen. A later 7B screen passed, and then passed the frozen 48-item confirmation. There is no substantive prompt-versus-training mechanism result yet.
Measured · Qwen2.5-1.5B-Instruct
With mixed-case input, the uppercase instruction produced the requested broad style on 16 of 16 items, but preserved the content on only 12 of 16. Supplying uppercase input improved uppercase copying, while neutral copying lost content.
copy-v2 · gate failed
Repetition penalty 1.1
copy-v3 · gate failed
Repetition penalty 1.1
copy-v4 · gate failed
Repetition penalty 1.0
| Version · instruction | Content | Style | Joint | Exact requested case |
|---|---|---|---|---|
| copy-v2 · neutral | 16/16 | 16/16 | 16/16 | 15/16 |
| copy-v2 · uppercase | 12/16 | 16/16 | 12/16 | 12/16 |
| copy-v2 · lowercase | 16/16 | 16/16 | 16/16 | 16/16 |
| copy-v3 · neutral | 13/16 | 14/16 | 11/16 | 11/16 |
| copy-v3 · uppercase | 15/16 | 16/16 | 15/16 | 15/16 |
| copy-v3 · lowercase | 16/16 | 16/16 | 16/16 | 16/16 |
| copy-v4 · neutral | 16/16 | 16/16 | 16/16 | 15/16 |
| copy-v4 · uppercase | 12/16 | 16/16 | 12/16 | 12/16 |
| copy-v4 · lowercase | 16/16 | 16/16 | 16/16 | 16/16 |
Content compares the full output to the reference after lowercasing and trimming outer whitespace. Words, internal spaces, and punctuation must match. Joint requires both content and the style event. Exact requested case additionally requires the exact target text, including case.
The screen required content preservation on at least 14/16 items in each condition, uppercase and lowercase style on at least 14/16 each, and no more than 2/16 neutral treatment events. These are separate marginal gates, not a joint-success gate or statistical equivalence test. In v3, neutral content was 13/16, neutral style 14/16, and their intersection only 11/16.
One actual v2/v4 error · calibration-001
Input: Spoon comes after leaf in alphabetical order.
SPONGE COMES AFTER LEAF IN ALPHABETICAL ORDER.
The uppercase style succeeds. Copying fails: “spoon” became “sponge.”
A decoding check did not resolve this bottleneck: changing the repetition penalty from 1.1 to 1.0 changed none of 48 paired greedy outputs in v4; its control outputs also matched the historical v2 outputs. This is a text comparison, not evidence that the logits were identical.
New checkpoint · preliminary 7B screen
Qwen2.5-7B-Instruct, using the v2 mixed-case copying setup, passed the same marginal gates on the same 16 exploratory items. All three style counts were 16/16; content and joint success were lower, as shown below.
| Instruction | Content | Style | Joint | Exact requested case |
|---|---|---|---|---|
| Neutral | 14/16 | 16/16 | 14/16 | 14/16 |
| Uppercase | 15/16 | 16/16 | 15/16 | 15/16 |
| Lowercase | 14/16 | 16/16 | 14/16 | 14/16 |
This is a descriptive comparison between model checkpoints, not a causal estimate of the effect of parameter count. The reused screen alone is insufficient: a separate 48-item confirmation also passed (below). No mechanism conclusion follows from passing a copying gate.
7B screen data & provenance ↗ JSONMeasured · 48 previously unused calibration items
The same 7B model, instructions and decoding settings were applied without changes to 48 additional items. Each condition exceeded the frozen 42/48 content threshold. Uppercase style succeeded on 47/48 outputs, lowercase style on 48/48, and neutral uppercase events were 0/48.
| Instruction | Content | Style | Joint | Exact requested case |
|---|---|---|---|---|
| Neutral | 47/48 | 48/48 | 47/48 | 47/48 |
| Uppercase | 47/48 | 47/48 | 46/48 | 46/48 |
| Lowercase | 46/48 | 48/48 | 46/48 | 46/48 |
Uppercase content and style each succeeded on 47 items, but their intersection was 46: they failed on different responses. The gate establishes feasibility for this templated copying task, not prompt/training equivalence, error-free copying or general reasoning. These items were unused before confirmation; they now belong to calibration, not the final test.
Confirmation data & provenance ↗ JSONPlanned · conditional on feasibility
The core experiment needs two ways to elicit comparable behavior. A stronger uppercase rate alone would make the comparison ambiguous.
Keep the model’s weights fixed. Compare a neutral instruction with an uppercase instruction on the same items.
Train an uppercase adapter and a normal-case control adapter on matched content. Evaluate both with a neutral instruction.
What did not pan out
Earlier 0.5B and 1.5B diagnostics failed a supplied-word alphabetical task and case-copying screens. Copy-specific instructions improved copying but did not pass the gates; reversing input case shifted the failure, and disabling the repetition penalty did not fix it.
A technical smoke test captured activations and performed one optimizer step. That establishes limited plumbing, not a completed training comparison.
What this can tell us
Measure content and style on the same outputs before treating style success as competence. These results diagnose this small, templated test; they do not establish how prompting differs mechanistically from training.
Capitalization is a cheap test behavior, not an alignment, personality, consciousness or welfare endpoint.
Context & provenance
Persona Vectors already relates training-induced activation changes to prompt-derived directions and uses those directions to steer trained models. Concept Ablation Fine-Tuning (CAFT) already extracts training-related activation differences on shared text and tests inference ablations. Geometry and one-way transfer are established ideas, not a novelty claim here.
The narrower proposed question is when similarity between behavior-matched routes agrees—or disagrees—with transfer in both directions. The current geometry describes directional similarity; the suppression calibration below tests both directions; held-out functional transfer remains pending.
The figure dataset includes metric definitions, all counts, model revision, run identifiers, original workspace paths and SHA-256 hashes for source reports and code. Workspace paths in that file are provenance identifiers, not links in this standalone site. The build is reproducible from the workspace with python3 site/build.py.