@ f8027fd0eaeea54caa13c31d31b9fdc459c38b49@ a10cc1512eabd3dde888204e902eca88bddb4951Included summaries Content and scope, not a quality or risk grade
| Record or summary | This report |
|---|---|
| Measurement summary | Included |
| Repeat-check summary | Not included |
| Same-artifact control summary | Not included |
| Underlying records | Retained in the VG1 archive; not included in this copy |
| Operational receipt | Linked · limited operational record |
| Statistical noise envelope | Not certified |
Repeat and control details
What this means for you
How the finding was decided
The reported internal finding uses the fixed, thresholded decision rule recorded for mv-1.4; it is not a test for any non-zero floating-point value. This rule is not, by itself, a measured noise floor for this model.
Text output is evaluated separately under exact greedy token-sequence comparison. The outcome is on the fixed input set, not a population rate or a capability score.
Structural coverage: 24 contributing probes per evaluated position. This is neither the number of output continuations nor a count of independent tasks.
Observed responses through the model
Boundary unresolved — difference already present at the first measured point (decoder block 4 output). Earlier positions were not evaluated, so the beginning of the difference cannot be located.
29 of 29 evaluated positions show a difference. This count describes the reach of the observed response, not its size or the number of changed weights.
| Recorded output | Observation |
|---|---|
| embedding output through decoder block 3 output | not evaluated |
| decoder block 4 output through decoder block 31 output | difference observed |
| final normalized output | difference observed |
Blocks are numbered from 1. Final is the output after the last normalization; the last block before normalization is not separately displayed. Not evaluated means no observation was made there; it is not a below-threshold result.
A difference can propagate from one output to later outputs. This view does not locate edited weights, identify a cause or show which capabilities changed. Controlled small-model findings are not a validation result for this model or execution environment.
What differs between A and B
| Artifact metadata | Recorded comparison |
|---|---|
| Model configuration semantics | different |
| Tokenizer semantics | different |
| Generation settings semantics | different |
Recorded relation: combined artifact change. Do not attribute the finding to fine-tuning or weights alone.
Matched refers only to the recorded semantic summaries. It does not mean the complete artifacts are identical. Different configurations, tokenizers or generation settings can matter alongside weight changes.
Measured weight changes Same artifact pair and recorded response measurement
Overall relative weight change: 7.938%. Internal response: Difference observed. Text output: Withheld: output contracts differ.
Coverage: 290 unique parameter tensors; 361821120 values. All declared floating parameters were checked; 1 shared names were counted once. Buffers are outside this weight summary.
Numerical validity: all checked values are finite. Observed precision: bfloat16. Zero values: reference 0, candidate 0.
| Parameter group | Relative change | Absolute L2 change | Recorded response after the block |
|---|---|---|---|
| Decoder block 1 | 5.541% | 28.53 | Not evaluated |
| Decoder block 2 | 6.879% | 38.43 | Not evaluated |
| Decoder block 3 | 7.122% | 39.34 | Not evaluated |
| Decoder block 4 | 7.078% | 39.02 | Difference observed |
| Decoder block 5 | 6.997% | 38.61 | Difference observed |
| Decoder block 6 | 7.220% | 39.79 | Difference observed |
| Decoder block 7 | 7.224% | 39.73 | Difference observed |
| Decoder block 8 | 7.209% | 39.59 | Difference observed |
| Decoder block 9 | 7.405% | 40.25 | Difference observed |
| Decoder block 10 | 7.212% | 39.59 | Difference observed |
| Decoder block 11 | 7.386% | 41.16 | Difference observed |
| Decoder block 12 | 7.463% | 41.53 | Difference observed |
| Decoder block 13 | 7.325% | 41.08 | Difference observed |
| Decoder block 14 | 7.097% | 40.11 | Difference observed |
| Decoder block 15 | 7.367% | 41.11 | Difference observed |
| Decoder block 16 | 7.636% | 42.67 | Difference observed |
| Decoder block 17 | 7.711% | 43.02 | Difference observed |
| Decoder block 18 | 7.688% | 41.92 | Difference observed |
| Decoder block 19 | 7.874% | 43.22 | Difference observed |
| Decoder block 20 | 7.684% | 43.04 | Difference observed |
| Decoder block 21 | 8.110% | 44.86 | Difference observed |
| Decoder block 22 | 8.100% | 45.27 | Difference observed |
| Decoder block 23 | 7.593% | 42.47 | Difference observed |
| Decoder block 24 | 8.249% | 46.32 | Difference observed |
| Decoder block 25 | 8.310% | 46.50 | Difference observed |
| Decoder block 26 | 8.403% | 48.09 | Difference observed |
| Decoder block 27 | 8.228% | 47.55 | Difference observed |
| Decoder block 28 | 8.065% | 46.67 | Difference observed |
| Decoder block 29 | 8.266% | 47.85 | Difference observed |
| Decoder block 30 | 7.998% | 46.26 | Difference observed |
| Decoder block 31 | 7.576% | 43.82 | Difference observed |
| Decoder block 32 | 7.042% | 39.76 | Final normalized output: difference observed |
| Other registered parameters | 12.06% | 100.1 | No matching block output |
Relative change uses the reference norm; a zero reference has no relative percentage. Shared parameters belong to their first recorded group. The last block output is not separately measured; its row shows the final normalized output. Weight changes locate parameters to inspect, while response differences can propagate from earlier blocks. Neither identifies the cause of a task outcome.
Use the changed groups to choose targeted follow-up evaluations. No observed response difference does not establish equivalence, and a larger weight change is not a quality or severity score.
Recommended next checks Guidance, not a measurement
- Control the comparison contractReview the metadata differences above. Establish comparable tokenizer and generation settings before interpreting an output comparison; preserve each artifact revision.
- Run the relevant regression setCompare A and B on the tasks your users rely on, including output format and failure cases. Set acceptance criteria before inspecting the results.
- Keep an auditable decisionRetain this report with the task results and deployment revision. Separate the measured observation from your release decision.
Provenance and record
receipt xrr_jDdntAQqtgwrCe9UHIFs6exqXY91dJj7Fuk0Q7876ck · opaque operational record with limited assurance; scope at https://www.tetracta.ai/llm_tomografi/attest/receipt/xrr_jDdntAQqtgwrCe9UHIFs6exqXY91dJj7Fuk0Q7876ck
Recorded context
Pin these fields when you diff this report against a future scan. Matching context is a prerequisite for interpretation, not proof of the cause of a difference. The public receipt is a limited signed locator. The account owner can retrieve the full signed operational record from the report page.
| Measurement time (UTC) | 2026-09-22T08:26:55Z |
| Model reference | HuggingFaceTB/SmolLM2-360M -> HuggingFaceTB/SmolLM2-360M-Instruct |
| Artifact evidence | Artifact-identity and measurement records are held in the owner-only operational record |
| HF revision | A: f8027fd0eaeea54caa13c31d31b9fdc459c38b49 · B: a10cc1512eabd3dde888204e902eca88bddb4951 |
| Probe set / scan config | ps-1.1 / sc-1.0 · mv-1.4 |
| Reference set | no release reference set; estimator validation pending |
| Report schema | rs-1.7 |
| Compute isolation | managed compute; infrastructure identity is withheld |
| Runtime envelope | fixed release configuration; bitwise equality across device types is not guaranteed |
| Operational receipt | receipt xrr_jDdntAQqtgwrCe9UHIFs6exqXY91dJj7Fuk0Q7876ck · opaque operational record with limited assurance; scope at https://www.tetracta.ai/llm_tomografi/attest/receipt/xrr_jDdntAQqtgwrCe9UHIFs6exqXY91dJj7Fuk0Q7876ck |
Reading this report alongside another scan
Check the artifact pair, revisions and measurement contract first. The categorical findings describe each recorded comparison; they do not rank different models. A different contract or execution environment needs its own comparability evidence. No numerical severity score is supplied.
Use as a checkpoint reference
For supported checkpoint comparisons, retain the starting revision, this measurement contract and the output settings with your training record. Compare the next checkpoint against that same reference, then retain its task results alongside the scan.
| For the next checkpoint | Record alongside this scan |
|---|---|
| Reference and candidate | Exact repository revisions |
| Comparable measurement | Same instrument, input contract and approved execution profile |
| Training change | Your fine-tune, merge or pruning record |
| Adoption decision | Your task acceptance results and deployment revision |
A reference record supports change tracking. It does not define an optimal model or a target score for training. Simulation results describe the simulated state; actual checkpoint changes require their own before/after comparison.
Your findings and the protected instrument
The observation, output-text status, artifact context and, where applicable, metadata comparison, plus provenance and categorical depth detail when included.
Raw profiles, per-item records, probe construction, thresholds, transformations and internal traces. The report contains no model weights, credentials or infrastructure identifiers.
This descriptive result does not identify cause, capability, quality, safety or deployment impact. It is not a task-quality measurement. Do not target a layer or make a shipping decision from this report alone. No calibrated risk or pass/fail grade is assigned. No cross-model ranking or cross-device bitwise-equality claim is made.
Historical results under superseded estimators do not validate this release. The categorical depth view, when included, describes recorded responses; it is not a diagnosis of edited weights.