Valid target tokens seen
Approved continued-training run scale.
Measured evidence
Loss signals, stability controls, generation gates, reload behavior, and token-scale receipts.
Recorded finding: The training continuation path is instrumented and produces governed stability evidence.
Decision signals
Each value is mapped directly from the recorded summary artifact and retains its experimental, historical, or control context.
Approved continued-training run scale.
First-to-last public aggregate evaluation signal.
All published reload checks passed.
6 hard failures across the approved prompts.
Source-linked visual evidence
Charts are generated from normalized source fields. Every visualization includes its exact values in an accessible table.
Status matrix
Passed and hard-failure counts in the current public generation gate.
Finding: The current generation gate fails; no prompt in the published set passed. That held result remains visible.
| Group | Count |
|---|---|
| Passed | 0 prompts |
| Hard failures | 6 prompts |
Method
Generation quality and broad model capability remain under review and are not mature public claims.
Inspection and reuse
Source data
The recorded JSON artifact is staged from the governed public evidence pack during the build. Its currency notice determines whether it may be read as current.
Open source JSONProvenance
Use the public build receipt to compare the deployed site source, evidence digest, claim posture, and inference release state.
Open build receiptContinue reviewing
Decision relevance: research · engineering
Review evidenceDecision relevance: research · diligence
Review evidence