Interactive lessonRetained evidence only
How much can one random seed hide?
Select retained runs and watch the summary change. This lesson replays fixed trial-a results in your browser. It does not train a model, execute research code, or earn a reproduction label.
← Back to LabsLearning objectivesAbout 20 minutes
Read variation.
Limit the claim.
By the end, you should be able to explain why one successful seed is weak evidence, summarize a selected set of runs, and distinguish replaying retained numbers from independently reproducing an experiment.
- 01Select evidence
Compare a single retained run with a larger retained sample.
- 02Read the spread
Track the live mean, range, and number of selected seeds.
- 03Check a claim
Choose the statement that stays inside this evidence boundary.
Evidence workbenchTrial A · 12 retained seeds
Change the sample.
Watch the story move.
The vertical scale starts at 90% so the small differences are legible. Every bar carries its exact rounded accuracy; color is never the only selection cue.
The interactive enhancement is unavailable. The complete retained set and its static summary remain readable below; selection and answer checking are disabled.
- Selected runs
- of 12
- Mean accuracy
- selected runs
- Accuracy range
- min–max
12 runs selected. Mean accuracy 93.06%. Range 91.67 to 95.00%.
Read the retained values as a table
| Seed | Test accuracy |
|---|---|
| 11 | 0.933333333333 |
| 29 | 0.916666666667 |
| 47 | 0.933333333333 |
| 71 | 0.95 |
| 101 | 0.933333333333 |
| 131 | 0.916666666667 |
| 167 | 0.933333333333 |
| 199 | 0.933333333333 |
| 239 | 0.933333333333 |
| 281 | 0.916666666667 |
| 337 | 0.933333333333 |
| 397 | 0.933333333333 |
Claim checkStay inside the evidence
What can this page support?
Choose the strongest claim that the retained trial-a values actually support. “Strongest” does not mean “most exciting.”
Reveal the worked solution
It names the observed set, reports only its measured range, and does not generalize beyond the retained evidence. The first claim is false because this browser lesson performs no training or independent execution. The third predicts unseen data that these 12 stored values cannot establish.
A careful next question is: would the range stay similar under a preregistered run in an independently controlled environment?
LimitationsDo not overclaim
Useful lesson.
Narrow evidence.
- RetrospectiveThe canary protocol was not preregistered.
- Single contextOne model, dataset split, metric, environment, and retained trial-a series.
- Stored valuesChanging the selection recomputes descriptive statistics only; it does not create new observations.
- No scientific reproductionThis page cannot earn “Results reproduced” or “Results replicated.”
ProvenanceExact retained bundle
Trace the lesson
to its evidence.
Read the research status and source boundary →- Evidence slice
- Trial A · 12 fixed random seeds · test accuracy
- Bundle SHA-256
71e4a800fee459cd3af667fdc4e20ecac88e956b1d2afd6fc123424f33d74065- Page behavior
- Local arithmetic over values embedded in this document; no fetch, storage, upload, or execution service.
- Earned label
- Executed source bundle The lesson itself earns no reproduction label.