L3
Verified — hub re-ran it
the hub re-executed the code and the headline numbers held
MABWiser — epsilon-greedy vs Thompson on a 3-arm Bernoulli design, best-arm pull rate
Automated re-run of the headline result of 'MABWiser — epsilon-greedy vs Thompson on a 3-arm Bernoulli design, best-arm pull rate' (arXiv:1909.04412) from its own repository. Pre-registered claim: best_arm_pull_rate_thompson = 0.9 (±15%). Hub verdict: REPRODUCED.
Claims
· unverified
c1
performance
The paper's own code (https://github.com/fidelity/mabwiser) reproduces best_arm_pull_rate_thompson = 0.9 for 'MABWiser — epsilon-greedy vs Thompson on a 3-arm Bernoulli design, best-arm pull rate'.
system ts-mabwiser-sim
metric best_arm_pull_rate_thompson
value 0.9
unit best_arm_pull_rate_thompson
higher_is_better True
hardware cpu-box
Artifacts 2 files · code, data, logs — integrity-checked
| role | location | size | integrity |
|---|---|---|---|
| code | https://github.com/fidelity/mabwiser | — | unchecked |
| paper | https://arxiv.org/pdf/1909.04412 | — | unchecked |
Verification runs 2 run(s) · mode script · 1 machine-checked assertions
failed · runner hub-local · level→L3 · 2026-07-02T07:47:22Z
runner log →
passed · runner hub:repro-study · level→L3 · 2026-07-02T07:47:22Z
runner log →
| claim | check | expected | actual | |
|---|---|---|---|---|
| c1 | best_arm_pull_rate_thompson approx | 0.9 | 0.986 | ✅ |