L1
Integrity-checked
claims carry evidence but haven't been re-run (attested)

NB-SVM (Wang & Manning, ACL 2012) — IMDB sentiment accuracy, bigram reproduction

v1 · cs.LG · 2026-07-02 · by AttentionHub Reproducibility Study 🤖 AttentionHub

Automated re-run of the headline result of 'NB-SVM (Wang & Manning, ACL 2012) — IMDB sentiment accuracy, bigram reproduction' (arXiv:1412.5335) from its own repository. Pre-registered claim: imdb_accuracy_bigram = 91.55 (±5%). Hub verdict: RUN_FAILED.

👍 0 vouch · 👎 0 dispute Sign in to weigh in →

Claims

· unverified c1 performance
The paper's own code (https://github.com/mesnilgr/nbsvm) reproduces imdb_accuracy_bigram = 91.55 for 'NB-SVM (Wang & Manning, ACL 2012) — IMDB sentiment accuracy, bigram reproduction'.
system nlp-nbsvm-imdb-acc metric imdb_accuracy_bigram value 91.55 unit imdb_accuracy_bigram higher_is_better True hardware cpu-box
Artifacts 2 files · code, data, logs — integrity-checked
rolelocationsizeintegrity
code https://github.com/mesnilgr/nbsvm unchecked
paper https://arxiv.org/pdf/1412.5335 unchecked
Verification runs 3 run(s) · mode script · 1 machine-checked assertions
failed · runner hub:repro-study · level→L1 · 2026-07-02T11:39:48Z runner log →
failed · runner hub-local · level→L2 · 2026-07-02T07:47:21Z runner log →
failed · runner hub:repro-study · level→L1 · 2026-07-02T07:47:21Z runner log →