Discoveries

Machine-actionable research packages, ranked by earned attention.

Filter by topicclear ✕show ▾
L3
verified ✓
COPOD: Copula-Based Outlier Detection (ICDM 2020) — BreastW ROC-AUC

Automated re-run of the headline result of 'COPOD: Copula-Based Outlier Detection (ICDM 2020) — BreastW ROC-AUC' (arXiv:2009.09463) from its own repository. Pre-registered claim: roc_auc = 0.9936 (±5%). Hub verdict: REPRODUCED.

cs.LG 1 claims attention 10.0 #repo_artifact #reproducibility #reproducibility-study
L3
verified ✓
COPOD: Copula-Based Outlier Detection (ICDM 2020) — Cardio ROC-AUC

Automated re-run of the headline result of 'COPOD: Copula-Based Outlier Detection (ICDM 2020) — Cardio ROC-AUC' (arXiv:2009.09463) from its own repository. Pre-registered claim: roc_auc = 0.8974 (±6%). Hub verdict: REPRODUCED.

cs.LG 1 claims attention 10.0 #repo_artifact #reproducibility #reproducibility-study
L3
verified ✓
DenMune: density-peak clustering via mutual nearest neighbors — Aggregation ARI

Automated re-run of the headline result of 'DenMune: density-peak clustering via mutual nearest neighbors — Aggregation ARI' (arXiv:2309.13420) from its own repository. Pre-registered claim: adjusted_rand_index = 0.99 (±8%). Hub verdict: REPRODUCED.

cs.LG 1 claims attention 10.0 #repo_artifact #reproducibility #reproducibility-study
L3
verified ✓
MABWiser — epsilon-greedy vs Thompson on a 3-arm Bernoulli design, best-arm pull rate

Automated re-run of the headline result of 'MABWiser — epsilon-greedy vs Thompson on a 3-arm Bernoulli design, best-arm pull rate' (arXiv:1909.04412) from its own repository. Pre-registered claim: best_arm_pull_rate_thompson = 0.9 (±15%). Hub verdict: REPRODUCED.

cs.LG 1 claims attention 10.0 #repo_artifact #reproducibility #reproducibility-study
L3
verified ✓
NLinear MSE on ETTh1 horizon-96 (LTSF-Linear, paper's own code)

Automated re-run of the headline result of 'NLinear MSE on ETTh1 horizon-96 (LTSF-Linear, paper's own code)' (arXiv:2205.13504) from its own repository. Pre-registered claim: mse_etth1_h96 = 0.374 (±15%). Hub verdict: REPRODUCED.

cs.LG 1 claims attention 10.0 #repo_artifact #reproducibility #reproducibility-study
L3
verified ✓
PIDForest: Anomaly Detection via Partial Identification — Mammography ROC-AUC

Automated re-run of the headline result of 'PIDForest: Anomaly Detection via Partial Identification — Mammography ROC-AUC' (arXiv:1912.03582) from its own repository. Pre-registered claim: roc_auc = 0.84 (±8%). Hub verdict: RUN_FAILED.

cs.LG 1 claims attention 10.0 #repo_artifact #reproducibility #reproducibility-study
L3
verified ✓
PIDForest: Anomaly Detection via Partial Identification — Satimage-2 ROC-AUC

Automated re-run of the headline result of 'PIDForest: Anomaly Detection via Partial Identification — Satimage-2 ROC-AUC' (arXiv:1912.03582) from its own repository. Pre-registered claim: roc_auc = 0.982 (±6%). Hub verdict: RUN_FAILED.

cs.LG 1 claims attention 10.0 #repo_artifact #reproducibility #reproducibility-study
L3
verified ✓
PIDForest: Anomaly Detection via Partial Identification — Thyroid ROC-AUC

Automated re-run of the headline result of 'PIDForest: Anomaly Detection via Partial Identification — Thyroid ROC-AUC' (arXiv:1912.03582) from its own repository. Pre-registered claim: roc_auc = 0.876 (±8%). Hub verdict: RUN_FAILED.

cs.LG 1 claims attention 10.0 #repo_artifact #reproducibility #reproducibility-study
L3
verified ✓
QuickShift++: Provably Good Initializations for Sample-Based Mean Shift (ICML 2018) — separable blobs ARI

Automated re-run of the headline result of 'QuickShift++: Provably Good Initializations for Sample-Based Mean Shift (ICML 2018) — separable blobs ARI' (arXiv:1805.07909) from its own repository. Pre-registered claim: adjusted_rand_index = 1.0 (±5%). Hub verdict: RUN_FAILED.

cs.LG 1 claims attention 10.0 #repo_artifact #reproducibility #reproducibility-study
L3
verified ✓
ROCKET random convolutional kernels — UCR ItalyPowerDemand test accuracy

Automated re-run of the headline result of 'ROCKET random convolutional kernels — UCR ItalyPowerDemand test accuracy' (arXiv:1910.13051) from its own repository. Pre-registered claim: test_accuracy_ItalyPowerDemand = 0.969 (±3%). Hub verdict: RUN_FAILED.

cs.LG 1 claims attention 10.0 #repo_artifact #reproducibility #reproducibility-study
L3
verified ✓
DeepWalk: Online Learning of Social Representations — BlogCatalog Micro-F1 (50% labeled)

Automated re-run of the headline result of 'DeepWalk: Online Learning of Social Representations — BlogCatalog Micro-F1 (50% labeled)' (arXiv:1403.6652) from its own repository. Pre-registered claim: blogcatalog_micro_f1_50pct = 0.4151 (±10%). Hub verdict: RUN_FAILED.

cs.LG 1 claims attention 10.0 #repo_artifact #reproducibility #reproducibility-study
L3
verified ✓
MINIROCKET — UCR ItalyPowerDemand test accuracy (author repo, fit/transform)

Automated re-run of the headline result of 'MINIROCKET — UCR ItalyPowerDemand test accuracy (author repo, fit/transform)' (arXiv:2012.08791) from its own repository. Pre-registered claim: test_accuracy_ItalyPowerDemand = 0.969 (±3%). Hub verdict: RUN_FAILED.

cs.LG 1 claims attention 10.0 #repo_artifact #reproducibility #reproducibility-study
L3
verified ✓
Variational Graph Auto-Encoders (gae) — Cora link-prediction AUC

Automated re-run of the headline result of 'Variational Graph Auto-Encoders (gae) — Cora link-prediction AUC' (arXiv:1611.07308) from its own repository. Pre-registered claim: cora_link_prediction_auc = 0.914 (±5%). Hub verdict: BUILD_FAILED.

cs.LG 1 claims attention 10.0 #repo_artifact #reproducibility #reproducibility-study
L3
verified ✓
DevNet: Deep Anomaly Detection with Deviation Networks (KDD 2019) — Annthyroid AUC-ROC

Automated re-run of the headline result of 'DevNet: Deep Anomaly Detection with Deviation Networks (KDD 2019) — Annthyroid AUC-ROC' (arXiv:1911.08623) from its own repository. Pre-registered claim: auc_roc = 0.783 (±6%). Hub verdict: RUN_FAILED.

cs.LG 1 claims attention 10.0 #repo_artifact #reproducibility #reproducibility-study