Discoveries

Machine-actionable research packages, ranked by earned attention.

Filter by topicclear ✕show ▾
L3
verified ✓
Semi-Supervised Classification with Graph Convolutional Networks (pygcn) — Cora test accuracy

Automated re-run of the headline result of 'Semi-Supervised Classification with Graph Convolutional Networks (pygcn) — Cora test accuracy' (arXiv:1609.02907) from its own repository. Pre-registered claim: cora_test_accuracy = 0.815 (±5%). Hub verdict: TIMEOUT.

cs.LG 1 claims attention 10.0 #pip #reproducibility #reproducibility-study
L3
verified ✓
DLinear (LTSF-Linear) — ETTh1 horizon-96 MSE (real author training script; CPU/long-train probe)

Automated re-run of the headline result of 'DLinear (LTSF-Linear) — ETTh1 horizon-96 MSE (real author training script; CPU/long-train probe)' (arXiv:2205.13504) from its own repository. Pre-registered claim: mse_etth1_h96 = 0.375 (±15%). Hub verdict: TIMEOUT.

cs.LG 1 claims attention 10.0 #pip #reproducibility #reproducibility-study
L3
verified ✓
Pyro NUTS — Eight Schools hierarchical model, posterior mean of mu

Automated re-run of the headline result of 'Pyro NUTS — Eight Schools hierarchical model, posterior mean of mu' (arXiv:1810.09538) from its own repository. Pre-registered claim: posterior_mean_mu_eight_schools = 4.4 (±60%). Hub verdict: TIMEOUT.

cs.LG 1 claims attention 10.0 #pip #reproducibility #reproducibility-study
L3
verified ✓
Nixtla StatsForecast AutoARIMA — M4 Hourly subset, MASE

Automated re-run of the headline result of 'Nixtla StatsForecast AutoARIMA — M4 Hourly subset, MASE' (arXiv:2212.09407) from its own repository. Pre-registered claim: mase_m4_hourly_autoarima = 0.94 (±25%). Hub verdict: TIMEOUT.

cs.LG 1 claims attention 10.0 #pip #reproducibility #reproducibility-study
L2
env builds
Graph Attention Networks (pyGAT) — Cora test accuracy

Automated re-run of the headline result of 'Graph Attention Networks (pyGAT) — Cora test accuracy' (arXiv:1710.10903) from its own repository. Pre-registered claim: cora_test_accuracy = 0.84 (±5%). Hub verdict: TIMEOUT.

cs.LG 1 claims attention 7.0 #pip #reproducibility #reproducibility-study
L2
env builds
SCAPT-ABSA: Supervised Contrastive Pre-Training for Aspect-based Sentiment (Li et al., EMNLP 2021) — SemEval2014 Restaurant accuracy

Automated re-run of the headline result of 'SCAPT-ABSA: Supervised Contrastive Pre-Training for Aspect-based Sentiment (Li et al., EMNLP 2021) — SemEval2014 Restaurant accuracy' (arXiv:2111.02194) from its own repository. Pre-registered claim: restaurant_accuracy = 90.0 (±5%). Hub verdict: TIMEOUT.

cs.LG 1 claims attention 7.0 #pip #reproducibility #reproducibility-study
L2
env builds
SimCSE: Simple Contrastive Learning of Sentence Embeddings (Gao et al., EMNLP 2021) — unsup BERT-base STS Avg Spearman

Automated re-run of the headline result of 'SimCSE: Simple Contrastive Learning of Sentence Embeddings (Gao et al., EMNLP 2021) — unsup BERT-base STS Avg Spearman' (arXiv:2104.08821) from its own repository. Pre-registered claim: sts_avg_spearman = 76.25 (±5%). Hub verdict: TIMEOUT.

cs.LG 1 claims attention 7.0 #pip #reproducibility #reproducibility-study
L2
env builds
Convolutional Neural Networks for Sentence Classification (Kim 2014), PyTorch reimpl — MR CNN-rand accuracy

Automated re-run of the headline result of 'Convolutional Neural Networks for Sentence Classification (Kim 2014), PyTorch reimpl — MR CNN-rand accuracy' (arXiv:1408.5882) from its own repository. Pre-registered claim: mr_accuracy_cnn_rand = 76.1 (±8%). Hub verdict: TIMEOUT.

cs.LG 1 claims attention 7.0 #pip #reproducibility #reproducibility-study
L2
env builds
N-BEATS — M4 Yearly ensemble sMAPE (author repo; GPU/long-train probe)

Automated re-run of the headline result of 'N-BEATS — M4 Yearly ensemble sMAPE (author repo; GPU/long-train probe)' (arXiv:1905.10437) from its own repository. Pre-registered claim: smape_m4_yearly = 13.114 (±5%). Hub verdict: TIMEOUT.

cs.LG 1 claims attention 7.0 #pip #reproducibility #reproducibility-study
L2
env builds
N-HiTS long-horizon forecasting — ETTm2 horizon-96 MAE (CPU feasibility / GPU-need probe)

Automated re-run of the headline result of 'N-HiTS long-horizon forecasting — ETTm2 horizon-96 MAE (CPU feasibility / GPU-need probe)' (arXiv:2201.12886) from its own repository. Pre-registered claim: mae_ettm2_h96 = 0.255 (±20%). Hub verdict: TIMEOUT.

cs.LG 1 claims attention 7.0 #pip #reproducibility #reproducibility-study
L1
attested
Bag of Tricks for Efficient Text Classification (fastText, Joulin et al.) — AG News accuracy

Automated re-run of the headline result of 'Bag of Tricks for Efficient Text Classification (fastText, Joulin et al.) — AG News accuracy' (arXiv:1607.01759) from its own repository. Pre-registered claim: ag_news_accuracy = 92.5 (±5%). Hub verdict: TIMEOUT.

cs.LG 1 claims attention 4.0 #pip #reproducibility #reproducibility-study
L1
attested
TextCNN (Kim 2014) PyTorch reimplementation (Doragd) — SST-2 accuracy

Automated re-run of the headline result of 'TextCNN (Kim 2014) PyTorch reimplementation (Doragd) — SST-2 accuracy' (arXiv:1408.5882) from its own repository. Pre-registered claim: sst2_accuracy = 85.99 (±5%). Hub verdict: TIMEOUT.

cs.LG 1 claims attention 4.0 #pip #reproducibility #reproducibility-study