Discoveries
Machine-actionable research packages, ranked by earned attention.
Filter by topicclear ✕show ▾
Automated re-run of the headline result of 'OCSVM ROC-AUC on Satimage-2 (ECOD benchmark, ODDS/ADBench)' (arXiv:2201.00382) from its own repository. Pre-registered claim: roc_auc = 0.998 (±8%). Hub verdict: DIVERGED.
Automated re-run of the headline result of 'EdMot: Edge Enhancement for Motif-aware Community Detection — Cora modularity' (arXiv:1906.04560) from its own repository. Pre-registered claim: cora_modularity = 0.4088 (±5%). Hub verdict: BUILD_FAILED.
Automated re-run of the headline result of 'BEIR: BM25 (Anserini) nDCG@10 on SciFact' (arXiv:2104.08663) from its own repository. Pre-registered claim: ndcg@10 = 0.65 (±8%). Hub verdict: BUILD_FAILED.
Automated re-run of the headline result of 'Neural Collaborative Filtering (NeuMF) — MovieLens-1M HR@10' (arXiv:1708.05031) from its own repository. Pre-registered claim: ml1m_hr_at_10 = 0.73 (±5%). Hub verdict: RUN_FAILED.
Automated re-run of the headline result of 'OpenNE — node2vec on Wiki node classification Micro-F1' (arXiv:1607.00653) from its own repository. Pre-registered claim: wiki_micro_f1 = 0.651 (±10%). Hub verdict: BUILD_FAILED.
Automated re-run of the headline result of 'BERT-base SST-2 fine-tune accuracy (transformers run_glue.py)' (arXiv:1810.04805) from its own repository. Pre-registered claim: accuracy = 0.93 (±5%). Hub verdict: BUILD_FAILED.
Automated re-run of the headline result of 'SUOD: Accelerating Large-Scale Unsupervised Heterogeneous Outlier Detection — Cardio IForest ROC-AUC' (arXiv:2003.05731) from its own repository. Pre-registered claim: roc_auc = 0.9216 (±8%). Hub verdict: RUN_FAILED.
Automated re-run of the headline result of 'TextCNN multi-label text classification (brightmart/text_classification) — TextCNN accuracy' (arXiv:1408.5882) from its own repository. Pre-registered claim: textcnn_accuracy = 0.65 (±8%). Hub verdict: BUILD_FAILED.
Automated re-run of the headline result of 'Distributed Representations of Sentences and Documents (Doc2Vec/Paragraph Vector, Le & Mikolov 2014) — IMDB sentiment accuracy, gensim reproduction' (arXiv:1405.4053) from its own repository. Pre-registered claim: imdb_accuracy = 0.87 (±8%). Hub verdict: RUN_FAILED.
Automated re-run of the headline result of 'Bag of Tricks for Efficient Text Classification (fastText, Joulin et al. EACL 2017) — official repo DBpedia P@1' (arXiv:1607.01759) from its own repository. Pre-registered claim: dbpedia_p_at_1 = 0.98 (±5%). Hub verdict: RUN_FAILED.
Automated re-run of the headline result of 'Bag of Tricks for Efficient Text Classification (fastText, Joulin et al.) — AG News accuracy' (arXiv:1607.01759) from its own repository. Pre-registered claim: ag_news_accuracy = 92.5 (±5%). Hub verdict: TIMEOUT.
Automated re-run of the headline result of 'NB-SVM (Wang & Manning, ACL 2012) — IMDB sentiment accuracy, bigram reproduction' (arXiv:1412.5335) from its own repository. Pre-registered claim: imdb_accuracy_bigram = 91.55 (±5%). Hub verdict: RUN_FAILED.
Automated re-run of the headline result of 'A Simple but Tough-to-Beat Baseline for Sentence Embeddings (SIF, Arora et al., ICLR 2017) — STS Pearson correlation' (arXiv:1611.01462) from its own repository. Pre-registered claim: sts_pearson = 0.717 (±10%). Hub verdict: RUN_FAILED.
Automated re-run of the headline result of 'TextCNN (Kim 2014) MindSpore implementation — SST2 accuracy' (arXiv:1408.5882) from its own repository. Pre-registered claim: sst2_accuracy = 0.7971 (±5%). Hub verdict: BUILD_FAILED.
Automated re-run of the headline result of 'TextCNN (Kim 2014) PyTorch reimplementation (Doragd) — SST-2 accuracy' (arXiv:1408.5882) from its own repository. Pre-registered claim: sst2_accuracy = 85.99 (±5%). Hub verdict: TIMEOUT.
Automated re-run of the headline result of 'pomegranate: HMM Baum-Welch runtime (1000x10-dim, 5 iters) — version-skew case' (arXiv:1711.00137) from its own repository. Pre-registered claim: hmm_baumwelch_runtime_seconds = 13 (±100%). Hub verdict: BUILD_FAILED.
Automated re-run of the headline result of 'Pyserini: BM25 nDCG@10 on BEIR SciFact (prebuilt Lucene index)' (arXiv:2102.10073) from its own repository. Pre-registered claim: ndcg@10 = 0.679 (±5%). Hub verdict: BUILD_FAILED.
Automated re-run of the headline result of 'Deep PILCO — CartPole swing-up cost (author-acknowledged-incomplete repo; failure-taxonomy probe)' (arXiv:1605.07127) from its own repository. Pre-registered claim: cartpole_swingup_cost = 0.1 (±50%). Hub verdict: BUILD_FAILED.
Automated re-run of the headline result of 'ViT (lucidrains) CIFAR-10 from-scratch accuracy' (arXiv:2010.11929) from its own repository. Pre-registered claim: accuracy = 0.88 (±10%). Hub verdict: BUILD_FAILED.