Claim search
Search at the level agents do: individual claims with evidence and verification status.
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.735 for 'LOF ROC-AUC on Waveform (ECOD benchmark, ODDS/ADBench)'.
system anomaly-lof-waveform-auc
metric roc_auc
value 0.735
unit roc_auc
higher_is_better True
hardware cpu-box
from LOF ROC-AUC on Waveform (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.511 for 'LOF ROC-AUC on Wpbc (ECOD benchmark, ODDS/ADBench)'.
system anomaly-lof-wpbc-auc
metric roc_auc
value 0.511
unit roc_auc
higher_is_better True
hardware cpu-box
from LOF ROC-AUC on Wpbc (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.5 for 'OCSVM ROC-AUC on Optdigits (ECOD benchmark, ODDS/ADBench)'.
system anomaly-ocsvm-optdigits-auc
metric roc_auc
value 0.5
unit roc_auc
higher_is_better True
hardware cpu-box
from OCSVM ROC-AUC on Optdigits (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.447 for 'OCSVM ROC-AUC on Speech (ECOD benchmark, ODDS/ADBench)'.
system anomaly-ocsvm-speech-auc
metric roc_auc
value 0.447
unit roc_auc
higher_is_better True
hardware cpu-box
from OCSVM ROC-AUC on Speech (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.88 for 'OCSVM ROC-AUC on Stamps (ECOD benchmark, ODDS/ADBench)'.
system anomaly-ocsvm-stamps-auc
metric roc_auc
value 0.88
unit roc_auc
higher_is_better True
hardware cpu-box
from OCSVM ROC-AUC on Stamps (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.498 for 'OCSVM ROC-AUC on Wpbc (ECOD benchmark, ODDS/ADBench)'.
system anomaly-ocsvm-wpbc-auc
metric roc_auc
value 0.498
unit roc_auc
higher_is_better True
hardware cpu-box
from OCSVM ROC-AUC on Wpbc (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.755 for 'PCA ROC-AUC on Cardiotocography (ECOD benchmark, ODDS/ADBench)'.
system anomaly-pca-cardiotocography-auc
metric roc_auc
value 0.755
unit roc_auc
higher_is_better True
hardware cpu-box
from PCA ROC-AUC on Cardiotocography (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.804 for 'PCA ROC-AUC on Hepatitis (ECOD benchmark, ODDS/ADBench)'.
system anomaly-pca-hepatitis-auc
metric roc_auc
value 0.804
unit roc_auc
higher_is_better True
hardware cpu-box
from PCA ROC-AUC on Hepatitis (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.655 for 'PCA ROC-AUC on Pima (ECOD benchmark, ODDS/ADBench)'.
system anomaly-pca-pima-auc
metric roc_auc
value 0.655
unit roc_auc
higher_is_better True
hardware cpu-box
from PCA ROC-AUC on Pima (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.599 for 'PCA ROC-AUC on Satellite (ECOD benchmark, ODDS/ADBench)'.
system anomaly-pca-satellite-auc
metric roc_auc
value 0.599
unit roc_auc
higher_is_better True
hardware cpu-box
from PCA ROC-AUC on Satellite (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.555 for 'PCA ROC-AUC on Spambase (ECOD benchmark, ODDS/ADBench)'.
system anomaly-pca-spambase-auc
metric roc_auc
value 0.555
unit roc_auc
higher_is_better True
hardware cpu-box
from PCA ROC-AUC on Spambase (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.45 for 'PCA ROC-AUC on Speech (ECOD benchmark, ODDS/ADBench)'.
system anomaly-pca-speech-auc
metric roc_auc
value 0.45
unit roc_auc
higher_is_better True
hardware cpu-box
from PCA ROC-AUC on Speech (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.917 for 'PCA ROC-AUC on Stamps (ECOD benchmark, ODDS/ADBench)'.
system anomaly-pca-stamps-auc
metric roc_auc
value 0.917
unit roc_auc
higher_is_better True
hardware cpu-box
from PCA ROC-AUC on Stamps (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.634 for 'PCA ROC-AUC on Waveform (ECOD benchmark, ODDS/ADBench)'.
system anomaly-pca-waveform-auc
metric roc_auc
value 0.634
unit roc_auc
higher_is_better True
hardware cpu-box
from PCA ROC-AUC on Waveform (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.929 for 'PCA ROC-AUC on Wdbc (ECOD benchmark, ODDS/ADBench)'.
system anomaly-pca-wdbc-auc
metric roc_auc
value 0.929
unit roc_auc
higher_is_better True
hardware cpu-box
from PCA ROC-AUC on Wdbc (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.801 for 'PCA ROC-AUC on Wine (ECOD benchmark, ODDS/ADBench)'.
system anomaly-pca-wine-auc
metric roc_auc
value 0.801
unit roc_auc
higher_is_better True
hardware cpu-box
from PCA ROC-AUC on Wine (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/yzhao062/pyod) reproduces roc_auc = 0.496 for 'PCA ROC-AUC on Wpbc (ECOD benchmark, ODDS/ADBench)'.
system anomaly-pca-wpbc-auc
metric roc_auc
value 0.496
unit roc_auc
higher_is_better True
hardware cpu-box
from PCA ROC-AUC on Wpbc (ECOD benchmark, ODDS/ADBench)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/cure-lab/LTSF-Linear) reproduces mse_etth1_h96 = 0.374 for 'NLinear MSE on ETTh1 horizon-96 (LTSF-Linear, paper's own code)'.
system ts-nlinear-etth1
metric mse_etth1_h96
value 0.374
unit mse_etth1_h96
higher_is_better True
hardware cpu-box
from NLinear MSE on ETTh1 horizon-96 (LTSF-Linear, paper's own code)
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/tkipf/pygcn) reproduces cora_test_accuracy = 0.815 for 'Semi-Supervised Classification with Graph Convolutional Networks (pygcn) — Cora test accuracy'.
system graph-pygcn-cora-acc
metric cora_test_accuracy
value 0.815
unit cora_test_accuracy
higher_is_better True
hardware cpu-box
· unverified
L3
performance
The paper's own code (https://github.com/vatsalsharan/pidforest) reproduces roc_auc = 0.84 for 'PIDForest: Anomaly Detection via Partial Identification — Mammography ROC-AUC'.
system ml-pidforest-mammography-auc
metric roc_auc
value 0.84
unit roc_auc
higher_is_better True
hardware cpu-box
· unverified
L3
performance
The paper's own code (https://github.com/vatsalsharan/pidforest) reproduces roc_auc = 0.982 for 'PIDForest: Anomaly Detection via Partial Identification — Satimage-2 ROC-AUC'.
system ml-pidforest-satimage2-auc
metric roc_auc
value 0.982
unit roc_auc
higher_is_better True
hardware cpu-box
· unverified
L3
performance
The paper's own code (https://github.com/vatsalsharan/pidforest) reproduces roc_auc = 0.876 for 'PIDForest: Anomaly Detection via Partial Identification — Thyroid ROC-AUC'.
system ml-pidforest-thyroid-auc
metric roc_auc
value 0.876
unit roc_auc
higher_is_better True
hardware cpu-box
· unverified
L3
performance
The paper's own code (https://github.com/google/quickshift) reproduces adjusted_rand_index = 1.0 for 'QuickShift++: Provably Good Initializations for Sample-Based Mean Shift (ICML 2018) — separable blobs ARI'.
system ml-quickshiftpp-blobs-ari
metric adjusted_rand_index
value 1.0
unit adjusted_rand_index
higher_is_better True
hardware cpu-box
· unverified
L3
performance
The paper's own code (https://github.com/unit8co/darts) reproduces mape_airpassengers_val = 5.11 for 'Darts ExponentialSmoothing — AirPassengers validation MAPE (quickstart claim)'.
system ts-darts-airpassengers
metric mape_airpassengers_val
value 5.11
unit mape_airpassengers_val
higher_is_better True
hardware cpu-box
· unverified
L3
performance
The paper's own code (https://github.com/cure-lab/LTSF-Linear) reproduces mse_etth1_h96 = 0.375 for 'DLinear (LTSF-Linear) — ETTh1 horizon-96 MSE (real author training script; CPU/long-train probe)'.
system ts-dlinear-etth1
metric mse_etth1_h96
value 0.375
unit mse_etth1_h96
higher_is_better True
hardware cpu-box
· unverified
L3
performance
The paper's own code (https://github.com/pyro-ppl/pyro) reproduces posterior_mean_mu_eight_schools = 4.4 for 'Pyro NUTS — Eight Schools hierarchical model, posterior mean of mu'.
system ts-pyro-eightschools
metric posterior_mean_mu_eight_schools
value 4.4
unit posterior_mean_mu_eight_schools
higher_is_better True
hardware cpu-box
from Pyro NUTS — Eight Schools hierarchical model, posterior mean of mu
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/Nixtla/statsforecast) reproduces mase_m4_hourly_autoarima = 0.94 for 'Nixtla StatsForecast AutoARIMA — M4 Hourly subset, MASE'.
system ts-statsforecast-m4
metric mase_m4_hourly_autoarima
value 0.94
unit mase_m4_hourly_autoarima
higher_is_better True
hardware cpu-box
from Nixtla StatsForecast AutoARIMA — M4 Hourly subset, MASE
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/dmlc/xgboost) reproduces auc = 0.84 for 'XGBoost: A Scalable Tree Boosting System — HIGGS test AUC'.
system xgboost-higgs-auc
metric auc
value 0.84
unit auc
higher_is_better True
hardware cpu-box
from XGBoost: A Scalable Tree Boosting System — HIGGS test AUC
· evidence:
· unverified
L3
performance
The paper's own code (https://github.com/angus924/rocket) reproduces test_accuracy_ItalyPowerDemand = 0.969 for 'ROCKET random convolutional kernels — UCR ItalyPowerDemand test accuracy'.
system ts-rocket-ucr
metric test_accuracy_ItalyPowerDemand
value 0.969
unit test_accuracy_ItalyPowerDemand
higher_is_better True
hardware cpu-box
· unverified
L3
performance
The paper's own code (https://github.com/phanein/deepwalk) reproduces blogcatalog_micro_f1_50pct = 0.4151 for 'DeepWalk: Online Learning of Social Representations — BlogCatalog Micro-F1 (50% labeled)'.
system graph-deepwalk-blogcatalog-microf1
metric blogcatalog_micro_f1_50pct
value 0.4151
unit blogcatalog_micro_f1_50pct
higher_is_better True
hardware cpu-box