Skip to content

Add Dummy's results on the eight largest Multiverse datasets - #32

Open
TonyBagnall wants to merge 1 commit into
mainfrom
dummy-large-datasets
Open

TonyBagnall wants to merge 1 commit into
mainfrom
dummy-large-datasets

Conversation

@TonyBagnall

Copy link
Copy Markdown
Contributor

DREAMERA, DREAMERV, InsectWingbeat, LenDB and the four S2Agri datasets, from the first-submission Dummy runs, now copied into D:/Results/Multiverse. They use the same configuration as the current Dummy (strategy prior, random_state 0), and the five whose data is held locally match the class labels exactly.

The two full S2Agri prediction files hold 12 million rows each, too large to load through tsml-eval here, so the metrics were computed from counts of each (true, predicted) pair with the one probability row every Dummy row shares, using tsml-eval's metric definitions as sample-weighted scikit-learn calls. The same computation reproduces all 125 existing Dummy results on all seven metrics to within 1e-15. Existing rows are unchanged.

No table changes, since none of the eight is complete for any other estimator. InsectWingbeat is added to DEFERRED_DATASETS: with a Dummy result it would otherwise be listed against every other estimator on the UEA page as not run outside Multiverse-core, which is wrong, as the paper's classifiers were run on it and did not finish.

Summary

  • What does this PR change?

Checklist

  • I have updated documentation if needed.
  • I have added tests or validation steps if needed.
  • If this PR submits results, it follows results/schema.md.

DREAMERA, DREAMERV, InsectWingbeat, LenDB and the four S2Agri datasets, from
the first-submission Dummy runs, now copied into D:/Results/Multiverse. They
use the same configuration as the current Dummy (strategy prior, random_state
0), and the five whose data is held locally match the class labels exactly.

The two full S2Agri prediction files hold 12 million rows each, too large to
load through tsml-eval here, so the metrics were computed from counts of each
(true, predicted) pair with the one probability row every Dummy row shares,
using tsml-eval's metric definitions as sample-weighted scikit-learn calls. The
same computation reproduces all 125 existing Dummy results on all seven metrics
to within 1e-15. Existing rows are unchanged.

No table changes, since none of the eight is complete for any other estimator.
InsectWingbeat is added to DEFERRED_DATASETS: with a Dummy result it would
otherwise be listed against every other estimator on the UEA page as not run
outside Multiverse-core, which is wrong, as the paper's classifiers were run on
it and did not finish.

Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>

This branch has not been deployed

No deployments
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant