Quickstart
Everything a participant runs is Python 3.12+ standard library only. Download the data, validate your files, and (for Task 1) score them locally.
Everything a participant runs is Python 3.12+ standard library only. Clone the distribution and run from its root.
1. Get the data
The 2026 datasets are published on the Hugging Face Hub under the OAEI-ML organisation — browse at huggingface.co/datasets/OAEI-ML/diso-oaei (the v2026 tag pins this edition):
| Download | Contents |
|---|---|
archives/ontologies.zip | the 10 task ontologies (also unpacked under ontologies/ in the repository) |
archives/repaired_silver_refs.zip | Task 1 headline references (repository: tasks/global/references/_for_use/; per-pair component files live under tasks/global/references/<pair>/) |
archives/unrepaired_silver_refs.zip | Task 1 secondary references (repository: tasks/global/references/_unrepaired/) |
pools/uco-stix/pools.jsonl · pools/stix-d3fend/pools.jsonl | Task 2 candidate pools (repository: tasks/ranking/candidates/<pair>/pools.jsonl) |
Verify your downloads against the SHA-256 checksums in the downloads table. The repository ships the same files unpacked, so cloning it also works.
Command-line download
pip install -U huggingface_hub
hf download OAEI-ML/diso-oaei --repo-type dataset --local-dir ./diso-oaei # everything
hf download OAEI-ML/diso-oaei ontologies/stix.owl --repo-type dataset # a single file
hf download OAEI-ML/diso-oaei --repo-type dataset --revision v2026 --local-dir ./diso-oaei-2026 # pin the 2026 edition
2. Task 1 — Global alignment
For each of the 6 pairs, emit one OAEI Alignment RDF (default xmlns = the alignment namespace without a trailing #; one <Cell> per = correspondence). The full spec and template are provided under tasks/global/submission-format.md.
Validate, then self-score — the references are public, with headline metrics computed using (with as secondary):
# structural check (zero-dependency); optional RelaxNG check needs libxml2-utils
python3 scripts/validate_global.py my-thinkhome-brick.rdf
xmllint --relaxng scripts/alignment.rng my-thinkhome-brick.rdf
# score one submission under the dual reference
python3 scripts/score_global.py my-submission.rdf \
--rplus tasks/global/references/_for_use/thinkhome-brick.silver.rdf \
--rapprox tasks/global/references/_unrepaired/thinkhome-brick.silver.unrepaired.rdf
Sanity check: a reference scored against itself gives P=R=F1=1. A MELT local-track driver is under construction (see the README).
3. Task 2 — Local equivalence ranking
For each pair, read tasks/ranking/candidates/<pair>/pools.jsonl (one JSON object per query, each with 50 candidates including the NIL IRI). Emit a JSONL submission, one line per qid, ranking that query’s candidates best-first (a permutation of the 50; rank NIL first to abstain):
{"qid": 0, "ranking": ["<best-IRI>", "...", "https://oaei.ontologymatching.org/2026/diso/NIL", "..."]}
The task is considered unsupervised; the answers (ground truths) are private, so there is no local scorer. Validate the format, then submit. We score Hits@, MRR, and macro-average over both pairs.
python3 scripts/validate_ranking.py tasks/ranking/candidates/uco-stix/pools.jsonl my_uco-stix.jsonl
python3 scripts/validate_ranking.py tasks/ranking/candidates/stix-d3fend/pools.jsonl my_stix-d3fend.jsonl
Full spec + worked example: tasks/ranking/submission-format.md.
4. Submit
The evaluation window runs from 12 July to 1 September 2026, 00:00 Anywhere on Earth (AoE).
Submit via CodaBench: Task 1 — Global Alignment · Task 2 — Local Equivalence Ranking.
Register, then upload one zip per task as described on each competition’s Overview page.
Results are automatically published as provisional to the leaderboard.
Organisers verify, reproduce where possible, mark participant results as accepted, and publish them alongside the organiser-run baselines.