Datasheet Summary
Reads from three samples aligned to the reference with bwa-mem, sorted and indexed with samtools, and jointly called with bcftools; a Snakemake run captured as an EVI RO-Crate.
Also 1 schemas
Formats bam fastq binary yaml fasta amb ann bwt ns-proxy-autoconfig sa +3 more
SubstantivePartialHuman reviewNo points
16 of 48 points over the 24 criteria with a mechanical estimate; 4 await human review and are not counted either way.
Release Overview
Human Subjects & Regulatory
AI-Ready Review
Rubric for Review of AI-readiness Evaluation Criteria, v1.8 (2026-09-10). Estimates are a mechanical application of the scoring rules to the extracted evidence; criteria calling for human judgment are left unscored rather than assumed. Generated without network checks. The review page lists every piece of evidence.
0. FAIRness · gating 4/6 pts
- PID present (scheme: ARK)
- PID present but publisher is not a recognized sustainable repository
7 evidence items collected.
4 evidence items collected.
- metadata is JSON-LD but references no standard vocabulary
- 1 machine-readable schema entities
8 evidence items collected.
- machine-readable license linked in the metadata
- no AI/ML prohibition language found in license or use terms
5 evidence items collected.
1. Provenance · gating 5/8 pts
- 12 datasets carry provenance links
- ground-truth elements missing: samples, instruments, experiments
8 evidence items collected.
- 13 machine-readable transformation steps
- every computation links its software
- completeness of the provenance record (or disclosure of known gaps) is asserted, not verified
- final score is capped at 1.a's score (rule 1.b ≤ 1.a)
6 evidence items collected.
- 1 on mutable code hosting only
- 7 with no link
9 evidence items collected.
- 0 of 1 authors carry a PID
- 1 named in free text only
8 evidence items collected.
2. Characterization 2/8 pts
- abstract and keywords present, no controlled-vocabulary terms
4 evidence items collected.
5 evidence items collected.
- 1 schema entities but no standard-vocabulary binding found
5 evidence items collected.
- no bias, missingness, limitations, or completeness description anywhere in the metadata
9 evidence items collected.
- no QC description and no quality-control language anywhere in the metadata
4 evidence items collected.
3. Pre-model Explainability 1/6 pts
- human-readable datasheet linked
- 1 of 7 machine-readable sections populated
4 evidence items collected.
- no use-case guidance in the metadata
6 evidence items collected.
- no checksums found
3 evidence items collected.
4. Ethics · gating 0/6 pts
- no acquisition, consent, or ethics-review description anywhere in the metadata
10 evidence items collected.
- no sensitivity classification, sensitivity statement, or management description
7 evidence items collected.
6 evidence items collected.
- no security-level metadata
4 evidence items collected.
5. Sustainability 3/8 pts
- PID present (scheme: ARK)
- 19 datasets have a contentUrl but no recognized archive detected
5 evidence items collected.
- no recognized repository host detected
7 evidence items collected.
- no DMP / governance plan in the metadata
7 evidence items collected.
- components associated machine-readably in the archived RO-Crate (hasPart + provenance links)
- accessibility of every component not verified — downgrade if pieces are missing
4 evidence items collected.
6. Computability 1/6 pts
- formal schema/standard declared, so structural validation is possible
- no populated standard-vocabulary bindings — semantic conformance cannot be deterministically validated (the 1-rule)
7 evidence items collected.
- no dataset has a remote distribution link — no programmatic access mechanism visible
5 evidence items collected.
5 evidence items collected.
- no splits, withheld-information statement, or example data anywhere in the metadata
7 evidence items collected.