Method
Medical AI Dataset & Ground Truth Assessor
Is the dataset credible enough to support clinical-AI validation?
Field scoring contract
| Field | Type | Role | Direction | Scored |
|---|---|---|---|---|
| Dataset/cohort | text | context | context | no |
| Data provenance | select | evidence_signal | higher_is_better | yes |
| Reference standard / ground truth | select | evidence_signal | higher_is_better | yes |
| Adjudication and disagreement process | select | evidence_signal | higher_is_better | yes |
| Population/site representativeness | select | evidence_signal | higher_is_better | yes |
| Leakage and duplicate-patient controls | select | evidence_signal | higher_is_better | yes |
| Subgroup coverage | select | evidence_signal | higher_is_better | yes |
Limits
- Preliminary output
- Human review required
- Not certification
- Context text is not averaged into numeric scores.
- Outputs require source/evidence review before decisions.