Skip to main content

SOUL: OCTA Human-Machine Collaborative Annotation Dataset

Longitudinal OCT angiography projection maps with vessel labels, clinical text, and treatment/follow-up groupings.

At a glance

FieldValue
Short namesoul_octa
Full nameSOUL: OCTA Human-Machine Collaborative Annotation Dataset
First published2023-12-22
Publication date precisionday
Publication date evidenceapi.figshare.com/v2
Publication date source fieldpublished_date (Figshare version 1)
Publication date reviewed2026-09-11
Primary categoryocta
Resource rolecurrent_dataset
Dataset familysoul_octa
Contained modalitiesocta, text, tabular
Taskssegmentation, classification, progression_analysis
Primary reported quantity178 longitudinal samples
ClassesNot reported (Not reported)
Splitsall
Size0.113 GB
Source-stated termsCC BY 4.0
Normalized termscc-by
Descriptive screening labelStandard label without an explicit NC clause; not a permission finding
Terms scopedataset_files
Access frictionself_service_authenticated
Route backendFigshare
Availabilityavailable (checked 2026-07-21)
Acquisition supportend_to_end_tested
Legacy sample-loader statusMetadata and access only

Reported quantities

RoleCountUnitScopeBasisEvidence
Primary178longitudinal_samplesSix source-reported longitudinal subsets The current archive contains 2,046 JPG assets including raw images and corresponding labels; these are not 2,046 independent samples.official_source_descriptionhttps://doi.org/10.6084/m9.figshare.24893358.v3

Counts retain their source-reported units. Additional rows can describe components, paired items, or derivative copies and are not automatically added to the primary quantity.

Notes

The six source-reported longitudinal subsets total 178 samples. The release includes raw images, vessel annotations, and clinical text.

Access information and download

# Read-only preflight
eyehub download soul_octa --data-dir ./data --dry-run --json

# Download, only when preflight reports supported behavior
eyehub download soul_octa --data-dir ./data

Upstream page: https://doi.org/10.6084/m9.figshare.24893358.v3

Source-term evidence: https://doi.org/10.6084/m9.figshare.24893358.v3

Loader status

This catalog record provides metadata and access instructions, but it does not yet include a standard DatasetSample loader. Inspect the source file structure or contribute a loader before using it in a training pipeline.

Citation

@misc{soul_octa,
title = { SOUL: OCTA Human-Machine Collaborative Annotation Dataset },
note = { Xue J, Feng Z, Zeng L, et al. Soul: An OCTA dataset based on Human Machine Collaborative Annotation Framework. Scientific Data. 2024;11. doi:10.1038/s41597-024-03665-7 },
year = { 2024 },
url = { https://doi.org/10.6084/m9.figshare.24893358.v3 },
}

Source-stated terms

  • Raw source string: CC BY 4.0
  • Normalized category: cc-by
  • Apparent scope: dataset_files
  • Descriptive screening label: Standard label without an explicit NC clause; not a permission finding

⚠️ Source-stated terms, scope, and normalized labels are curation metadata, not legal advice or a permission finding. Review the current official source before transfer or reuse.

Similar resources by shared modality

  • ocular_chat_vqa: OcularChat-VQA: AREDS-Derived Patient-Physician Dialogue Dataset (844,000 records, cc-by-nc-sa)
  • lmod_plus: LMOD+ Multimodal Ophthalmology Benchmark (32,633 annotated instances, unknown)
  • fairvlmed: FairVLMed: Fair Vision-Language Medical Ophthalmic Dataset (10,000 images, cc-by-nc-nd)
  • ophtho_readability: Language and Readability Barriers in Ophthalmology Dataset (139 documents, cc-by)
  • dryad_preeclampsia_ocular_octa: Plane wave ultrasound and OCT angiography of the eye in preeclampsia (Not reported, cc0)
  • fundus_cc_2_5m: Fundus-CC-2.5M Text Corpus (2,500,000 text items, unknown)
  • ophora: Ophora-160K: Ophthalmic Surgical Video Instruction Dataset (162,185 video clip instruction pairs, unknown)
  • fundus_105k: Fundus-105K Text Dataset (105,000 text items, unknown)