Skip to main content

OphthalVQA Dataset

Ophthalmic visual question-answering benchmark dataset released as supplementary data for evaluating multimodal language models in ophthalmology.

At a glance

FieldValue
Short nameophthalvqa
Full nameOphthalVQA Dataset
Primary categorymultimodal
Contained modalitiesfundus, fundus_angiography, oct, ocular_ultrasound, external_eye, text
Tasksvisual_question_answering, classification
SamplesNot reported
ClassesNot reported (Not reported)
Splitsall
Size0.03 GB
Source-stated termsCC BY 4.0
Normalized termscc-by
Descriptive screening labelStandard label without an explicit NC clause; not a permission finding
Terms scopedataset_files
Access frictionanonymous_direct
Route backendFigshare
Availabilityavailable (checked 2026-07-21)
Acquisition supportstandard_platform_supported
Legacy sample-loader statusMetadata and access only

Notes

Supplementary benchmark for ophthalmic VQA and multimodal model evaluation; inspect image provenance and splits before leakage-sensitive validation.

Access preflight and acquisition

# Read-only preflight
eyehub download ophthalvqa --data-dir ./data --dry-run --json

# Explicit transfer, only when preflight reports supported behavior
eyehub download ophthalvqa --data-dir ./data

Upstream page: https://doi.org/10.6084/m9.figshare.25624917

Source-term evidence: https://doi.org/10.6084/m9.figshare.25624917

Loader status

This catalog record provides metadata and access instructions, but it does not yet include a standard DatasetSample loader. Inspect the source file structure or contribute a loader before using it in a training pipeline.

Citation

@misc{ophthalvqa,
title = { OphthalVQA Dataset },
note = { Xu P, Chen X, Zhao Z, Shi D. OphthalVQA Dataset. Figshare, 2025. doi:10.6084/m9.figshare.25624917 },
year = { 2025 },
url = { https://doi.org/10.6084/m9.figshare.25624917 },
}

Source-stated terms

  • Raw source string: CC BY 4.0
  • Normalized category: cc-by
  • Apparent scope: dataset_files
  • Descriptive screening label: Standard label without an explicit NC clause; not a permission finding

⚠️ Source-stated terms, scope, and normalized labels are curation metadata, not legal advice or a permission finding. Review the current official source before transfer or reuse.

  • eyecare_100k: Eyecare-100K: Multimodal Ophthalmology VQA Corpus (102,000 records, unknown)
  • x_pcr: X-PCR Ophthalmology Progressive Clinical Reasoning Benchmark (18,700 records, unknown)
  • lmod_plus: LMOD+ Multimodal Ophthalmology Benchmark (32,633 records, unknown)
  • mm_retinal_reason: MM-Retinal-Reason: Ophthalmology Multimodal Reasoning Dataset (count not reported records, unknown)
  • angioreport: AngioReport Fundus Angiography Report Dataset (55,361 records, unknown)
  • ffa_ir: FFA-IR Medical Report Dataset (47,247 records, unknown)
  • multieye: MultiEYE: OCT-Enhanced Fundus Multi-Disease Benchmark (58,036 records, mit)
  • harvard_fairvision: Harvard-FairVision (AMD + DR + Glaucoma, paired SLO + OCT) (30,000 records, cc-by-nc-nd)