Skip to main content

Language and Readability Barriers in Ophthalmology Dataset

Text/tabular dataset supporting readability and language-access analyses in ophthalmology.

At a glance

FieldValue
Short nameophtho_readability
Full nameLanguage and Readability Barriers in Ophthalmology Dataset
Primary categorytext
Contained modalitiestext, tabular
Taskstext_generation, classification
SamplesNot reported
ClassesNot reported (Not reported)
Splitsall
Size0.01 GB
Source-stated termsCC BY 4.0
Normalized termscc-by
Descriptive screening labelStandard label without an explicit NC clause; not a permission finding
Terms scopedataset_files
Access frictionanonymous_direct
Route backendZenodo
Availabilityavailable (checked 2026-07-21)
Acquisition supportstandard_platform_supported
Legacy sample-loader statusMetadata and access only

Access preflight and acquisition

# Read-only preflight
eyehub download ophtho_readability --data-dir ./data --dry-run --json

# Explicit transfer, only when preflight reports supported behavior
eyehub download ophtho_readability --data-dir ./data

Upstream page: zenodo.org/records

Source-term evidence: zenodo.org/records

Loader status

This catalog record provides metadata and access instructions, but it does not yet include a standard DatasetSample loader. Inspect the source file structure or contribute a loader before using it in a training pipeline.

Citation

@misc{ophtho_readability,
title = { Language and Readability Barriers in Ophthalmology Dataset },
note = { Language and readability barriers in ophthalmology dataset. Zenodo, 2025. doi:10.5281/zenodo.16592100 },
year = { 2025 },
url = { https://zenodo.org/records/16592100 },
}

Source-stated terms

  • Raw source string: CC BY 4.0
  • Normalized category: cc-by
  • Apparent scope: dataset_files
  • Descriptive screening label: Standard label without an explicit NC clause; not a permission finding

⚠️ Source-stated terms, scope, and normalized labels are curation metadata, not legal advice or a permission finding. Review the current official source before transfer or reuse.

  • ocular_chat_vqa: OcularChat-VQA: AREDS-Derived Patient-Physician Dialogue Dataset (844,000 records, cc-by-nc-sa)
  • lmod_plus: LMOD+ Multimodal Ophthalmology Benchmark (32,633 records, unknown)
  • soul_octa: SOUL: OCTA Human-Machine Collaborative Annotation Dataset (178 records, cc-by)
  • fairvlmed: FairVLMed: Fair Vision-Language Medical Ophthalmic Dataset (count not reported records, unknown)
  • fundus_cc_2_5m: Fundus-CC-2.5M Text Corpus (2,500,000 records, unknown)
  • ophora: Ophora-160K: Ophthalmic Surgical Video Instruction Dataset (160,185 records, unknown)
  • fundus_105k: Fundus-105K Text Dataset (105,000 records, unknown)
  • eyecare_100k: Eyecare-100K: Multimodal Ophthalmology VQA Corpus (102,000 records, unknown)