Skip to main content

LMOD+ Multimodal Ophthalmology Benchmark

Composite multimodal ophthalmology benchmark with multi-granular anatomical, diagnostic, staging, demographic, and text annotations.

At a glance

FieldValue
Short namelmod_plus
Full nameLMOD+ Multimodal Ophthalmology Benchmark
First publishedUnknown
Publication date precisionUnknown
Publication date evidenceUnknown
Publication date source fieldUnknown
Publication date reviewedUnknown
Primary categorymultimodal
Resource rolederivative_dataset
Dataset familylmod_plus
Contained modalitiesfundus, oct, external_eye, surgical_video, text, tabular
Tasksvisual_question_answering, classification, grading, detection
Primary reported quantity32,633 annotated instances
ClassesNot reported (Not reported)
Splitsall
SizeNot reported
Source-stated termsMixed upstream licenses; verify each component
Normalized termsunknown
Descriptive screening labelUnknown or unclear; do not assume permission
Terms scopemixed_components
Access frictionanonymous_direct
Route backendManual (upstream-gated)
Availabilityavailable (checked 2026-07-21)
Acquisition supportguided_instructions_only
Legacy sample-loader statusMetadata and access only

Reported quantities

RoleCountUnitScopeBasisEvidence
Primary32,633annotated_instancesComposite benchmark instancesofficial_source_descriptionkfzyqin.github.io/lmod_plus

Counts retain their source-reported units. Additional rows can describe components, paired items, or derivative copies and are not automatically added to the primary quantity.

Notes

Integrates 32,633 instances from nine public datasets across five modalities and 12 conditions. It adds a distinct benchmark and annotation surface but does not replace the source datasets; license and redistribution terms remain component-specific.

Documented relationships

These links record source-supported lineage or overlap, not merely similar modality tags.

  • This record is derived from cataract1k: The LMOD+ project page lists nine component datasets, including the five cataloged targets represented by these edges. (evidence)
  • This record is derived from g1020: The LMOD+ project page lists nine component datasets, including the five cataloged targets represented by these edges. (evidence)
  • This record is derived from idrid: The LMOD+ project page lists nine component datasets, including the five cataloged targets represented by these edges. (evidence)
  • This record is derived from oimhs: The LMOD+ project page lists nine component datasets, including the five cataloged targets represented by these edges. (evidence)
  • This record is derived from refuge2: The LMOD+ project page lists nine component datasets, including the five cataloged targets represented by these edges. (evidence)

Access information and download

# Read-only preflight
eyehub download lmod_plus --data-dir ./data --dry-run --json

# Download, only when preflight reports supported behavior
eyehub download lmod_plus --data-dir ./data

Upstream page: kfzyqin.github.io/lmod_plus

Source-term evidence: kfzyqin.github.io/lmod_plus

Loader status

This catalog record provides metadata and access instructions, but it does not yet include a standard DatasetSample loader. Inspect the source file structure or contribute a loader before using it in a training pipeline.

Citation

@misc{lmod_plus,
title = { LMOD+ Multimodal Ophthalmology Benchmark },
note = { Qin Z, Liu Y, Yin Y, et al. LMOD+: A Comprehensive Multimodal Dataset and Benchmark for Developing and Evaluating Multimodal Large Language Models in Ophthalmology. ACM Transactions on Computing for Healthcare. 2026. doi:10.1145/3801746 },
year = { 2026 },
url = { https://kfzyqin.github.io/lmod_plus/ },
}

Source-stated terms

  • Raw source string: Mixed upstream licenses; verify each component
  • Normalized category: unknown
  • Apparent scope: mixed_components
  • Descriptive screening label: Unknown or unclear; do not assume permission

⚠️ Source-stated terms, scope, and normalized labels are curation metadata, not legal advice or a permission finding. Review the current official source before transfer or reuse.

Similar resources by shared modality

  • eyecare_100k: Eyecare-100K: Multimodal Ophthalmology VQA Corpus (102,000 question answer pairs, unknown)
  • x_pcr: X-PCR Ophthalmology Progressive Clinical Reasoning Benchmark (18,735 rows, unknown)
  • ophthalvqa: OphthalVQA Dataset (600 question answer pairs, cc-by)
  • fairvlmed: FairVLMed: Fair Vision-Language Medical Ophthalmic Dataset (10,000 images, cc-by-nc-nd)
  • olives: OLIVES: Ophthalmic Labels for Investigating Visual Eye Semantics (9,408 b scans, cc-by)
  • grape: GRAPE: Glaucoma Real-world Appraisal Progression Ensemble (1,115 examinations, cc0)
  • mm_retinal_reason: MM-Retinal-Reason: Ophthalmology Multimodal Reasoning Dataset (130 question answer pairs, unknown)
  • ocular_chat_vqa: OcularChat-VQA: AREDS-Derived Patient-Physician Dialogue Dataset (844,000 records, cc-by-nc-sa)