Evidence map›Paper›PMID 42649801›Full record

ArticleBioengineering (Basel, Switzerland)2026

A Rationale-Conditioned Image-Contrast Audit of OCT Dependence in Vision-Language Models for Anti-VEGF Treatment-Response Prediction.

Wonbong Jang, Gwon Yul Jo, Siyun Lee, Shin Jeong Yoon, Gi Young Lee, Hee-Eun Lee, Tae Hyung Kim, Jong Won Baek, Joonhyung Kim

Abstract read
In one paragraph

Article in Bioengineering (Basel, Switzerland), 2026. The graph could read no effect estimate from its abstract, so it casts no vote on the map. Not yet cited in PubMed.

0numbers the graph read from it
0cells of the map it votes in
0citing papers in PubMed
–field-weighted citation impact
1 · What the graph read from it

What it found

Each row is one number read from the abstract, on the scale the paper reported it, with its interval. Left of the dashed line favours the treatment, right favours the comparator. Under each row is the sentence it came from. New to these charts? A ten-minute tutorial.

The abstract states no effect estimate the extractor could read, or names no intervention and outcome on the map, so this paper lights no cell and moves no belief. It is still indexed, cited and linked below.

2 · The registry

The trial behind it

Trials whose registry record cites this paper, or whose number appears in the abstract. A trial that started after this paper was published is citing it as background, not reporting it.

Neither the registry nor the abstract names a trial number. If this is a trial report, that itself is worth knowing.

3 · Its place in the literature

Who cites it

0 citing papers in PubMed.

No citing paper in PubMed yet.

4 · The record

Corrections and comments

PubMed lists nothing against this paper. Absence here is not a guarantee, only a check that was made.

5 · Who and what money

Authors and funding

9 authors.

Wonbong JangClear Eye Clinic, 181-26 Jijedongsak 2-ro, Pyeongtaek 18029, Republic of Korea.
Gwon Yul JoBioNexus Inc., Gwanggyo Business Center, 156 Gwanggyo-ro, Yeongtong-gu, Suwon 16506, Republic of Korea.ORCID 0009-0001-1532-8194
Siyun LeeDepartment of Ophthalmology, CHA Bundang Medical Center, CHA University School of Medicine, 59 Yatap-ro, Bundang-gu, Seongnam 13496, Republic of Korea.ORCID 0009-0003-1089-9443
Shin Jeong YoonDepartment of Ophthalmology, CHA Bundang Medical Center, CHA University School of Medicine, 59 Yatap-ro, Bundang-gu, Seongnam 13496, Republic of Korea.
Gi Young LeeBioNexus Inc., Gwanggyo Business Center, 156 Gwanggyo-ro, Yeongtong-gu, Suwon 16506, Republic of Korea.ORCID 0009-0009-0431-5572
Hee-Eun LeeBioNexus Inc., Gwanggyo Business Center, 156 Gwanggyo-ro, Yeongtong-gu, Suwon 16506, Republic of Korea.ORCID 0000-0002-7485-9083
Tae Hyung KimBioNexus Inc., Gwanggyo Business Center, 156 Gwanggyo-ro, Yeongtong-gu, Suwon 16506, Republic of Korea.
Jong Won BaekBioNexus Inc., Gwanggyo Business Center, 156 Gwanggyo-ro, Yeongtong-gu, Suwon 16506, Republic of Korea.
Joonhyung KimDepartment of Ophthalmology, CHA Bundang Medical Center, CHA University School of Medicine, 59 Yatap-ro, Bundang-gu, Seongnam 13496, Republic of Korea.ORCID 0000-0002-3717-2223

Funding

No grant is acknowledged in the PubMed record.

6 · The paper itself

Abstract

Vision-language models (VLMs) are increasingly evaluated for retinal image interpretation, but end-task discrimination does not establish whether a prediction depends on the optical coherence tomography (OCT) image. We evaluated RetinaVLM, LLaVA-Med, and Qwen3.6-27B under zero-shot and parameter-efficient adaptation using an APTOS-2021 cohort (128 training, 21 validation, and 69 test eyes) and a 100-eye cross-site stress-test cohort. The primary readout compared forced-choice continue/stop scores obtained with a real OCT B-scan and a uniform-grey image while holding the generated rationale fixed; it therefore estimates a rationale-conditioned direct image contrast rather than total image dependence. Across 22 internal cells, 21 confidence intervals included an AUC of 0.5, while one Qwen zero-shot cell was inversely aligned (AUC 0.354, 95% CI 0.224-0.493). No positively aligned cells survived the Benjamini-Hochberg adjustment. Selected backbones produced image-responsive biomarker outputs, including pigment epithelial detachment balanced accuracy up to 0.93, whereas a five-field tabular reference model achieved decision AUC 0.731 (95% CI 0.603-0.846). Similar direct-contrast findings occurred in the second-site stress test. Because the generated rationales were not verified as faithful representations of latent computation, the grey image was out of distribution, and the cohorts were small, null direct contrasts cannot show that the models ignored OCT or establish equivalence to no discrimination.

Indexed as

anti-VEGFclinical AI evaluationdirect image contrastdistribution shiftgenerated rationalemultimodal fusionoptical coherence tomographytreatment-response predictionvision-language models

Identifiers

PMID42649801
PMCPMC13509487

What OpenQuestion holds

Textmetadata
Read underepoch 390

Registered trials

None linked

Read under generation 80e0d062 · epoch 390. Bibliography from PubMed, PubMed Central and OpenAlex; grants from NIH RePORTER; trial links from ClinicalTrials.gov; estimates, votes and beliefs from the OpenQuestion graph.