CORTEXA
← Browse

Yingying Zhao

4 papers indexed

arxivcs.CVcs.AI2026-07-31

QR-Structured Thermal Triggers for Targeted Semantic Attacks on Infrared Vision-Language Models

Xiang Chen, Yingying Zhao, Chao Li, Jiaju Han, Ben Zhang, Ang Li, et al.

Infrared vision-language models (IR-VLMs) extend thermal perception to open-vocabulary classification, image captioning, and visual question answering. However, their robustness to structured thermal perturbations and the stability of cross-modal semantic alignment remain insuffi…

View free PDFSource page
openalexScientific Reports2026-07-23

Multimodal learning for clinically consistent RGP fitting in keratoconus

Hongbiao Xie, Yingying Zhao, Peifang Xu, J P Ye, Gangyong Jia, Lin An

Abstract Keratoconus is a progressive corneal disorder characterized by highly heterogeneous corneal morphology, which makes rigid gas permeable (RGP) lens fitting strongly dependent on clinician experience and iterative trial processes. This procedure is often time-consuming and…

View free PDFSource page
arxivcs.CV2026-07-08

InfraQR: Edge-Placed QR-Inspired Structured Patch Attacks on Infrared Vision-Language Models

Xin Li, Jiaju Han, Ma Yaqi, Chengyin Hu, Yingying Zhao, Jiahuan Long, et al.

Infrared vision-language models are increasingly used for perception under low-light and adverse visual conditions, yet their robustness to localized structured perturbations remains underexplored. Existing infrared adversarial studies mainly focus on object detectors, leaving th…

View free PDFSource page
arxivcs.CV2026-07-07

MonoIR-RS: Infrared Remote Sensing Vision-Language Learning with CLIP and VLM Adaptation

Jiaju Han, Ma Yaqi, Yahui Chai, Xuemeng Sun, Xin Li, Qike Zhang, et al.

Infrared remote-sensing imagery captures intensity structure, object-background contrast, and illumination-invariant cues often invisible in RGB imagery. Yet, most remote-sensing vision-language resources and models focus on visible-band semantics, leaving infrared vision-languag…

View free PDFSource page