CORTEXA
← Browse

Yan Gao

3 papers indexed

arxivcs.CV2026-06-26

ReScene: Structured Indoor Scene Reconstruction from Multi-View Captures

Haoran Xu, Lechao Zhang, Daoguo Dong, Yan Gao, Xin Tan

Constructing simulation-ready 3D scenes from multi-view captures is a key bottleneck for Embodied Artificial Intelligence, as downstream tasks require object-level structure, explicit inter-object relations, and physical plausibility. Existing approaches either rely on specialize…

View free PDFSource page
arxivcs.CVcs.CL2026-06-25

HarmVideoBench: Benchmarking Harmful Video Understanding in Large Multimodal Models

Jiajun Wu, Haoyu Kang, Yining Sun, Jiacheng Hou, Heng Zhang, Danyang Zhang, et al.

Large vision-language models (LVLMs) have recently shown immense potential in automated content moderation, sparking growing interest in developing harmful-video benchmarks. However, we identify two primary limitations in existing works: 1) The multi-layered characteristics of ha…

View free PDFSource page