From vision–language models to structure-aware visual reasoning: foundations, representations, and consistency checking

Xiaoyan Dai

Open source

DOI
10.3389/frobt.2026.1781874
Published
2026-06-24
Container
Frontiers in Robotics and AI
Publisher
Frontiers Media SA
Open access
unknown

Credibility signals

uncertain Score 64/100 under policy 1.0.0. This is a metadata assessment, not a judgment of the paper's conclusions.

Show all credibility signals

Cite this work

BibTeX

@article{allodium:10.3389/frobt.2026.1781874,
  title = {From vision–language models to structure-aware visual reasoning: foundations, representations, and consistency checking},
  author = {Xiaoyan Dai},
  year = {2026},
  journal = {Frontiers in Robotics and AI},
  doi = {10.3389/frobt.2026.1781874},
  url = {https://doi.org/10.3389/frobt.2026.1781874}
}

RIS

TY  - JOUR
TI  - From vision–language models to structure-aware visual reasoning: foundations, representations, and consistency checking
AU  - Xiaoyan Dai
PY  - 2026
JO  - Frontiers in Robotics and AI
DO  - 10.3389/frobt.2026.1781874
UR  - https://doi.org/10.3389/frobt.2026.1781874
ER  - 

APA

Dai, X. (2026). From vision–language models to structure-aware visual reasoning: foundations, representations, and consistency checking. Frontiers in Robotics and AI. https://doi.org/10.3389/frobt.2026.1781874

Source records