Machine learningLLMs & Text
An Investigation on LLMs' Visual Understanding Ability Using SVG for Image-Text Bridging
The research examines the capability of Large Language Models (LLMs) to interpret images by transforming them into Scalable Vector Graphics (SVG) and assessing the LLMs on various computer vision tasks.
Featured in No. 57 on 17 Jul 2024 · · 8 citations today · published in 2025 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV)
- Released
- 9 Jun 2023
- First featured
- No. 57 · 17 Jul 2024
- Citations (Semantic Scholar)
- 8
- Influential citations
- 0
- Published in
- 2025 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV)
- Shares when featured
- 65
- Identifier
- arXiv:2306.06094
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).