ML-QuantSubscribe

Machine learningLLMs & Text

An Investigation on LLMs' Visual Understanding Ability Using SVG for Image-Text Bridging

The research examines the capability of Large Language Models (LLMs) to interpret images by transforming them into Scalable Vector Graphics (SVG) and assessing the LLMs on various computer vision tasks.

Featured in No. 57 on 17 Jul 2024 · · 8 citations today · published in 2025 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV)

Released
9 Jun 2023
First featured
No. 57 · 17 Jul 2024
Citations (Semantic Scholar)
8
Influential citations
0
Published in
2025 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV)
Shares when featured
65
Identifier
arXiv:2306.06094

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page