Machine learningLLMs & Text
Long-Form Answers to Visual Questions from Blind and Low Vision People
The study introduces VizWiz-LF, a dataset of long-form answers to visual questions asked by blind and low vision users, and assesses the ability of vision language models to provide accurate and useful responses.
Featured in No. 61 on 15 Aug 2024 · 3 days after release · 28 citations today
- Released
- 12 Aug 2024
- First featured
- No. 61 · 15 Aug 2024
- Citations (Semantic Scholar)
- 28
- Influential citations
- 8
- Published in
- Not yet, as far as Semantic Scholar knows
- Shares when featured
- 3
- Identifier
- arXiv:2408.06303
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).