ML-QuantSubscribe

Machine learningLLMs & Text

Long-Form Answers to Visual Questions from Blind and Low Vision People

The study introduces VizWiz-LF, a dataset of long-form answers to visual questions asked by blind and low vision users, and assesses the ability of vision language models to provide accurate and useful responses.

Featured in No. 61 on 15 Aug 2024 · 3 days after release · 28 citations today

Released
12 Aug 2024
First featured
No. 61 · 15 Aug 2024
Citations (Semantic Scholar)
28
Influential citations
8
Published in
Not yet, as far as Semantic Scholar knows
Shares when featured
3
Identifier
arXiv:2408.06303

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page