ML-QuantSubscribe

Machine learningOther

VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks

VisualWebArena, a benchmark for assessing the performance of multimodal web agents on visually grounded tasks, is introduced, highlighting gaps in current multimodal language agents.

Featured in No. 35 on 30 Jan 2024 · 6 days after release · 0 citations today

Released
24 Jan 2024
First featured
No. 35 · 30 Jan 2024
Citations (Semantic Scholar)
0
Influential citations
0
Published in
Not yet, as far as Semantic Scholar knows
Shares when featured
56
Identifier
arXiv:2401.13649

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page