Machine learningLLMs & Text
Unveiling Encoder-Free Vision-Language Models
The EVE model, a vision-language model without an encoder, performs well across multiple benchmarks, offering an efficient way to develop a decoder-only architecture.
Featured in No. 54 on 20 Jun 2024 · 3 days after release · 102 citations today · published in Neural Information Processing Systems
- Released
- 17 Jun 2024
- First featured
- No. 54 · 20 Jun 2024
- Citations (Semantic Scholar)
- 102
- Influential citations
- 17
- Published in
- Neural Information Processing Systems
- Shares when featured
- 73
- Identifier
- arXiv:2406.11832
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).