Machine learningLLMs & Text
Sparrow: Data-Efficient Video-LLM With Text-to-Image Augmentation
The T2Vid method, developed by researchers, uses pre-trained image-LLMs to enhance video understanding, performing as well or better than full video datasets with only 15% of the sample size.
Featured in No. 77 on 4 Dec 2024 · 5 days after release · 0 citations today · published in IEEE Transactions on Multimedia
- Released
- 29 Nov 2024
- First featured
- No. 77 · 4 Dec 2024
- Citations (Semantic Scholar)
- 0
- Influential citations
- 0
- Published in
- IEEE Transactions on Multimedia
- Shares when featured
- 6
- Identifier
- arXiv:2411.19951
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).