Machine learningLLMs & Text
SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living
The study presents SKI models, which incorporate 3D skeletons into the vision-language embedding space, using a skeleton-language model to enhance Vision Language Models and Large Vision Language Models.
Featured in No. 95 on 30 Apr 2025 · · 6 citations today · published in AAAI Conference on Artificial Intelligence
- Released
- 5 Feb 2025
- First featured
- No. 95 · 30 Apr 2025
- Citations (Semantic Scholar)
- 6
- Influential citations
- 0
- Published in
- AAAI Conference on Artificial Intelligence
- Shares when featured
- 3
- Identifier
- arXiv:2502.03459
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).