VL-JEPA: Joint Embedding Predictive Architecture for Vision-Language (arxiv.org) 3 points by hbarka 8mo ago ↗ HN
0 comments
[ 3.2 ms ] story [ 9.7 ms ] threadNo comments yet.