mirror of
https://github.com/zenlm/vjepa2.git
synced 2026-07-26 22:28:08 +00:00
main
VJEPA2
Self-supervised visual representation learning from video. Part of the Zen LM ecosystem.
Overview
VJEPA2 implements Video Joint-Embedding Predictive Architecture for learning visual representations from unlabeled video data without relying on hand-crafted augmentations.
Features
- Self-supervised learning from video
- No hand-crafted augmentations required
- Pre-trained visual encoder for downstream tasks
- Efficient training with masking strategies
Related
License
See LICENSE file.
Languages
Python
95.5%
Jupyter Notebook
4.5%