V-JEPA2 (MLX)
Collection
Apple MLX fp16 ports of Meta V-JEPA2 ViT-L — video embeddings, JEPA predictor, SSv2 classifier. MIT. • 3 items • Updated • 2
How to use mlx-community/V-JEPA2-vitl-fpc64-256 with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] hf download mlx-community/V-JEPA2-vitl-fpc64-256 --local-dir V-JEPA2-vitl-fpc64-256
Apple MLX fp16 port of Meta's V-JEPA2
ViT-L/16 (facebook/vjepa2-vitl-fpc64-256): video/image embedding extraction
plus the JEPA latent-space predictor (masked world-model). MIT.
pip install vjepa2-mlx # https://github.com/xocialize/vjepa2-mlx
vjepa2-mlx -i clip.mp4 --task embed -o emb.npy
from vjepa2_mlx.pipeline_mlx import embed_video
emb = embed_video("clip.mp4", num_frames=16) # (1024,)
MIT (© Meta Platforms). Action-conditioned (robotics) predictor not included — separate ViT-g model.
Quantized