★ 54,951GitHub describes it as The repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.. The repository metadata lists Jupyter Notebook as its primary language. The metadata lists the Apache-2.0 license. This article stays within the project description and details documented in the GitHub repository README.
Jupyter Notebook6,381 forks
★ 40,980A library for efficient similarity search and clustering of dense vectors.
C++Developer Tools4,530 forks
★ 34,736Detectron2 is a platform for object detection, segmentation and other visual recognition tasks.
Python7,933 forks
★ 23,651Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressor / tokenizer, along with MusicGen, a simple and controllable music generation LM with textual and melodic conditioning.
Jupyter Notebook2,697 forks
★ 19,938The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
Jupyter Notebook2,557 forks
★ 14,442[CVPR 2025 Best Paper Award] VGGT: Visual Geometry Grounded Transformer
Python1,557 forks
★ 13,380PyTorch code and models for the DINOv2 self-supervised learning method.
Jupyter Notebook1,270 forks
★ 11,883Foundational Models for State-of-the-Art Speech and Text Translation
Jupyter Notebook1,184 forks
★ 11,840The repository provides code for running inference and finetuning with the Meta Segment Anything Model 3 (SAM 3), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
Python1,796 forks
★ 11,468Reference PyTorch implementation and models for DINOv3
Jupyter Notebook968 forks
★ 10,553Hackable and optimized Transformers building blocks, supporting a composable construction.
Python794 forks
★ 9,972PyTorch3D is FAIR's library of reusable components for deep learning with 3D data
Python1,467 forks
★ 7,426PySlowFast: video understanding codebase from FAIR for reproducing state-of-the-art video models.
Python1,296 forks
★ 5,633A modular framework for vision & language multimodal research from Facebook AI Research (FAIR)
Python938 forks
★ 5,425High-resolution models for human tasks.
Python323 forks
★ 5,128CoTracker is a model for tracking any point (pixel) on a video.
Jupyter Notebook392 forks
★ 5,099A data augmentations library for audio, image, text, and video.
Python312 forks
★ 4,692PyTorch code and models for VJEPA2 self-supervised learning from video.
Python584 forks
★ 4,601[CVPR 2026 Oral] VGGT Omega
Python360 forks
★ 4,210A Python toolbox for performing gradient-free optimization
Python371 forks