Effortless AI-assisted data labeling with AI support from YOLO, Segment Anything (SAM+SAM2/2.1+SAM3), MobileSAM!!
-
Updated
Aug 30, 2026 - Python
Effortless AI-assisted data labeling with AI support from YOLO, Segment Anything (SAM+SAM2/2.1+SAM3), MobileSAM!!
Labeling tool with SAM(segment anything model),supports SAM, SAM2, SAM3, sam-hq, MobileSAM EdgeSAM etc.交互式半自动图像标注工具
Tailor是一款视频智能裁剪、视频生成和视频优化的视频剪辑工具。目前的目标是通过人工智能技术减少视频剪辑的繁琐操作,让普通人也能简单实现专业剪辑人的水准!长远目标是让视频剪辑实现真正的AIGC!
SimpleAICV: pytorch training examples.
[CVPR 2025] Official PyTorch implementation of "EdgeTAM: On-Device Track Anything Model"
ComfyUI nodes for vision-language models: Qwen3-VL, Moondream 3, Florence-2, SmolVLM2, InternVL, Gemma 3, MiniCPM-V. Plus open-vocabulary detection, SAM2/SAM3 segmentation, video temporal reasoning, GGUF via llama.cpp, and hosted LLM/VLM APIs.
Export and run SAM, MobileSAM, EfficientSAM, SAM 2/2.1, and SAM 3 as ONNX for portable image segmentation
[CVPR 2025] Code for Segment Any Motion in Videos
The code for PixelRefer & VideoRefer
An open-source studio for prompt-driven video segmentation. Powered by SAM2 & Grounding DINO with a hybrid Cloud-Local architecture.
[CVPR 2026] OccAny: Generalized Unconstrained Urban 3D Occupancy. The first Unified Framework for Generalized 3D Occupancy Prediction. Supports SAM2/SAM3, MUSt3R & Depth Anything 3.
Video-Inpaint-Anything: This is the inference code for our paper CoCoCo: Improving Text-Guided Video Inpainting for Better Consistency, Controllability and Compatibility.
Grounded Tracking for Streaming Videos
Zero-shot segmentation of very large remote sensing images with SAM2: multi-pass coverage maximization and parameter-free tile-boundary merging (Remote SAMsing)
Playground Web UI using segment-anything-2 models from the Meta.
PEANUT (Prompt-Enhanced Ablation with Optical Flow-Based Neural Unit) designed to enhance video restoration by combining spatial and temporal consistency with clarity optimization. The core innovation lies in Prompt-Guided Mask Self-Generation and leveraging optical flow-based neural units to generate high-fidelity video sequences
A system for barcode detection and decoding using YOLO, SAM2, and pyzbar designed to extract barcode data from images.
Interactive vision input toolkit for CV frameworks, generate bbox, polygons, lines and points for Ultralytics, SAM, Supervision and many more.
[CVPR'26 Highlight] Official Code for “V²-SAM: Marrying SAM2 with Multi-Prompt Experts for Cross-View Object Correspondence”
To associate your repository with the sam2 topic, visit your repo's landing page and select "manage topics."