An open API service indexing awesome lists of open source software.

Projects in Awesome Lists tagged with segment-anything

A curated list of projects in awesome lists tagged with segment-anything .

https://github.com/gaomingqi/track-anything

Track-Anything is a flexible and interactive tool for video object tracking and segmentation, based on Segment Anything, XMem, and E2FGVI.

inpaint-anything interactive-tracking segment-anything track-anything video-object-segmentation video-object-tracking

Last synced: 14 May 2025

https://github.com/gaomingqi/Track-Anything

Track-Anything is a flexible and interactive tool for video object tracking and segmentation, based on Segment Anything, XMem, and E2FGVI.

inpaint-anything interactive-tracking segment-anything track-anything video-object-segmentation video-object-tracking

Last synced: 20 Mar 2025

https://github.com/syscv/sam-hq

Segment Anything in High Quality [NeurIPS 2023]

high-quality sam segment-anything segment-anything-model segmentation zero-shot-segmentation

Last synced: 13 May 2025

https://github.com/SysCV/SAM-HQ

Segment Anything in High Quality [NeurIPS 2023]

high-quality sam segment-anything segment-anything-model segmentation zero-shot-segmentation

Last synced: 06 May 2025

https://github.com/SysCV/sam-hq

Segment Anything in High Quality [NeurIPS 2023]

high-quality sam segment-anything segment-anything-model segmentation zero-shot-segmentation

Last synced: 02 Apr 2025

https://github.com/opengeos/segment-geospatial

A Python package for segmenting geospatial data with the Segment Anything Model (SAM)

artificial-intelligence deep-learning geopython geospatial machine-learning segment-anything segmentation

Last synced: 12 May 2025

https://github.com/OpenGVLab/InternGPT

InternGPT (iGPT) is an open source demo platform where you can easily showcase your AI models. Now it supports DragGAN, ChatGPT, ImageBind, multimodal chat like GPT-4, SAM, interactive image editing, etc. Try it at igpt.opengvlab.com (支持DragGAN、ChatGPT、ImageBind、SAM的在线Demo系统)

chatgpt click draggan foundation-model gpt gpt-4 gradio husky image-captioning imagebind internimage langchain llama llm multimodal sam segment-anything vicuna video-generation vqa

Last synced: 27 Mar 2025

https://github.com/opengvlab/interngpt

InternGPT (iGPT) is an open source demo platform where you can easily showcase your AI models. Now it supports DragGAN, ChatGPT, ImageBind, multimodal chat like GPT-4, SAM, interactive image editing, etc. Try it at igpt.opengvlab.com (支持DragGAN、ChatGPT、ImageBind、SAM的在线Demo系统)

chatgpt click draggan foundation-model gpt gpt-4 gradio husky image-captioning imagebind internimage langchain llama llm multimodal sam segment-anything vicuna video-generation vqa

Last synced: 14 May 2025

https://github.com/z-x-yang/segment-and-track-anything

An open-source project dedicated to tracking and segmenting any objects in videos, either automatically or interactively. The primary algorithms utilized include the Segment Anything Model (SAM) for key-frame segmentation and Associating Objects with Transformers (AOT) for efficient tracking and propagation purposes.

interactive-segmentation segment-anything segment-anything-model video-object-segmentation visual-object-tracking

Last synced: 13 Apr 2025

https://github.com/z-x-yang/Segment-and-Track-Anything

An open-source project dedicated to tracking and segmenting any objects in videos, either automatically or interactively. The primary algorithms utilized include the Segment Anything Model (SAM) for key-frame segmentation and Associating Objects with Transformers (AOT) for efficient tracking and propagation purposes.

interactive-segmentation segment-anything segment-anything-model video-object-segmentation visual-object-tracking

Last synced: 16 Mar 2025

https://github.com/vietanhdev/anylabeling

Effortless AI-assisted data labeling with AI support from YOLO, Segment Anything (SAM+SAM2), MobileSAM!!

auto-labeling computer-vision labeling labeling-tool mobilesam onnx sam2 segment-anything segment-anything-2 yolo yolov8

Last synced: 13 May 2025

https://github.com/ttengwang/caption-anything

Caption-Anything is a versatile tool combining image segmentation, visual captioning, and ChatGPT, generating tailored captions with diverse controls for user preferences. https://huggingface.co/spaces/TencentARC/Caption-Anything https://huggingface.co/spaces/VIPLab/Caption-Anything

chatgpt controllable-generation controllable-image-captioning image-captioning segment-anything

Last synced: 15 May 2025

https://github.com/ttengwang/Caption-Anything

Caption-Anything is a versatile tool combining image segmentation, visual captioning, and ChatGPT, generating tailored captions with diverse controls for user preferences. https://huggingface.co/spaces/TencentARC/Caption-Anything https://huggingface.co/spaces/VIPLab/Caption-Anything

chatgpt controllable-generation controllable-image-captioning image-captioning segment-anything

Last synced: 12 Mar 2025

https://github.com/anything-of-anything/anything-3d

Segment-Anything + 3D. Let's lift anything to 3D.

3d computer-vision reconstruction segment segment-anything

Last synced: 16 May 2025

https://github.com/Anything-of-anything/Anything-3D

Segment-Anything + 3D. Let's lift anything to 3D.

3d computer-vision reconstruction segment segment-anything

Last synced: 20 Mar 2025

https://github.com/yatenglg/isat_with_segment_anything

Labeling tool with SAM(segment anything model),supports SAM, SAM2, sam-hq, MobileSAM EdgeSAM etc.交互式半自动图像标注工具

annotation-tool computer-vision labeling labeling-tool sam sam2 segment-anything segment-anything-2 video-segmentation

Last synced: 24 Dec 2025

https://github.com/openadaptai/openadapt

Open Source Generative Process Automation (i.e. Generative RPA). AI-First Process Automation with Large ([Language (LLMs) / Action (LAMs) / Multimodal (LMMs)] / Visual Language (VLMs)) Models

agents ai-agents ai-agents-framework anthropic computer-use generative-process-automation google-gemini gpt4o huggingface large-action-model large-language-models large-multimodal-models omniparser openai process-automation process-mining python segment-anything transformers ultralytics

Last synced: 04 Mar 2026

https://github.com/DWCTOD/CVPR2024-Papers-with-Code-Demo

收集 CVPR 最新的成果,包括论文、代码和demo视频等,欢迎大家推荐!Collect the latest CVPR (Conference on Computer Vision and Pattern Recognition) results, including papers, code, and demo videos, etc., and welcome recommendations from everyone!

computer-vision cvpr cvpr2021 cvpr2022 cvpr2023 cvpr2024 llm multimodal-deep-learning object-detection segment-anything segmentation

Last synced: 29 Mar 2025

https://github.com/dwctod/cvpr2024-papers-with-code-demo

收集 CVPR 最新的成果,包括论文、代码和demo视频等,欢迎大家推荐!Collect the latest CVPR (Conference on Computer Vision and Pattern Recognition) results, including papers, code, and demo videos, etc., and welcome recommendations from everyone!

computer-vision cvpr cvpr2021 cvpr2022 cvpr2023 cvpr2024 llm multimodal-deep-learning object-detection segment-anything segmentation

Last synced: 26 Jan 2026

https://github.com/yatengLG/ISAT_with_segment_anything

Labeling tool with SAM(segment anything model),supports SAM, SAM2, sam-hq, MobileSAM EdgeSAM etc.交互式半自动图像标注工具

annotation-tool computer-vision labeling labeling-tool sam sam2 segment-anything segment-anything-2 video-segmentation

Last synced: 09 Mar 2025

https://github.com/wladradchenko/wunjo.wladradchenko.ru

Wunjo CE: Face Swap, Lip Sync, Control Remove Objects & Text & Background, Restyling, Audio Separator, Clone Voice, Video Generation. Open Source, Local & Free.

controlnet deepfake diffusion-models face-animation face-swap free img2video lip-sync photo-editing public-api remove-background remover restyle segment-anything txt2video video-editing video-generation voice-clone wunjo

Last synced: 30 Apr 2026

https://github.com/OpenAdaptAI/OpenAdapt

Open Source Generative Process Automation (i.e. Generative RPA). AI-First Process Automation with Large ([Language (LLMs) / Action (LAMs) / Multimodal (LMMs)] / Visual Language (VLMs)) Models

agents ai-agents ai-agents-framework anthropic computer-use generative-process-automation google-gemini gpt4o huggingface large-action-model large-language-models large-multimodal-models omniparser openai process-automation process-mining python segment-anything transformers ultralytics

Last synced: 05 Apr 2025

https://github.com/kadirnar/segment-anything-video

MetaSeg: Packaged version of the Segment Anything repository

object-detection object-segmentation segment-anything segmentation yolov5 yolov6 yolov7 yolov8

Last synced: 14 May 2025

https://github.com/chongzhou96/EdgeSAM

Official PyTorch implementation of "EdgeSAM: Prompt-In-the-Loop Distillation for On-Device Deployment of SAM"

coreml on-device-ai segment-anything

Last synced: 20 Mar 2025

https://github.com/storyicon/comfyui_segment_anything

Based on GroundingDino and SAM, use semantic strings to segment any element in an image. The comfyui version of sd-webui-segment-anything.

comfyui custom-nodes groundingdino sam segment-anything stable-diffusion

Last synced: 16 May 2025

https://github.com/ziqi-jin/finetune-anything

Fine-tune SAM (Segment Anything Model) for computer vision tasks such as semantic segmentation, matting, detection ... in specific scenarios

computer-vision deep-learning fine-tune segment-anything

Last synced: 06 May 2025

https://github.com/Jumpat/SegmentAnythingin3D

Segment Anything in 3D with NeRFs (NeurIPS 2023)

3d 3d-segmentation computer-vision deep-learning nerf segment-anything segmentation

Last synced: 20 Mar 2025

https://github.com/nvidia-ai-iot/nanosam

A distilled Segment Anything (SAM) model capable of running real-time with NVIDIA TensorRT

jetson-orin jetson-orin-nano nvidia real-time segment-anything tensorrt

Last synced: 28 Jun 2025

https://github.com/NVIDIA-AI-IOT/nanosam

A distilled Segment Anything (SAM) model capable of running real-time with NVIDIA TensorRT

jetson-orin jetson-orin-nano nvidia real-time segment-anything tensorrt

Last synced: 20 Mar 2025

https://github.com/leondgarse/keras_cv_attention_models

Keras beit,caformer,CMT,CoAtNet,convnext,davit,dino,efficientdet,edgenext,efficientformer,efficientnet,eva,fasternet,fastervit,fastvit,flexivit,gcvit,ghostnet,gpvit,hornet,hiera,iformer,inceptionnext,lcnet,levit,maxvit,mobilevit,moganet,nat,nfnets,pvt,swin,tinynet,tinyvit,uniformer,volo,vanillanet,yolor,yolov7,yolov8,yolox,gpt2,llama2, alias kecam

attention clip coco ddpm detection imagenet keras model recognition segment-anything stable-diffusion tensorflow tf tf2 visualizing

Last synced: 08 Apr 2025

https://github.com/dvlab-research/3d-box-segment-anything

We extend Segment Anything to 3D perception by combining it with VoxelNeXt.

3d autonomous-driving segment-anything

Last synced: 03 Jul 2025

https://github.com/yeungchenwa/OCR-SAM

Combining MMOCR with Segment Anything & Stable Diffusion. Automatically detect, recognize and segment text instances, with serval downstream tasks, e.g., Text Removal and Text Inpainting

mmocr segment-anything text-detection text-inpainting text-recognition text-removal

Last synced: 02 May 2025

https://github.com/SuperMedIntel/Medical-SAM2

Medical SAM 2: Segment Medical Images As Video Via Segment Anything Model 2

deep-learning medical medical-imaging segment-anything segment-anything-2 segment-anything-model segmentation

Last synced: 24 Jul 2025

https://github.com/hitachinsk/SAMed

The implementation of the technical report: "Customized Segment Anything Model for Medical Image Segmentation"

medical-imaging segment-anything segment-anything-model segmentation

Last synced: 03 Apr 2025

https://github.com/dvlab-research/3D-Box-Segment-Anything

We extend Segment Anything to 3D perception by combining it with VoxelNeXt.

3d autonomous-driving segment-anything

Last synced: 20 Mar 2025

https://github.com/xinghaochen/tinysam

[AAAI 2025] Official PyTorch implementation of "TinySAM: Pushing the Envelope for Efficient Segment Anything Model"

efficient-sam sam segment-anything tinysam

Last synced: 16 May 2025

https://github.com/hustvl/evf-sam

Official code of "EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model"

multimodal multimodal-large-language-models referring-image-segmentation segment-anything segmentation

Last synced: 16 May 2025

https://github.com/xmed-lab/CLIP_Surgery

CLIP Surgery for Better Explainability with Enhancement in Open-Vocabulary Tasks

clip explainability interpretability multilabel multimodal open-vocabulary sam segment-anything segmentation vision-transformer

Last synced: 16 Mar 2025

https://github.com/xinghaochen/TinySAM

Official PyTorch implementation of "TinySAM: Pushing the Envelope for Efficient Segment Anything Model"

efficient-sam sam segment-anything tinysam

Last synced: 20 Mar 2025

https://github.com/OpenGVLab/Instruct2Act

Instruct2Act: Mapping Multi-modality Instructions to Robotic Actions with Large Language Model

chatgpt clip llm robotics segment-anything

Last synced: 06 May 2025

https://github.com/opengvlab/instruct2act

Instruct2Act: Mapping Multi-modality Instructions to Robotic Actions with Large Language Model

chatgpt clip llm robotics segment-anything

Last synced: 20 Apr 2025

https://github.com/vietanhdev/samexporter

Export Segment Anything Models to ONNX

onnx segment-anything segment-anything-model

Last synced: 04 Apr 2025

https://github.com/hustvl/EVF-SAM

Official code of "EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model"

multimodal multimodal-large-language-models referring-image-segmentation segment-anything segmentation

Last synced: 07 May 2025

https://github.com/RockeyCoss/Prompt-Segment-Anything

This is an implementation of zero-shot instance segmentation using Segment Anything.

instance-segmentation mmdetection segment-anything

Last synced: 06 May 2025

https://github.com/zibojia/COCOCO

Video-Inpaint-Anything: This is the inference code for our paper CoCoCo: Improving Text-Guided Video Inpainting for Better Consistency, Controllability and Compatibility.

cococo diffusion inpainting pytorch sam2 segment segment-anything text-guided text-guided-video-inpainting video-inpainting video-inpainting-with-prompt video-sam2-inpaint

Last synced: 24 Jul 2025

https://github.com/jasonaidm/ai_webui

AI-WEBUI: A universal web interface for AI creation, 一款好用的图像、音频、视频AI处理工具

ai chatbot chatglm chatgpt inpainting sam segment-anything speech-recognition speech-synthesis video-clip

Last synced: 24 Mar 2025

https://github.com/adaptivemotorcontrollab/amadeusgpt

We turn natural language descriptions of behaviors into machine-executable code

amadeusgpt cebra chatgpt deeplabcut llms segment-anything

Last synced: 16 May 2025

https://github.com/ymy-k/Hi-SAM

[IEEE TPAMI] Hi-SAM: Marrying Segment Anything Model for Hierarchical Text Segmentation

hierarchical-text-segmentation high-quality-text-stroke-segmentation sam segment-anything segment-anything-model

Last synced: 24 Jul 2025

https://github.com/halleewong/ScribblePrompt

[ECCV 2024] ScribblePrompt: Fast and Flexible Interactive Segmentation for Any Medical Image

interactive-segmentation medical-imaging segment-anything segmentation

Last synced: 05 Mar 2026

https://github.com/pengtaojiang/segment-anything-clip

Connecting segment-anything's output masks with the CLIP model; Awesome-Segment-Anything-Works

classification clip segment-anything semantic-segmentation

Last synced: 04 Apr 2025

https://github.com/positive666/prompt-can-anything

You can do anything by sota AI with prompt ,auto AI tools , VL larger model fine and project

automatic chatgpt detection grouned python segment-anything stable-diffusion visualglm

Last synced: 14 Aug 2025

https://github.com/MaybeShewill-CV/segment-anything-u-specify

using clip and sam to segment any instance you specify with text prompt of any instance names

deep-learning instance-segmentation object-detection sam-model segment-anything

Last synced: 06 May 2025

https://github.com/ngthanhtin/owlvit_segment_anything

Combining OwlViT with Segment Anything - Open-vocabulary Detection and Segmentation (Text-conditioned, and Image-conditioned)

detection open-vocabulary owl-vit segment-anything segmentation small-object stable-diffusion

Last synced: 06 May 2025

https://github.com/ylqi/Count-Anything

This method uses Segment Anything and CLIP to ground and count any object that matches a custom text prompt, without requiring any point or box annotation.

clip count-anything segment-anything

Last synced: 23 Aug 2025

https://github.com/AnyLoc/Revisit-Anything

Code release for Revisit Anything: Visual Place Recognition via Image Segment Retrieval (ECCV 2024)

deep-learning descriptors image-retrieval place-recognition revisit-anything robotics sam segment-anything

Last synced: 24 Jul 2025

https://paulpanwang.github.io/POPE/

Welcome to the project repository for POPE (Promptable Pose Estimation), a state-of-the-art technique for 6-DoF pose estimation of any object in any scene using a single reference.

dinov2 image-matching pose-estimation segment-anything

Last synced: 24 Jul 2025

https://github.com/uncbiag/SegNext

Rethinking Interactive Image Segmentation with Low Latency, High Quality, and Diverse Prompts (CVPR 2024)

interactive-image-segmentation segment-anything vision-transformers

Last synced: 24 Jul 2025

https://github.com/GAP-LAB-CUHK-SZ/SAMPro3D

SAMPro3D: Locating SAM Prompts in 3D for Zero-Shot Scene Segmentation

3d-scene-understanding segment-anything zero-shot-segmentation

Last synced: 20 Mar 2025

https://github.com/kadirnar/comfyui-yolo

ComfyUI-YOLO: Ultralytics-Powered Object Recognition for ComfyUI

comfyu object-detection segment-anything ultralytics yolo

Last synced: 09 Apr 2025

https://github.com/sunu/sam-in-browser

A PoC to run Segment Anything Model (SAM) entirely in the browser without any backend

onnx onnxruntime-web segment-anything segment-anything-model

Last synced: 12 Feb 2026

https://github.com/eth-siplab/RAP

Restore Anything Pipeline (RAP): Segment Anything Meets Image Restoration

computer-vision deep-learning image-deblurring image-denoising image-restoration jpeg-artifact-removal segment-anything

Last synced: 24 Jul 2025

https://github.com/wkentaro/osam

Get up and running with SAM, EfficientSAM, YOLO-World, and other promptable vision models locally.

computer-vision deep-learning foundation-models onnx segment-anything

Last synced: 16 Jan 2026

https://github.com/kadirnar/ComfyUI-YOLO

ComfyUI-YOLO: Ultralytics-Powered Object Recognition for ComfyUI

comfyu object-detection segment-anything ultralytics yolo

Last synced: 19 Aug 2025

https://github.com/aim-uofa/segagent

[CVPR2025] SegAgent: Exploring Pixel Understanding Capabilities in MLLMs by Imitating Human Annotator Trajectories

agent mllms segment-anything vlms

Last synced: 28 Jan 2026

https://github.com/licksylick/AutoTrackAnything

AutoTrackAnything is a universal, flexible and interactive tool for insane automatic object tracking over thousands of frames. It is developed upon XMem, Yolov8 and MobileSAM (Segment Anything), can track anything which detect Yolov8.

mobilesam multi-object-tracker multi-object-tracking re-id re-identification reid sam segment-anything track-anything tracking tracking-by-detection xmem yolov8

Last synced: 08 Jul 2025

https://github.com/jhj0517/sam2-playground

Playground Web UI using segment-anything-2 models from the Meta.

ai gradio open-source sam sam2 segment-anything segment-anything-2 segmentation torch webui

Last synced: 17 Mar 2025

https://github.com/capjamesg/sam-clip

Use Grounding DINO, Segment Anything, and CLIP to label objects in images.

clip computer-vision segment-anything zero-shot-object-detection

Last synced: 19 Aug 2025

https://github.com/AIDajiangtang/Segment-Anything-CPP

segment anything(SAM) for CPP Inference

deep-learning segment-anything

Last synced: 18 Mar 2025

https://github.com/rekalantar/medsegmentanything_sam_lungct

The code to finetune SAM with bounding box prompt for segmentation of the lungs on CT

ct deep-learning huggingface lung-segmentation medical-image-segmentation sam segment-anything transformers

Last synced: 08 Sep 2025

https://github.com/zhudongwork/SAM_TensorRT

This is the code to implement Segment Anything (SAM) using TensorRT(C++).

segment-anything tensorrt

Last synced: 18 Mar 2025

https://github.com/prateekralhan/Segment-Anything-Streamlit

Streamlit based implementation for the The Segment Anything Model (SAM) developed by Meta AI research

facebook-research opensourceforgood python3 segment-anything segmentation-models streamlit-webapp

Last synced: 26 Sep 2025

https://github.com/neka-nat/mylangrobot

Language instructions to mycobot using GPT-4V

chatgpt gpt-4-vision gpt-4-vision-preview gpt4v mycobot segment-anything whisper

Last synced: 20 Jun 2025

https://github.com/lygitdata/garmentiq

Free & Open Source. Precise and flexible garment measurements from images - no tape measures, no delays, just fashion - forward automation.

artificial-intelligence computer-vision fashion garments hrnet measurement segment-anything transformers vision-transformer

Last synced: 16 Apr 2026