Skip to content
View whikqp's full-sized avatar
  • Alibaba Cloud
  • Hangzhou, China
  • 22:47 (UTC +08:00)

Block or report whikqp

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Stars

CV Related

21 repositories

[CVPR 2024 Highlight] Putting the Object Back Into Video Object Segmentation

Python 1,118 118 Updated Nov 8, 2024

This is an automatic full segmentation tool based on Segment-Anything-2 and Segment-Anything-1. Our tool performs automatic full segmentation of the video, enabling the tracking of each object and …

Python 230 22 Updated Jul 9, 2025

Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything

Jupyter Notebook 17,751 1,595 Updated Sep 5, 2024

Based on GroundingDino and SAM, use semantic strings to segment any element in an image. The comfyui version of sd-webui-segment-anything.

Python 1,116 110 Updated Jul 12, 2024

ComfyUI nodes to use segment-anything-2

Python 1,221 87 Updated Sep 28, 2025

Fast and flexible image augmentation library. Paper about the library: https://www.mdpi.com/2078-2489/11/2/125

Python 15,308 1,708 Updated Jun 25, 2025

[CAAI AIR'24] Bilateral Reference for High-Resolution Dichotomous Image Segmentation

Python 4,263 336 Updated Sep 2, 2026

[CVPR 2025 Best Paper Award] VGGT: Visual Geometry Grounded Transformer

Python 14,462 1,560 Updated May 19, 2026

CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image

Jupyter Notebook 34,427 4,055 Updated Mar 25, 2026

An open source implementation of CLIP.

Python 14,188 1,319 Updated Sep 30, 2026

Semantic segmentation models with 500+ pretrained convolutional and transformer-based backbones.

Python 11,756 1,844 Updated Oct 5, 2026

[ECCV 2024] Official implementation of the paper "Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection"

Python 10,657 1,081 Updated Aug 12, 2024

[ECCV 2024] Official implementation of the paper "Semantic-SAM: Segment and Recognize Anything at Any Granularity"

Python 2,854 146 Updated Jul 10, 2025

real time face swap and one-click video deepfake with only a single image

Python 96,943 14,122 Updated Oct 5, 2026
Jupyter Notebook 235 33 Updated Aug 5, 2025

[ECCV 2024] The official code of paper "Open-Vocabulary SAM".

Python 1,034 36 Updated Sep 8, 2026

[CVPR 2025] MIDI: Multi-Instance Diffusion for Single Image to 3D Scene Generation

Python 951 75 Updated Jun 12, 2025

[CVPR 2024 🔥] GeoChat, the first grounded Large Vision Language Model for Remote Sensing

Python 759 67 Updated Nov 28, 2024
92 1 Updated Dec 12, 2024

Detectron2 is a platform for object detection, segmentation and other visual recognition tasks.

Python 34,766 7,935 Updated Sep 30, 2026

We write your reusable computer vision tools. 💜

Python 51,154 4,866 Updated Oct 8, 2026