3,940 skills found · Page 1 of 132
ultralytics / UltralyticsUltralytics YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking
HumanSignal / Label StudioLabel Studio is a multi-type data labeling and annotation tool with standardized output format
amusi / CVPR2026 Papers With CodeCVPR 2026 论文和开源项目合集
cvat-ai / CvatComputer Vision Annotation Tool (CVAT) is a leading platform for building high-quality visual datasets for vision AI. It offers open-source, cloud, and enterprise products, as well as labeling services, for image, video, and 3D annotation with AI-assisted labeling, quality assurance, team collaboration, analytics, and developer APIs.
wkentaro / LabelmeImage annotation with Python. Supports polygon, rectangle, circle, line, point, and AI-assisted annotation.
microsoft / Swin TransformerThis is an official implementation for "Swin Transformer: Hierarchical Vision Transformer using Shifted Windows".
qubvel-org / Segmentation Models.pytorchSemantic segmentation models with 500+ pretrained convolutional and transformer-based backbones.
milesial / Pytorch UNetPyTorch implementation of the U-Net for image semantic segmentation with high quality images
mrgloom / Awesome Semantic Segmentation:metal: awesome-semantic-segmentation
OpenGVLab / InternVL[CVPR 2024 Oral] InternVL Family: A Pioneering Open-Source Alternative to GPT-4o. 接近GPT-4o表现的开源多模态对话模型
open-mmlab / MmsegmentationOpenMMLab Semantic Segmentation Toolbox and Benchmark.
PaddlePaddle / PaddleSegEasy-to-use image segmentation library with awesome pre-trained model zoo, supporting wide-range of practical tasks in Semantic Segmentation, Interactive Segmentation, Panoptic Segmentation, Image Matting, 3D Segmentation, etc.
dmlc / Gluon CvGluon CV Toolkit
ashishpatel26 / Tools To Design Or Visualize Architecture Of Neural NetworkTools to Design or Visualize Architecture of Neural Network
CSAILVision / Semantic Segmentation PytorchPytorch implementation for Semantic Segmentation/Scene Parsing on MIT ADE20K dataset
fundamentalvision / BEVFormer[ECCV 2022] This is the official implementation of BEVFormer, a camera-only framework for autonomous driving perception, e.g., 3D object detection and semantic map segmentation.
HumanSignal / Awesome Data LabelingA curated list of awesome data labeling tools
NVlabs / SegFormerOfficial PyTorch implementation of SegFormer
meetps / Pytorch SemsegSemantic Segmentation Architectures Implemented in PyTorch
luigifreda / PyslampySLAM is a hybrid Python/C++ Visual SLAM pipeline supporting monocular, stereo, and RGB-D cameras. It provides a broad set of modern local and global feature extractors, multiple loop-closure strategies, a volumetric reconstruction module, integrated depth-prediction models, and semantic segmentation capabilities for enhanced scene understanding.