Skip to main content

YOLO Vision 2026

X-AnyLabeling interface

🥳 What's New

  • 2026-08-19: Add support for image tagging, with tag creation, editing, reordering, and batch deletion.
  • 2026-08-12: Add support for D-FINE-seg instance segmentation models.
  • 2026-08-08: Add support for the RT-DETRv2-OBB rotated object detection model.
  • 2026-08-08: Add the Magic Wand tool for quickly creating polygons from contiguous color regions.
  • 2026-08-05: Release X-AnyLabeling v4.0.0.
  • For more details, please refer to the CHANGELOG

Introduction

X-AnyLabeling is a lightweight, efficient, and unified cross-platform desktop application for AI-assisted annotation of text, image, video, and multimodal data. It combines versatile built-in tools, automated labeling workflows, state-of-the-art deep learning models, and flexible multi-format import and export. For remote inference, X-AnyLabeling-Server provides a lightweight, extensible backend for connecting custom models and compute resources.

Key Features

  • Unified support for annotating and processing text, image, video, and multimodal data.
  • Covers tasks such as image classification, object detection, instance segmentation, pose estimation, oriented object detection, multi-object tracking, optical character recognition, lane annotation, image captioning, visual question answering, and document parsing.
  • Provides polygons, rectangles, cuboids, rotated boxes, quadrilaterals, circles, lines, polylines, points, masks, and task-specific tools for text detection, text recognition, and KIE.
  • Integrates a wide range of state-of-the-art deep learning models for AI-assisted annotation, automated labeling, and batch dataset prediction.
  • Supports both local and remote inference through engines and serving frameworks such as ONNX Runtime, TensorRT, OpenCV DNN, vLLM, and SGLang.
  • Supports importing and exporting formats such as COCO, VOC, YOLO, DOTA, MOT, MASK, PPOCR, MMGD, VLM-R1, and ShareGPT.
  • Runs on Windows, Linux, and macOS, with interfaces available in English, Simplified Chinese, Japanese, and Korean.
  • Supports custom model integration, flexible extension, and secondary development.

Model library

Task Category Supported Models
🖼️ Image Classification YOLOv5-Cls, YOLOv8-Cls, YOLO11-Cls, InternImage, PULC
🎯 Object Detection YOLOv5/6/7/8/9/10, YOLO11/12/26, YOLOX, YOLO-NAS, D-FINE, DAMO-YOLO, Gold_YOLO, RT-DETR, RF-DETR, DEIMv2
🖌️ Instance Segmentation YOLOv5-Seg, YOLOv8-Seg, YOLO11-Seg, YOLO26-Seg, Hyper-YOLO-Seg, RF-DETR-Seg, D-FINE-seg
🏃 Pose Estimation YOLOv8-Pose, YOLO11-Pose, YOLO26-Pose, DWPose, RTMO
😀 Face Estimation SCRFD, YOLOv6Lite-Face
👣 Tracking TrackTrack, Bot-SORT, ByteTrack, SAM2/3-Video
🔄 Rotated Object Detection YOLOv5-Obb, YOLOv8-Obb, YOLO11-Obb, YOLO26-Obb, RT-DETRv2-OBB
📏 Depth Estimation Depth Anything
🧩 Segment Anything SAM 1/2/3, SAM-HQ, SAM-Med2D, EdgeSAM, EfficientViT-SAM, MobileSAM
✂️ Image Matting RMBG 1.4/2.0
💡 Proposal UPN
🏷️ Tagging RAM, RAM++
📄 OCR PP-OCRv4, PP-OCRv5, PP-OCRv6
🧾 Layout Analysis PP-DocLayoutV3
📑 Document Parsing PaddleOCR-VL, PaddleOCR-VL-1.6
🗣️ Vision Foundation Models Rex-Omni, Florence2
👁️ Vision Language Models Qwen3-VL, Gemini, ChatGPT, GLM
🛣️ Lane Detection CLRNet
🔢 Object Counting CountGD, GeCO, GeCo2
📍 Grounding Grounding DINO, YOLO-World, YOLOE, SAM 3, LocateAnything
📚 Other 👉 model_zoo 👈

Docs

  1. Remote Inference Service
  2. Installation & Quickstart
  3. Usage
  4. Command Line Interface
  5. Customize a model
  6. Chatbot
  7. VQA
  8. Image Classifier
  9. Video Classifier
  10. Document Parsing and Intelligent Text Recognition

Examples

Contribute

We believe in open collaboration! X‑AnyLabeling continues to grow with the support of the community. Whether you're fixing bugs, improving documentation, or adding new features, your contributions make a real impact.

To get started, please read our Contributing Guide and make sure to agree to the Contributor License Agreement (CLA) before submitting a pull request.

If you find this project helpful, please consider giving it a ⭐️ star! Have questions or suggestions? Open an issue or email us at cv_hub@163.com.

A huge thank you 🙏 to everyone helping to make X‑AnyLabeling better.

License

This project is licensed under the GNU General Public License v3.0. You may use, modify, and redistribute the software, including for commercial purposes, provided that you comply with the terms of the license.

Sponsor

X-AnyLabeling is an actively maintained open-source project. Your sponsorship helps support feature development, model integration, documentation, and community support.

Sponsor the X-AnyLabeling project

Click the image above to visit the sponsorship page.

Acknowledgement

I extend my heartfelt thanks to the developers and contributors of AnyLabeling, LabelMe, LabelImg, roLabelImg, PPOCRLabel and CVAT, whose work has been crucial to the success of this project.

Citing

If you use this software in your research, please cite it as below:

@misc{X-AnyLabeling,
  year = {2023},
  author = {Wei Wang},
  publisher = {Github},
  organization = {CVHub},
  journal = {Github repository},
  title = {X-AnyLabeling: A Unified Desktop Platform for AI-Assisted Data Annotation},
  howpublished = {\url{https://github.com/CVHub520/X-AnyLabeling}}
}

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

x_anylabeling_cvhub-4.0.5.tar.gz (3.0 MB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

x_anylabeling_cvhub-4.0.5-py3-none-any.whl (3.4 MB view details)

Uploaded Python 3

File details

Details for the file x_anylabeling_cvhub-4.0.5.tar.gz.

File metadata

  • Download URL: x_anylabeling_cvhub-4.0.5.tar.gz
  • Upload date:
  • Size: 3.0 MB
  • Tags: Source
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/6.1.0 CPython/3.13.13

File hashes

Hashes for x_anylabeling_cvhub-4.0.5.tar.gz
Algorithm Hash digest
SHA256 6cb60c62371837c717a3cced91fb3fe44b383a1bb3b85506a42e106fe787fa65
MD5 f5a38f3903d5a498b7912d73398044fc
BLAKE2b-256 57a4fe3a876917a56eee61d8022eba9be00b1a579ff9ea22e8a90f6ead60024c

See more details on using hashes here.

Provenance

The following attestation bundles were made for x_anylabeling_cvhub-4.0.5.tar.gz:

Publisher: release.yml on CVHub520/X-AnyLabeling

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

File details

Details for the file x_anylabeling_cvhub-4.0.5-py3-none-any.whl.

File metadata

File hashes

Hashes for x_anylabeling_cvhub-4.0.5-py3-none-any.whl
Algorithm Hash digest
SHA256 bc1dbf45684872d9367d3c38f0264e18ff1ad01652f37ef5745c3299afe38ace
MD5 0e3fb73b79003e6425b49e646dfb1d1e
BLAKE2b-256 edd6c2bd2044e31e38058396dd5e16bd6b393052a3eca01298b19c301428ec0d

See more details on using hashes here.

Provenance

The following attestation bundles were made for x_anylabeling_cvhub-4.0.5-py3-none-any.whl:

Publisher: release.yml on CVHub520/X-AnyLabeling

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Release history Release notifications | RSS feed

4.0.6

2 files

This release

4.0.5 This release

2 files

4.0.4

2 files

4.0.3

2 files

4.0.2

2 files

4.0.1

2 files

4.0.0

2 files

3.3.10

2 files

3.3.9

2 files

3.3.8

2 files

3.3.7

2 files

3.3.6

2 files

3.3.5

2 files

3.3.4

2 files

3.3.3

2 files

3.3.2

2 files

3.3.1

2 files

3.3.0

2 files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page