🥳 What's New
2026-08-05: Release X-AnyLabeling v4.0.0.- For more details, please refer to the CHANGELOG
Introduction
X-AnyLabeling is a lightweight, efficient, and unified cross-platform desktop application for AI-assisted annotation of text, image, video, and multimodal data. It combines versatile built-in tools, automated labeling workflows, state-of-the-art deep learning models, and flexible multi-format import and export. For remote inference, X-AnyLabeling-Server provides a lightweight, extensible backend for connecting custom models and compute resources.
Key Features
- Unified support for annotating and processing text, image, video, and multimodal data.
- Covers tasks such as image classification, object detection, instance segmentation, pose estimation, oriented object detection, multi-object tracking, optical character recognition, lane annotation, image captioning, visual question answering, and document parsing.
- Provides polygons, rectangles, cuboids, rotated boxes, quadrilaterals, circles, lines, polylines, points, masks, and task-specific tools for text detection, text recognition, and KIE.
- Integrates a wide range of state-of-the-art deep learning models for AI-assisted annotation, automated labeling, and batch dataset prediction.
- Supports both local and remote inference through engines and serving frameworks such as
ONNX Runtime,TensorRT,OpenCV DNN,vLLM, andSGLang. - Supports importing and exporting formats such as
COCO,VOC,YOLO,DOTA,MOT,MASK,PPOCR,MMGD,VLM-R1, andShareGPT. - Runs on Windows, Linux, and macOS, with interfaces available in English, Simplified Chinese, Japanese, and Korean.
- Supports custom model integration, flexible extension, and secondary development.
Model library
| Task Category | Supported Models |
|---|---|
| 🖼️ Image Classification | YOLOv5-Cls, YOLOv8-Cls, YOLO11-Cls, InternImage, PULC |
| 🎯 Object Detection | YOLOv5/6/7/8/9/10, YOLO11/12/26, YOLOX, YOLO-NAS, D-FINE, DAMO-YOLO, Gold_YOLO, RT-DETR, RF-DETR, DEIMv2 |
| 🖌️ Instance Segmentation | YOLOv5-Seg, YOLOv8-Seg, YOLO11-Seg, YOLO26-Seg, Hyper-YOLO-Seg, RF-DETR-Seg |
| 🏃 Pose Estimation | YOLOv8-Pose, YOLO11-Pose, YOLO26-Pose, DWPose, RTMO |
| 😀 Face Estimation | SCRFD, YOLOv6Lite-Face |
| 👣 Tracking | TrackTrack, Bot-SORT, ByteTrack, SAM2/3-Video |
| 🔄 Rotated Object Detection | YOLOv5-Obb, YOLOv8-Obb, YOLO11-Obb, YOLO26-Obb |
| 📏 Depth Estimation | Depth Anything |
| 🧩 Segment Anything | SAM 1/2/3, SAM-HQ, SAM-Med2D, EdgeSAM, EfficientViT-SAM, MobileSAM |
| ✂️ Image Matting | RMBG 1.4/2.0 |
| 💡 Proposal | UPN |
| 🏷️ Tagging | RAM, RAM++ |
| 📄 OCR | PP-OCRv4, PP-OCRv5, PP-OCRv6 |
| 🧾 Layout Analysis | PP-DocLayoutV3 |
| 📑 Document Parsing | PaddleOCR-VL, PaddleOCR-VL-1.6 |
| 🗣️ Vision Foundation Models | Rex-Omni, Florence2 |
| 👁️ Vision Language Models | Qwen3-VL, Gemini, ChatGPT, GLM |
| 🛣️ Lane Detection | CLRNet |
| 🔢 Object Counting | CountGD, GeCO, GeCo2 |
| 📍 Grounding | Grounding DINO, YOLO-World, YOLOE, SAM 3, LocateAnything |
| 📚 Other | 👉 model_zoo 👈 |
Docs
- Remote Inference Service
- Installation & Quickstart
- Usage
- Command Line Interface
- Customize a model
- Chatbot
- VQA
- Image Classifier
- Video Classifier
- Document Parsing and Intelligent Text Recognition
Examples
- Classification
- Detection
- Segmentation
- Description
- Estimation
- OCR
- MOT
- iVOS
- Matting
- Vision-Language
- Counting
- Grounding
- Training
Contribute
We believe in open collaboration! X‑AnyLabeling continues to grow with the support of the community. Whether you're fixing bugs, improving documentation, or adding new features, your contributions make a real impact.
To get started, please read our Contributing Guide and make sure to agree to the Contributor License Agreement (CLA) before submitting a pull request.
If you find this project helpful, please consider giving it a ⭐️ star! Have questions or suggestions? Open an issue or email us at cv_hub@163.com.
A huge thank you 🙏 to everyone helping to make X‑AnyLabeling better.
License
This project is licensed under the GNU General Public License v3.0. You may use, modify, and redistribute the software, including for commercial purposes, provided that you comply with the terms of the license.
Sponsor
X-AnyLabeling is an actively maintained open-source project. Your sponsorship helps support feature development, model integration, documentation, and community support.
Click the image above to visit the sponsorship page.
Acknowledgement
I extend my heartfelt thanks to the developers and contributors of AnyLabeling, LabelMe, LabelImg, roLabelImg, PPOCRLabel and CVAT, whose work has been crucial to the success of this project.
Citing
If you use this software in your research, please cite it as below:
@misc{X-AnyLabeling,
year = {2023},
author = {Wei Wang},
publisher = {Github},
organization = {CVHub},
journal = {Github repository},
title = {X-AnyLabeling: A Unified Desktop Platform for AI-Assisted Data Annotation},
howpublished = {\url{https://github.com/CVHub520/X-AnyLabeling}}
}
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
Built Distribution
Filter files by name, interpreter, ABI, and platform.
If you're not sure about the file name format, learn more about wheel file names.
Copy a direct link to the current filters
File details
Details for the file x_anylabeling_cvhub-4.0.0.tar.gz.
File metadata
- Download URL: x_anylabeling_cvhub-4.0.0.tar.gz
- Upload date:
- Size: 2.9 MB
- Tags: Source
- Uploaded using Trusted Publishing? Yes
- Uploaded via: twine/6.1.0 CPython/3.13.13
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
7db6ad9345d101c7565fb5a0dc43d4f001d83b649d093f5b8b9e09f8a9343915
|
|
| MD5 |
07cdb6aade840b45b6cca5ee90f8e266
|
|
| BLAKE2b-256 |
4989daf7be4209c79ff43bf5b4a8598f163d303d341f157d736f1b6d8f7d9aec
|
Provenance
The following attestation bundles were made for x_anylabeling_cvhub-4.0.0.tar.gz:
Publisher:
release.yml on CVHub520/X-AnyLabeling
-
Statement:
-
Statement type:
https://in-toto.io/Statement/v1 -
Predicate type:
https://docs.pypi.org/attestations/publish/v1 -
Subject name:
x_anylabeling_cvhub-4.0.0.tar.gz -
Subject digest:
7db6ad9345d101c7565fb5a0dc43d4f001d83b649d093f5b8b9e09f8a9343915 - Sigstore transparency entry: 2340947804
- Sigstore integration time:
-
Permalink:
CVHub520/X-AnyLabeling@95fc0f9859d20cfa99082f335318cff1fd001dd8 -
Branch / Tag:
refs/tags/v4.0.0 - Owner: https://github.com/CVHub520
-
Access:
public
-
Token Issuer:
https://token.actions.githubusercontent.com -
Runner Environment:
github-hosted -
Publication workflow:
release.yml@95fc0f9859d20cfa99082f335318cff1fd001dd8 -
Trigger Event:
push
-
Statement type:
File details
Details for the file x_anylabeling_cvhub-4.0.0-py3-none-any.whl.
File metadata
- Download URL: x_anylabeling_cvhub-4.0.0-py3-none-any.whl
- Upload date:
- Size: 3.3 MB
- Tags: Python 3
- Uploaded using Trusted Publishing? Yes
- Uploaded via: twine/6.1.0 CPython/3.13.13
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
3990844a032cd06040f39797e3f13a0f7e0eeebd88e636cbd106c032f5a9ad33
|
|
| MD5 |
e48618c1d1bb522f7c5e6240f136e032
|
|
| BLAKE2b-256 |
055ec462c95150d172a8ff3cdc59ab7590a55c040bcc22fc89c6c7bea0a7e998
|
Provenance
The following attestation bundles were made for x_anylabeling_cvhub-4.0.0-py3-none-any.whl:
Publisher:
release.yml on CVHub520/X-AnyLabeling
-
Statement:
-
Statement type:
https://in-toto.io/Statement/v1 -
Predicate type:
https://docs.pypi.org/attestations/publish/v1 -
Subject name:
x_anylabeling_cvhub-4.0.0-py3-none-any.whl -
Subject digest:
3990844a032cd06040f39797e3f13a0f7e0eeebd88e636cbd106c032f5a9ad33 - Sigstore transparency entry: 2340947812
- Sigstore integration time:
-
Permalink:
CVHub520/X-AnyLabeling@95fc0f9859d20cfa99082f335318cff1fd001dd8 -
Branch / Tag:
refs/tags/v4.0.0 - Owner: https://github.com/CVHub520
-
Access:
public
-
Token Issuer:
https://token.actions.githubusercontent.com -
Runner Environment:
github-hosted -
Publication workflow:
release.yml@95fc0f9859d20cfa99082f335318cff1fd001dd8 -
Trigger Event:
push
-
Statement type: