Skip to main content

X-AnyLabeling interface

🥳 What's New

  • 2026-08-19: Add support for image tagging, with tag creation, editing, reordering, and batch deletion.
  • 2026-08-12: Add support for D-FINE-seg instance segmentation models.
  • 2026-08-08: Add support for the RT-DETRv2-OBB rotated object detection model.
  • 2026-08-08: Add the Magic Wand tool for quickly creating polygons from contiguous color regions.
  • 2026-08-05: Release X-AnyLabeling v4.0.0.
  • For more details, please refer to the CHANGELOG

Introduction

X-AnyLabeling is a lightweight, efficient, and unified cross-platform desktop application for AI-assisted annotation of text, image, video, and multimodal data. It combines versatile built-in tools, automated labeling workflows, state-of-the-art deep learning models, and flexible multi-format import and export. For remote inference, X-AnyLabeling-Server provides a lightweight, extensible backend for connecting custom models and compute resources.

Key Features

  • Unified support for annotating and processing text, image, video, and multimodal data.
  • Covers tasks such as image classification, object detection, instance segmentation, pose estimation, oriented object detection, multi-object tracking, optical character recognition, lane annotation, image captioning, visual question answering, and document parsing.
  • Provides polygons, rectangles, cuboids, rotated boxes, quadrilaterals, circles, lines, polylines, points, masks, and task-specific tools for text detection, text recognition, and KIE.
  • Integrates a wide range of state-of-the-art deep learning models for AI-assisted annotation, automated labeling, and batch dataset prediction.
  • Supports both local and remote inference through engines and serving frameworks such as ONNX Runtime, TensorRT, OpenCV DNN, vLLM, and SGLang.
  • Supports importing and exporting formats such as COCO, VOC, YOLO, DOTA, MOT, MASK, PPOCR, MMGD, VLM-R1, and ShareGPT.
  • Runs on Windows, Linux, and macOS, with interfaces available in English, Simplified Chinese, Japanese, and Korean.
  • Supports custom model integration, flexible extension, and secondary development.

Model library

Task Category Supported Models
🖼️ Image Classification YOLOv5-Cls, YOLOv8-Cls, YOLO11-Cls, InternImage, PULC
🎯 Object Detection YOLOv5/6/7/8/9/10, YOLO11/12/26, YOLOX, YOLO-NAS, D-FINE, DAMO-YOLO, Gold_YOLO, RT-DETR, RF-DETR, DEIMv2
🖌️ Instance Segmentation YOLOv5-Seg, YOLOv8-Seg, YOLO11-Seg, YOLO26-Seg, Hyper-YOLO-Seg, RF-DETR-Seg, D-FINE-seg
🏃 Pose Estimation YOLOv8-Pose, YOLO11-Pose, YOLO26-Pose, DWPose, RTMO
😀 Face Estimation SCRFD, YOLOv6Lite-Face
👣 Tracking TrackTrack, Bot-SORT, ByteTrack, SAM2/3-Video
🔄 Rotated Object Detection YOLOv5-Obb, YOLOv8-Obb, YOLO11-Obb, YOLO26-Obb, RT-DETRv2-OBB
📏 Depth Estimation Depth Anything
🧩 Segment Anything SAM 1/2/3, SAM-HQ, SAM-Med2D, EdgeSAM, EfficientViT-SAM, MobileSAM
✂️ Image Matting RMBG 1.4/2.0
💡 Proposal UPN
🏷️ Tagging RAM, RAM++
📄 OCR PP-OCRv4, PP-OCRv5, PP-OCRv6
🧾 Layout Analysis PP-DocLayoutV3
📑 Document Parsing PaddleOCR-VL, PaddleOCR-VL-1.6
🗣️ Vision Foundation Models Rex-Omni, Florence2
👁️ Vision Language Models Qwen3-VL, Gemini, ChatGPT, GLM
🛣️ Lane Detection CLRNet
🔢 Object Counting CountGD, GeCO, GeCo2
📍 Grounding Grounding DINO, YOLO-World, YOLOE, SAM 3, LocateAnything
📚 Other 👉 model_zoo 👈

Docs

  1. Remote Inference Service
  2. Installation & Quickstart
  3. Usage
  4. Command Line Interface
  5. Customize a model
  6. Chatbot
  7. VQA
  8. Image Classifier
  9. Video Classifier
  10. Document Parsing and Intelligent Text Recognition

Examples

Contribute

We believe in open collaboration! X‑AnyLabeling continues to grow with the support of the community. Whether you're fixing bugs, improving documentation, or adding new features, your contributions make a real impact.

To get started, please read our Contributing Guide and make sure to agree to the Contributor License Agreement (CLA) before submitting a pull request.

If you find this project helpful, please consider giving it a ⭐️ star! Have questions or suggestions? Open an issue or email us at cv_hub@163.com.

A huge thank you 🙏 to everyone helping to make X‑AnyLabeling better.

License

This project is licensed under the GNU General Public License v3.0. You may use, modify, and redistribute the software, including for commercial purposes, provided that you comply with the terms of the license.

Sponsor

X-AnyLabeling is an actively maintained open-source project. Your sponsorship helps support feature development, model integration, documentation, and community support.

Sponsor the X-AnyLabeling project

Click the image above to visit the sponsorship page.

Acknowledgement

I extend my heartfelt thanks to the developers and contributors of AnyLabeling, LabelMe, LabelImg, roLabelImg, PPOCRLabel and CVAT, whose work has been crucial to the success of this project.

Citing

If you use this software in your research, please cite it as below:

@misc{X-AnyLabeling,
  year = {2023},
  author = {Wei Wang},
  publisher = {Github},
  organization = {CVHub},
  journal = {Github repository},
  title = {X-AnyLabeling: A Unified Desktop Platform for AI-Assisted Data Annotation},
  howpublished = {\url{https://github.com/CVHub520/X-AnyLabeling}}
}

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

x_anylabeling_cvhub-4.0.4.tar.gz (3.0 MB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

x_anylabeling_cvhub-4.0.4-py3-none-any.whl (3.4 MB view details)

Uploaded Python 3

File details

Details for the file x_anylabeling_cvhub-4.0.4.tar.gz.

File metadata

  • Download URL: x_anylabeling_cvhub-4.0.4.tar.gz
  • Upload date:
  • Size: 3.0 MB
  • Tags: Source
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/6.1.0 CPython/3.13.13

File hashes

Hashes for x_anylabeling_cvhub-4.0.4.tar.gz
Algorithm Hash digest
SHA256 10464b73d64c67ec841307e28549545051c0488f29c17967da1dc33fb95c170c
MD5 e3d1cb211b5eee1df733147d1a4ab74e
BLAKE2b-256 3e0d46d9780a10f34641e976a26b702a2857c5fa4218cd7fe6111caf9bce0133

See more details on using hashes here.

Provenance

The following attestation bundles were made for x_anylabeling_cvhub-4.0.4.tar.gz:

Publisher: release.yml on CVHub520/X-AnyLabeling

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

File details

Details for the file x_anylabeling_cvhub-4.0.4-py3-none-any.whl.

File metadata

File hashes

Hashes for x_anylabeling_cvhub-4.0.4-py3-none-any.whl
Algorithm Hash digest
SHA256 423d9c968b546eb5a257bc76dd12cab23c6b99f687c01b7904527a3214229bdc
MD5 c7ed78b940c8190d356c612b65aa02bb
BLAKE2b-256 3dddea4319f523e562ba35296d672f7757222a2168aea3ace6dca7d69fb65c6b

See more details on using hashes here.

Provenance

The following attestation bundles were made for x_anylabeling_cvhub-4.0.4-py3-none-any.whl:

Publisher: release.yml on CVHub520/X-AnyLabeling

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Release history Release notifications | RSS feed

4.0.6

2 files

4.0.5

2 files

This release

4.0.4 This release

2 files

4.0.3

2 files

4.0.2

2 files

4.0.1

2 files

4.0.0

2 files

3.3.10

2 files

3.3.9

2 files

3.3.8

2 files

3.3.7

2 files

3.3.6

2 files

3.3.5

2 files

3.3.4

2 files

3.3.3

2 files

3.3.2

2 files

3.3.1

2 files

3.3.0

2 files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page