Image Classifier
The X-AnyLabeling Image Classifier is a dedicated annotation window for multi-class (single-label) and multi-label classification. It supports label management, AI-assisted classification for individual images or batches, dataset statistics, keyboard navigation, and category-based image export.
PaddleOCR
PaddleOCR is an OCR and document intelligence toolkit in the Baidu PaddlePaddle ecosystem. It covers general text recognition, document layout analysis, table parsing, formula recognition, and other capabilities for common document-processing scenarios such as scanned documents, photographed documents, multi-page PDFs, and technical documents.
Video Classifier
The Video Classifier in X-AnyLabeling is a dedicated annotation window for action recognition and video clip classification datasets. It lets you load videos, define class labels, mark time segments on a timeline, assign each segment to a label, preview frames, and export class-organized video clips or raw frame sequences.
Visual Question Answering
The X-AnyLabeling Visual Question Answering (VQA) tool annotates multimodal image-question answering datasets. It supports image-based question-and-answer pairs, configurable input components, and AI assistance for preparing structured data for supervised fine-tuning, reinforcement-learning post-training, and similar tasks.
Chatbot
The X-AnyLabeling Chatbot is an AI assistant integrated into the annotation workflow. It supports natural-language conversations, batch image-question answering, and importing or exporting single-turn and multi-turn multimodal data in ShareGPT format for frameworks such as LLaMA-Factory.