Ready to use Computer Vision algorithms and Apps.
Generalized-scale instance segmentation with GECO2 from bbox exemplars.
Run LocateAnything-3B for open-vocabulary object detection
Unlimited-OCR document OCR to Markdown
Inference with RF-DETR segmentation models
Classify document image orientation as 0, 90, 180, or 270 degrees
Train RF-DETR instance segmentation models
Inference with RF-DETR models
Train RF-DETR models
Inference with YOLOv11 models
Inference with YOLOv11 pose estimation models
Inference with YOLOv11 segmentation models
Inference for MMDET from OneDL MMDetection models
Inference for pose estimation models from mmpose
Inference for MMOCR from MMLAB text detection models
Train for MMOCR from MMLAB KIE models
Training process for MMOCR from MMLAB in text detection
Train MMDetection models
Inference for MMLAB segmentation models
Inference for MMOCR from MMLAB text recognition models
Inference for MMOCR from MMLAB KIE models
Train for MMLAB segmentation models
Training process for MMOCR from MMLAB in text recognition
Run florence 2 segmentation with or without text prompt
Run florence 2 object detection with or without text prompt
Inference for text recognition (OCR) with Florence-2
Image captioning with Florence-2
Inference with D-FINE models
Train D-FINE models
Inference with YOLO26 models (Ultralytics)
Inference with YOLO26 pose estimation models