Generates consistent characters across a series of images without requiring additional training.
Grounding dino is an open vocabulary zero-shot object detection model.
Creates diverse synthetic data that mimics the characteristics of real-world data.
OCDNet and OCRNet are pre-trained models designed for optical character detection and recognition respectively.
EfficientDet-based object detection network to detect 100 specific retail objects from an input video.