Robust image classification model for detecting and managing AI-generated content.
NV-DINOv2 is a visual foundation model that generates vector embeddings for the input image.
Grounding dino is an open vocabulary zero-shot object detection model.
OCDNet and OCRNet are pre-trained models designed for optical character detection and recognition respectively.
EfficientDet-based object detection network to detect 100 specific retail objects from an input video.