Skip to main content
NVIDIA
Explore
Models
Skills
Blueprints
GPUs
Docs
Help Center
Getting Started
  1. Create and verify your account to unlock full access to NVIDIA NIM APIs.
ResourcesDeveloper ForumsContact Support
FAQs
  • Terms of Use
    Privacy Policy
    Your Privacy Choices
    Contact

    Copyright © 2026 NVIDIA Corporation

    11 results for

    Filters

    Use Case
    Publisher
    Audience
    Domain
    NIM Container GPUs
    Library

    Search results

    • NVIDIA
      Downloadable

      nemoretriever-ocr

      Powerful OCR model for fast, accurate real-world image text extraction, layout, and structure analysis.
      Model
      • Table Extraction
      • nemo retriever
      • data ingestion
      • extraction
      • Optical Character Recognition
      14K API calls in the last 30 days
      Last updated on July 24, 2025
    Items per page
    of 1 pages
  • NVIDIA
    Downloadable

    nemotron-ocr-v1

    Powerful OCR model for fast, accurate real-world image text extraction, layout, and structure analysis.
    Model
    • Table Extraction
    • nemo retriever
    • data ingestion
    • extraction
    • Optical Character Recognition
    206K API calls in the last 30 days
    Last updated on March 12, 2026
  • NVIDIA
    Downloadable

    nemotron-ocr-v2

    Nemotron OCR v2 is a state-of-the-art multilingual text recognition model designed for robust end-to-end optical character recognition (OCR) on complex real-world images.
    Model
    • Table Extraction
    • nemo retriever
    • data ingestion
    • extraction
    • Optical Character Recognition
    338K API calls in the last 30 days
    Last updated on June 24, 2026
  • Baidu
    Downloadable

    paddleocr

    Model for table extraction that receives an image as input, runs OCR on the image, and returns the text within the image and its bounding boxes.
    Model
    • Optical Character Recognition
    • Table Extraction
    • Optical Character Detection
    • run-on-rtx
    • extraction
    • nemo retriever
    • data ingestion
    303K API calls in the last 30 days
    Last updated on July 9, 2025
  • NVIDIA
    Downloadable

    nemotron-parse

    Cutting-edge vision-language model exceling in retrieving text and metadata from images.
    Model
    • text and table extraction
    • document parsing
    • supported language - english
    1M API calls in the last 30 days
    Last updated on October 28, 2025
  • NVIDIA
    Downloadable

    nemotron-parse-2.0

    Cutting-edge vision-language model excelling in retrieving text and metadata from images.
    Model
    • text and table extraction
    • document parsing
    • supported language - english
    Last updated on September 11, 2026
  • Download NVIDIA Jetson Linux BSP artifacts (BSP tarball, sample rootfs, public_sources, x-tools, guides) for the active target. Use for Auto Setup; not for extraction or profile edits.
    Skill
    • Developer
    • DevOps Engineer
    • Platform Engineer
    • Hands On Builder
    • Application Developer
    • Jetson
    • Physical AI
    920 downloads in the last 30 days
    Last updated on June 22, 2026
  • Use to call the VIOS REST API (sensor list, timelines, clip extraction, snapshots, add/delete sensors and streams). Not for VLM inference or search.
    Skill
    • Developer
    • DevOps Engineer
    • Platform Engineer
    • Application Developer
    • Solutions Architect
    • Video Search and Summarization (VSS)
    • AI And Machine Learning
    2K downloads in the last 30 days
    Last updated on June 13, 2026
  • Extract per-molecule embeddings from any encoder-bearing KERMT checkpoint. Use a local checkpoint or optionally download a pinned Hugging Face model bundle using HF_TOKEN if configured. Run containerized embedding extraction and write model bundles, per-r
    Skill
    • Bioinformatician
    • Data Scientist
    • Computational Biologist
    • Ml Engineer
    • Research Academic
    • AI And Machine Learning
    • BioNeMo
    Last updated on September 15, 2026
  • CLIP vision-language model for image-text retrieval, zero-shot classification, embedding extraction, ONNX export, and TensorRT deployment. Use when fine-tuning or training CLIP, running zero-shot classification, computing image embeddings, or deploying CL
    Skill
    • AI Engineer
    • Data Scientist
    • Ml Engineer
    • Application Developer
    • AI And Machine Learning
    • TAO Toolkit
    1K downloads in the last 30 days
    Last updated on June 13, 2026
  • SegFormer for semantic segmentation. Lightweight transformer-based architecture with hierarchical feature extraction, efficient for real-time segmentation tasks. Use when training, evaluating, exporting, quantizing, or running inference for a TAO SegForme
    Skill
    • Developer
    • AI Engineer
    • Ml Engineer
    • AI And Machine Learning
    • TAO Toolkit
    1K downloads in the last 30 days
    Last updated on June 13, 2026