Skip to main content
NVIDIA
Explore
Models
Skills
Blueprints
GPUs
Docs
Help Center
Getting Started
  1. Create and verify your account to unlock full access to NVIDIA NIM APIs.
ResourcesDeveloper ForumsContact Support
FAQs
  • DiscoverModelsSkillsBlueprintsGPUsDocsForums

    workstations

    • Run on RTX
    • Run on Spark
    • Run on Station

    models

    • Reasoning
    • Vision
    • Visual Design
    • Retrieval
    • Speech
    • Biology
    • Simulation
    • Climate & Weather
    • Safety & Moderation

    industries

    • Automotive
    • Financial Services
    • Gaming
    • Healthcare
    • Industrial
    • Robotics

    Financial Services

    Terms of Use
    Privacy Policy
    Your Privacy Choices
    Contact

    Copyright © 2026 NVIDIA Corporation

    Deploy Models Now with NVIDIA NIM

    Optimized inference for the world’s leading models
    Free serverless APIs for development
    Accelerated by DGX Cloud
    Self-Host on your GPU infrastructure
    Continuous vulnerability fixes

    Accelerate Financial Workflows With NVIDIA Technology

    Developer examples designed for quick-start AI development in financial services, including artifacts like Docker containers and Jupyter Notebooks, allowing for fast deployment with tools like Docker compose and Brev Launchable.

    nvidiaQuantitative Signal Discovery Agent

    Automate and scale the discovery, testing, and refinement of trading signals for quantitative research.

    • NeMo Agent Toolkit
    • Nemotron
    • algorithmic trading

    nvidiaBuild Your Own Transaction Foundation Model

    Create intelligent embeddings by using transformer architecture on tabular data.

    • banking
    • capital markets
    • financial services
    • fraud
    • nemo
    • payments
    • personalization
    • transformers

    nvidiaQuantitative Portfolio Optimization

    Enable fast, scalable, and real-time portfolio optimization for financial institutions.

    • algorithmic trading
    • cuopt
    • developer example
    • financial services
    • portfolio optimization

    nvidiaAI Model Distillation for Financial Data

    Distill and deploy domain-specific AI models from unstructured financial data to generate market signals efficiently—scaling your workflow with the NVIDIA Data Flywheel Blueprint for high-performance, cost-efficient experimentation.

    • Nemotron
    • algorithmic trading
    • data flywheel
    • developer example
    • financial services
    • launchable
    • llm
    • nim
    • nvidia ai

    nvidiaFinancial Fraud Detection

    Detect and prevent sophisticated fraudulent activities for financial services with high accuracy.

    • Financial Services
    • Fraud Detection
    • GNN
    • Payments

    Explore NVIDIA Blueprints

    Comprehensive reference workflows that accelerate application development and deployment, featuring NVIDIA acceleration libraries, APIs, and microservices for AI agents, digital twins, and more.

    Enterprise

    nvidiaNVIDIA AI-Q Blueprint for intelligent agents

    AI agents that connect, retrieve, and reason on enterprise data—making information accessible, actionable, and intelligent.

    • Agents
    • Enterprise
    • NIM
    • NeMo
    • Nemotron

    nvidiaBuild Continuous Refining AI Agents with Data Flywheels

    Build a production-integrated data flywheel that continuously optimizes AI agents for latency, accuracy, and cost — built with NVIDIA NeMo and open NVIDIA Nemotron models.

    • Data Flywheel
    • NIM
    • NeMo microservices
    • enterprise
    • nemotron
    • nvidia ai

    nvidiaBuild an Enterprise RAG Pipeline Blueprint

    Power fast, accurate semantic search across multimodal enterprise data with NVIDIA’s RAG Blueprint—built on NeMo Retriever and Nemotron models—to connect your agents to trusted, authoritative sources of knowledge.

    • NIM
    • NeMo Retriever
    • Nemotron
    • Retrieval-Augmented Generation

    nvidiaBuild a Video Search and Summarization (VSS) Agent

    Ingest massive volumes of live or archived videos and extract insights for summarization and interactive Q&A

    • chat
    • generative AI
    • video-to-text
    • vision

    Chat With Your Industry Domain Expertise

    Leverage retrieval-augmented generation to ground large language models in your proprietary data.

    • NVIDIA
      DownloadableFree Endpoint

      nvidia-nemotron-nano-9b-v2

      High‑efficiency LLM with hybrid Transformer‑Mamba design, excelling in reasoning and agentic tasks.
      Model
      • reasoning
      • thinking budget
      2M API calls in the last 30 days
      Last updated on August 18, 2025
    • NVIDIA
      Downloadable

      cuopt

      World-record accuracy and performance for complex route optimization.
      Model
      • NVIDIA
        Downloadable

        parakeet-1.1b-rnnt-multilingual-asr

        High accuracy and optimized performance for transcription in 25 languages
        Model
        • Automatic Speech Recognition
        • NVIDIA NIM
        • NVIDIA Riva
      • OpenAI
        DownloadableFree Endpoint

        gpt-oss-120b

        Mixture of Experts (MoE) reasoning LLM (text-only) designed to fit within 80GB GPU.
        Model
        • chat
        • math
        • reasoning
        • text-to-text
      59K API calls in the last 30 days
      Last updated on May 21, 2025
      66K API calls in the last 30 days
      Last updated on April 30, 2025
      45M API calls in the last 30 days
      Last updated on August 5, 2025