Skip to main content
Explore
Models
Skills
Blueprints
GPUs
Docs
Search
⌘K
Ctrl+K
?
Forums
Support
Login
2 results for
Filters (2)
Models (2)
Blueprints (0)
Skills (0)
Other (0)
Sort By
Best Match
Select item
Best Match
Most Popular
Most Downloaded
Alphabetical (A-Z)
Alphabetical (Z-A)
Best Match
Most Popular
Most Downloaded
Alphabetical (A-Z)
Alphabetical (Z-A)
Google
Free Endpoint
paligemma
Vision language model adept at comprehending text and visual inputs to produce informative responses
Model
image
+8
cv
Vision Assistant
vlm
Visual Question Answering
computer vision
Language Generation
video
Image-to-Text
Items per page
24
12
24
48
96
1
1
of 1 pages
12K
12K API calls in the last 30 days
1y
Last updated on August 26, 2024
NVIDIA
Downloadable
Free Endpoint
nemotron-nano-12b-v2-vl
Nemotron Nano 12B v2 VL enables multi-image and video understanding, along with visual Q&A and summarization capabilities.
Model
language generation
+3
vision assistant
visual question answering
Image-to-Text
5M
5M API calls in the last 30 days
9mo
Last updated on October 28, 2025