PaliGemma 3B

Google's 3B vision-language model for image captioning and OCR on NIM.

Updated 8/30/20260 code examples

About the Model

为什么要选择 PaliGemma?

Google 紧凑型 3B 视觉语言基座模型,在轻量视觉识别、图像描述和简单目标定位上表现优异。

How to Access for Free (via NVIDIA NIM)

No detailed access description available.

Try it in your browser

Pick a model, paste your own free key, and run. Your key is sent once to call the provider and never stored on our servers.

Pick a model, paste your own free key, and run. Your key is sent once to call the provider and never stored on our servers.

Code Examples

No examples provided yet.

Community Feedback

Comments are tied to your GitHub account — sign in to join the discussion.