Tech Product

PaliGemma

Overview

最終更新: 2026年8月14日

PaliGemma is a specialized vision-language model (VLM) within the Gemma family. It is designed for tasks that require understanding both images and text, such as image captioning, visual question answering, and object detection. It serves as a pre-trained foundation that developers can fine-tune for specific computer vision use cases.

Mentioned Articles

1 件

External Mentions

3 件