Google Cloud Vision logo
ConnectorGoogle Cloud Vision

Google Cloud Vision integration for Claude and Codex

Create a Google Cloud Vision connection, then control which Spaces can use its approved tools and data without pasting credentials into prompts or threads.

Use Google Cloud Vision from Claude

Add the Google Cloud Vision connection to a Space that runs Claude, then use it from every channel in that Space.

Use Google Cloud Vision from Codex

Add the Google Cloud Vision connection to a Space that runs Codex, then use it from every channel in that Space.

Use Google Cloud Vision from Sidekick

Connect your own OpenClaw to Sidekick while Type provides the conversations, permissions, connections, skills, and automations.

Design & Media

What the Google Cloud Vision connector can do

Google Cloud Vision API enables developers to integrate vision detection features into applications, including image labeling, face and landmark detection, optical character recognition (OCR), and explicit content tagging.

One connection, many Spaces

Connect Google Cloud Vision once, then decide which Spaces can use it in threads, skills, automations, and coding work.

Representative actions

  • Annotate Files with Vision API

    Tool to perform image detection and annotation for batch files in Google Cloud Vision. Supports PDF, TIFF, and GIF files. Extracts up to 5 frames (GIF) or pages (PDF/TIFF) from each file and performs detection for each image. Use when you need to analyze documents or multi-page images with features like text detection, label detection, face detection, or other Vision API capabilities.

  • Async Batch Annotate Files

    Tool to run asynchronous image detection and annotation for a list of generic files (PDF, TIFF, GIF). Use when processing multi-page documents that may contain multiple images per page. Results are written to Google Cloud Storage and progress can be tracked via the returned operation name using VisionGetOperation.

  • Annotate Images

    Run image detection and annotation for a batch of images using Google Cloud Vision API. Performs various types of image analysis including face detection, landmark detection, logo detection, label detection, text detection (OCR), safe search detection, image properties, crop hints, web detection, product search, and object localization. Supports up to 16 images in a single batch request. Each image can have multiple feature types analyzed simultaneously.

  • Annotate Images Async Batch

    Tool to run asynchronous image detection and annotation for a batch of images. Use when processing multiple images or large images that require longer processing time. Results are written to Google Cloud Storage as JSON files.

  • Annotate Location Images

    Tool to run image detection and annotation for a batch of images scoped to a specific project and location. Performs various types of image analysis including label detection, face detection, landmark detection, logo detection, OCR text detection, safe search detection, image properties, crop hints, web detection, product search, and object localization. Supports processing up to 16 images per request with regional endpoint routing (us, asia, eu). Use this when you need to analyze images with location-specific processing for content extraction, text recognition, object detection, face identification, or landmark/logo recognition.

Connection

API and auth details

Google Cloud Vision API analyzes images with pretrained ML features. Integrations can perform label, text/OCR, document text, face, landmark, logo, object localization, crop hint, image property, and safe-search detection, using Google Cloud API credentials and project-level IAM or API-key configuration as appropriate.

FAQ

Questions people ask before connecting Google Cloud Vision

Can Claude use Google Cloud Vision?

Yes. Add a Google Cloud Vision connection to a Space that runs Claude, then use its approved tools and data from any channel in that Space.

Can Codex work with Google Cloud Vision through Type?

Yes. Add a Google Cloud Vision connection to a Space that runs Codex, then use its approved tools and data from any channel in that Space.

Is this the same as a Google Cloud Vision MCP server?

Not exactly. Type uses connectors and connections to give selected Spaces access to approved app tools and data. Some connectors use hosted MCP, while others use OAuth, API keys, service accounts, or custom APIs.

More design & media connectors