Google Cloud Vision logo
ConnectorGoogle Cloud Vision

Google Cloud Vision integration for Claude and Codex

Create a Google Cloud Vision connection, then control which Spaces can use its approved tools and data without pasting credentials into prompts or threads.

Use Google Cloud Vision from Claude

Add the Google Cloud Vision connection to a Space that runs Claude, then use it from every channel in that Space.

Use Google Cloud Vision from Codex

Add the Google Cloud Vision connection to a Space that runs Codex, then use it from every channel in that Space.

Use Google Cloud Vision from Sidekick

Connect your own OpenClaw to Sidekick while type.com provides the conversations, permissions, connections, skills, and automations.

Design & Media

What the Google Cloud Vision connector can do

Google Cloud Vision reads images using machine learning, covering text extraction, labeling, face and logo detection, and safe-search flags. In type.com, a content or operations Space can ask the agent to pull text out of a scanned document, label a batch of images, or check uploaded photos for inappropriate content, from one thread.

One connection, many Spaces

Connect Google Cloud Vision once, then decide which Spaces can use it in threads, skills, automations, and coding work.

What teams do with Google Cloud Vision in type.com

Extract text from a scanned document

Run OCR on a PDF or image file so its text becomes searchable and usable without retyping it by hand.

Try asking
Run Google Cloud Vision OCR on this scanned invoice and give me the extracted text.

Label a batch of product photos

Annotate a set of images for labels and objects, useful for organizing a large photo library by content.

Try asking
Annotate these 20 product photos with Google Cloud Vision and tag them by detected objects.

Screen uploads for inappropriate content

Check a batch of user-submitted images for safe-search flags before they go live on a public page.

Try asking
Run safe-search detection on these uploaded community photos with Google Cloud Vision.

Representative actions

  • Annotate Files with Vision API

    Detects and annotates content in batch files such as PDF, TIFF, or GIF, extracting up to five frames or pages per file.

  • Async Batch Annotate Files

    Runs asynchronous detection on a list of multi-page files, writing results to Google Cloud Storage with progress tracking.

  • Annotate Images

    Runs image analysis such as face, landmark, logo, label, and text detection on a batch of images.

  • Annotate Images Async Batch

    Runs asynchronous detection on a large batch of images, writing results to Google Cloud Storage as JSON files.

  • Annotate Location Images

    Runs image analysis such as label, face, landmark, logo, and OCR detection scoped to a specific project and location.

Connection

API and auth details

Google Cloud Vision API analyzes images with pretrained ML features. Integrations can perform label, text/OCR, document text, face, landmark, logo, object localization, crop hint, image property, and safe-search detection, using Google Cloud API credentials and project-level IAM or API-key configuration as appropriate.

FAQ

Questions people ask before connecting Google Cloud Vision

Can Claude analyze images with Google Cloud Vision in type.com?

Yes. With a Google Cloud Vision connection shared to a Space, the agent can run OCR, labeling, or safe-search checks on images and post the results in a thread.

How does type.com authenticate with Google Cloud Vision?

Google Cloud Vision uses an API key, added once by an admin when the connection is created, and shared only with Spaces that need it.

Can image annotation jobs run on a recurring schedule?

Yes. A recurring batch, such as labeling newly uploaded photos each day, can be set up as a type.com automation.

Is this the same as a Google Cloud Vision MCP server?

Not exactly. type.com uses connectors and connections to give selected Spaces access to approved app tools and data. Some connectors use hosted MCP, while others use OAuth, API keys, service accounts, or custom APIs.

More design & media connectors

All Design & Media integrations