# Google Vision AI

> Google Cloud's image recognition and OCR API

- Category: [Multimodal & Perception](https://ailandscape.org/category/multimodal-perception) › Computer Vision
- Homepage: https://cloud.google.com/vision
- Tags: ocr, vision, api
- Added to the landscape: 2026-03-18

## Similar tools in Computer Vision

- [Detectron2](https://ailandscape.org/tool/detectron2): Meta's object detection and segmentation framework
- [DETR](https://ailandscape.org/tool/detr): End-to-end object detection with transformers by Meta
- [PaddleOCR](https://ailandscape.org/tool/paddleocr): Multilingual OCR toolkits based on PaddlePaddle
- [SAM](https://ailandscape.org/tool/sam): Segment Anything Model by Meta
- [SAM 2](https://ailandscape.org/tool/sam-2): Meta's model for image and video segmentation
- [Tesseract](https://ailandscape.org/tool/tesseract): Open-source OCR engine
- [YOLO](https://ailandscape.org/tool/yolo): Real-time object detection by Ultralytics

---

Part of [AI Landscape](https://ailandscape.org), an open map of the AI ecosystem. Web page: https://ailandscape.org/tool/google-vision-ai · Index for AI assistants: https://ailandscape.org/llms.txt
