Google Cloud Vision API enables developers to integrate vision detection features into applications, including image labeling, face and landmark detection, optical character recognition (OCR), and explicit content tagging.
Every capability is a discrete, logged action the agent calls by name — scoped to what you authorize and recorded in the run trace.