Keyflow

Search Keyflow

Find a shortcut, workflow, MCP server or skill

Prompt-based Object Detection with Gemini 2.0

n8n template #2649
This n8n template demonstrates how to get started with Gemini 2.0's new Bounding Box detection capabilities in your workflows. The key difference being this enables prompt-based object detection for images which is pretty powerful for things like contextual search over an image. eg. "Put a bounding box around all adults with children in this image" or "Put a bounding box around cars parked out of bounds of a parking space". How it works An image is downloaded via the HTTP node and an "Edit Image" node is used to extract the file's width and height. The image is then given to the Gemini 2.0 API to parse and return coordinates of the bounding box of the requested subjects. In this demo, we've

Hand off to your agent

Prompt

Help me set up the n8n workflow "Prompt-based Object Detection with Gemini 2.0" (https://n8n.io/workflows/2649). It uses: Edit Image, HTTP Request, Code. Import the template JSON into my n8n instance, list every credential I need to create, and walk me through testing it.

Paste into Claude Code and it will do the rest.

Apps and nodes

Similar workflows

NameViews
  1. Transcribing Bank Statements To Markdown Using Gemini Vision AI15K
  2. Easy Image Captioning with Gemini 1.5 Pro12K
  3. Narrating over a Video using Multimodal AI8,840
  4. Scale Deal Flow with a Pitch Deck AI Vision, Chatbot and QDrant Vector Store6,822
  5. Overlay or Watermark Images by Merging with Another Image3,881
  6. WordPress blog automation with Airtable interface, human review & AI research v2174
  7. Automate SEO-Optimized blog creation with GPT-4o, Perplexity AI & multi-language support166
  8. Generate AI Viral Videos with Seedance and Upload to TikTok, YouTube & Instagram215K
Buy me a coffee