CV Resume PDF Parsing with Multimodal Vision AI
n8n template #2416Summary
Use templateThis n8n workflow demonstrates how we can use Multimodal LLMs to parse and extract from PDF documents in n8n. In this particular scenario, we're passing a candidate's CV/resume to an AI which filters out unqualified applications. However, this sneaky candidate has added in hidden prompt to bypass our bot! Whatever will we do? No fret, using AI Vision is one approach to solve this problem... read on! How it works Our candidate's CV/Resume is a PDF downloaded via Google Drive for this demonstration. The PDF is then converted into an image PNG using a tool called Stirling PDF. Since the hidden prompt has a white font color, it is is invisible in the converted image. The image is then forwarded
Hand off to your agent
Prompt
Help me set up the n8n workflow "CV Resume PDF Parsing with Multimodal Vision AI" (https://n8n.io/workflows/2416). It uses: Edit Image, HTTP Request, Google Drive, Basic LLM Chain, Structured Output Parser, Google Gemini Chat Model. Import the template JSON into my n8n instance, list every credential I need to create, and walk me through testing it.
Paste into Claude Code and it will do the rest.
Apps and nodes
Similar workflows
NameViews
- Automate Image Validation Tasks using AI Vision22K
- Transcribing Bank Statements To Markdown Using Gemini Vision AI15K
- Easy Image Captioning with Gemini 1.5 Pro12K
- Resume Screening & Behavioral Interviews with Gemini, Elevenlabs, & Notion ATS5,127
- Visual Regression Testing with Apify and AI Vision Model4,738
- Automate data extraction from faxes & PDFs using Google Gemini and Google Sheets726
- Automated Viral Content Engine for LinkedIn & X with AI Generation & Publishing402
- WordPress blog automation with Airtable interface, human review & AI research v2174