Evaluation metric example: Check if tool was called
n8n template #4268Summary
Use templateAI evaluation in n8n This is a template for n8n's evaluation feature. Evaluation is a technique for getting confidence that your AI workflow performs reliably, by running a test dataset containing different inputs through the workflow. By calculating a metric (score) for each input, you can see where the workflow is performing well and where it isn't. How it works This template shows how to calculate a workflow evaluation metric: whether a specific tool was called by an agent. We use an evaluation trigger to read in our dataset It is wired up in parallel with the regular trigger so that the workflow can be started from either one. More info We make sure that the agent outputs the list of too
Hand off to your agent
Prompt
Help me set up the n8n workflow "Evaluation metric example: Check if tool was called" (https://n8n.io/workflows/4268). It uses: AI Agent, OpenAI Chat Model, Calculator, Evaluation. Import the template JSON into my n8n instance, list every credential I need to create, and walk me through testing it.
Paste into Claude Code and it will do the rest.
Apps and nodes
Similar workflows
NameViews
- Build Your First AI Data Analyst Chatbot134K
- AI marketing report (Google Analytics & Ads, Meta Ads), sent via email/Telegram36K
- ERP AI chatbot for Odoo sales module with OpenAI19K
- Personal Shopper Chatbot for WooCommerce with RAG using Google Drive and openAI12K
- Turn Emails into AI-Enhanced Tasks in Notion (Multi-User Support) with Gmail, Airtable and Softr7,647
- Create a Pizza Ordering Chatbot with GPT-3.5 - Menu, Orders & Status Tracking5,534
- Explore n8n Nodes in a Visual Reference Library5,197
- Time logging on Clockify using Slack4,274