n8n-nodes-base.evaluation
23 workflows using this node.
What is the n8n-nodes-base.evaluation node used for in n8n workflows?
n8n-nodes-base.evaluation n8n workflows are automation templates that use the n8n-nodes-base.evaluation node as part of a reusable workflow. Builders use these templates to inspect how the node connects with triggers, data transformation steps, app integrations, notifications, storage, or AI-powered actions inside n8n. Each workflow listing includes a title, summary, author attribution, categories, complexity level, source links when available, and a downloadable JSON file for review before import. Use this page to compare practical examples, understand the surrounding nodes, and choose a template that can be safely tested and adapted in your own n8n workspace.
n8n-nodes-base-evaluation Monitor AI quality drift with GPT-4o-mini evaluations and Slack alerts
Catch AI quality drift before your users do. This template ties scheduled evaluation, LLM as a Judge scoring, and thr...
Score customer support AI responses with GPT‐4 judge metrics
Score open ended AI responses with a judge model. This template shows how to evaluate a customer support agent using...
Evaluate a support ticket classifier with OpenAI GPT-4o-mini and n8n evaluations
Measure how well your AI classifier actually performs. This template shows how to evaluate a support ticket classifie...
Route and qualify email leads with Gmail, Gemini, Slack, Sheets and Salesforce
Email Lead Router: Gmail → Gemini → Salesforce Pipeline Who is this for? Event sales teams & conference organizers pr...
Route event sales leads with Gmail, Google Gemini, Sheets and Salesforce
Email Sentiment Router for Event Sales Leads Who is this for? Event organizers, conference managers, and sales teams...
Evaluate AI workflows using Google Sheets, Gemini, Claude, GPT, and Perplexity
This template and YouTube video goes over 5 different implementations of evaluations within n8n. Categorization Corre...
Extract meeting details with GPT-4.1-mini and evaluate accuracy in Google Sheets
Who's it for Developers building AI powered workflows who want to ensure their agents work reliably. If you need to v...
Sales lead routing with Gemini Sentiment Analysis & Model Evaluation Framework
This n8n template demonstrates how to deploy an AI workflow in production while simultaneously running a robust, data...
Automate Reddit replies with F5Bot alerts & GPT-5 personalized comments
Reddit Auto Comment Assistant (AI Driven Marketing Workflow) Automate how you reply to Reddit posts using AI generate...
My solution for the "Agentic Arena Community Contest" (RAG, Qdrant, Mistral OCR)
This workflow is my personal solution for the Agentic Arena Community Contest , where the goal is to build a Retrieva...
🎓 Learn evaluate tool. Tutorial for beginners with Gemini and Google Sheets
This workflow is a beginner friendly tutorial demonstrating how to use the Evaluation tool to automatically score the...
Evaluate tool usage accuracy in multi-agent AI workflows using evaluation nodes
Who's it for This workflow is ideal for AI developers running multi agent systems in n8n who need to quantitatively e...
Custom Discord notifications for Radarr, Sonarr, Bazarr etc.
This is a simple temlate that will allow you to customise the notifications in Radarr, Sonarr, Bazarr and similar. By...
Evaluation metric: summarization
This n8n template demonstrates how to calculate the evaluation metric "Summarization" which in this scenario, measure...
Evaluate RAG response accuracy with OpenAI: document groundedness metric
This n8n template demonstrates how to calculate the evaluation metric "RAG document groundedness" which in this scena...
Evaluate AI agent response relevance using OpenAI and cosine similarity
This n8n template demonstrates how to calculate the evaluation metric "Relevance" which in this scenario, measures th...
Evaluate AI agent response correctness with OpenAI and RAGAS methodology
This n8n template demonstrates how to calculate the evaluation metric "Correctness" which in this scenario, measures...
Evaluations metric: answer similarity
This n8n template demonstrates how to calculate the evaluation metric "Similarity" which in this scenario, measures t...
Evaluation metric example: String similarity
AI evaluation in n8n This is a template for n8n's evaluation feature. Evaluation is a technique for getting confidenc...
Evaluation metric example: RAG document relevance
AI evaluation in n8n This is a template for n8n's evaluation feature. Evaluation is a technique for getting confidenc...
Evaluation metric example: Correctness (judged by AI)
AI evaluation in n8n This is a template for n8n's evaluation feature. Evaluation is a technique for getting confidenc...
Evaluation metric example: categorization
AI evaluation in n8n This is a template for n8n's evaluation feature. Evaluation is a technique for getting confidenc...
Evaluation metric example: Check if tool was called
AI evaluation in n8n This is a template for n8n's evaluation feature. Evaluation is a technique for getting confidenc...