Multimodal Slack AI assistant with voice, image & video processing
n8n template #9149Summary
Use templateQuick overview This is a starting point for building a Slack AI agent. The base handles four input types: voice, pictures, video, and text, through the AI model of your choice. From here you connect tools to expand what the agent can do inside your n8n workflows. How it works Input: a Slack message that mentions the bot in a channel. A Switch node sorts the message by type: Voice message Picture message Video message Text message It currently uses OpenAI and Gemini to analyze voice, photos, and video, but you can swap in other models. The model reads the message, generates a response from the system prompt, and posts it back to Slack. Setup Create the Slack bot and token. At https://api.slac
Hand off to your agent
Prompt
Help me set up the n8n workflow "Multimodal Slack AI assistant with voice, image & video processing" (https://n8n.io/workflows/9149). It uses: HTTP Request, Slack, AI Agent, Anthropic Chat Model, Simple Memory, OpenAI, Google Gemini. Import the template JSON into my n8n instance, list every credential I need to create, and walk me through testing it.
Paste into Claude Code and it will do the rest.
Apps and nodes
Similar workflows
NameViews
- Advanced AI Demo (Presented at AI Developers #14 meetup)24K
- Slack AI Chatbot for business team with RAG, Claude 3.7 Sonnet and Google Drive7,366
- Explore n8n Nodes in a Visual Reference Library5,197
- Multimodal telegram bot with voice, image & video analysis using Claude & Gemini1,724
- 💅 AI Agents Generate Content & Automate Posting for Beauty Salon Social Media 📲129
- AI-Powered WhatsApp Chatbot 🤖📲 for Text, Voice, Images & PDFs with memory 🧠71K
- AI-Powered WhatsApp Chatbot for Text, Voice, Images, and PDF with RAG47K
- Automated Stock Analysis Reports with Technical & News Sentiment using GPT-4o37K