Skip to main content

Create faceless videos with Gemini, ElevenLabs, Leonardo AI & Shotstack

Workflow preview

Workflow preview
100%
Create faceless videos with Gemini, ElevenLabs, Leonardo AI & Shotstack preview
Open on n8n.io

Important notice

This workflow is provided as-is. Please review and test before using in production.

1. Workflow Overview

This n8n template demonstrates walks you through a fully automated process to generate faceless videos from script creation to final download using AI generated voice, images, and smart video editi...

Best for

  • Content Creation automation workflows
  • Multimodal AI automation workflows
  • advanced n8n builders looking for reusable templates

Tools used

n8n-nodes-base.googledrive, n8n-nodes-base.merge, @n8n/n8n-nodes-langchain.outputparserstructured, n8n-nodes-base.stickynote, n8n-nodes-base.manualtrigger, @n8n/n8n-nodes-langchain.chainllm, @n8n/n8n-nodes-langchain.outputparserautofixing, n8n-nodes-base.splitout

Source and attribution

This workflow is cataloged by N8N Workflows and links back to its original n8n.io source page by Agent Circle.

Original n8n.io source

1.1 Workflow description

Title
Create faceless videos with Gemini, ElevenLabs, Leonardo AI & Shotstack
Workflow name
Create faceless videos with Gemini, ElevenLabs, Leonardo AI & Shotstack

This n8n template demonstrates walks you through a fully automated process to generate faceless videos - from script creation to final download - using AI-generated voice, images, and smart video editing.

Use cases are many: This tool is perfect for YouTube and Shorts creators who want to publish daily content without showing their face, TikTok and Reels marketers automating voice-over-driven videos, and solopreneurs scaling up their content without hiring a team. It’s also ideal for agencies producing batches of faceless video ads, automation enthusiasts building smart media workflows in n8n, and anyone who’s rich in ideas but tired of spending hours editing.

How It Works

  • Phase 1: Provide Topic Input
    • A short topic and idea should be entered into the Idea part in Node Fields - Set Idea inside the workflow in n8n.
    • Trigger the process manually by clicking Test Workflow or Execute Workflow.
  • Phase 2: Script Generation
    • Your idea is passed to Google Gemini's chat model. The model returns a concise, 60-second faceless video script.
    • The script is then reformatted into a structured layout optimized for voice generation and visual synchronization.
  • Phase 3: Audio Generation
    • The formatted script is passed to ElevenLabs, which turns the text into a high-quality voiceover audio.
    • The generated audio is uploaded to Google Drive and made publicly accessible.
    • At the same time, the audio is sent to OpenAI Whisper via a POST request to generate a transcription of the voiceover.
  • Phase 4: Timestamps Generation
    • The tool merges the original script and the OpenAI Whisper-generated transcription.
    • The merged data is passed to Google Gemini's chat model to generate image prompts with precise timestamps.
    • The output is parsed and cleaned using a structured parser to ensure it's in ready-to-use JSON format for image generation.
  • Phase 5: Images Generation
    • The full list of timestamped prompts is is split into individual entries.
    • Each prompt is sent to Leonardo's API that turns text descriptions into visuals.
    • A delay of 30 seconds is added to give the image generation engine enough time to complete rendering.
    • Once completed, the workflow retrieves all final images for the next stage.
  • Phase 6: Images To Video Conversion
    • All generated images are sent to Leonardo's API, which stitches them together based on the structured prompts and timing.
    • A 5-minute wait allows time for rendering.
    • After the wait, the workflow retrieves the generated small videos and makes them downloadable.
    • Then, the tool aggregates all downloaded videos into a single unified structure, preparing them for the final editing.
  • Phase 7: Video Editing and Downloading
    • The raw video, along with timestamps or subtitles, is sent to Shotstack, a video editing tool that supports advanced edits.
    • A delay of 1 minute allows Shotstack to process the edit.
    • Then, the tool checks whether the edited video is finished by Shotstack and ready to be downloaded.
    • Once completed, you can download the final polished video to your local storage for later use.

How To Use

  • Download the workflow package.
  • Import the package into your n8n interface.
  • Set up necessary credentials for tools access and usability:
    • For Google Gemini access, please connect to its API in the following nodes: Node Google Gemini Chat Model 1 Node Google Gemini Chat Model 2
    • For Google Drive access, please ensure connection in the following nodes: Node Upload Audio to Drive Node Make Audio File Public
    • For ElevenLabs access, please connect to its API in the following node: Node Generate Voice
    • For OpenAI Whisper access, please connect to its API in the following node: Node Transcribe Audio with OpenAI Whisper
    • For Leonardo access, please allow connection to its API in the following nodes: Node Generate Images Node Generate Videos/Scenes
    • For Shortstack access, please connect to its API in the following nodes: Node Edit with Shotstack Node Render Final Video with Shotstack
  • Input your video idea or short description as a string in Node Fields - Set Idea in n8n.
  • Run the workflow by clicking Execute Workflow or Test Workflow.
  • Wait the process to run and finish.
  • View the result in Node Download Final Video and download it in your local storage for later use.

Requirements

  • Basic setup in Google Cloud Console (OAuth or API Key method enabled) with enabled access to Google Drive.
  • Google Gemini API access with permission to use chat-based large language models.
  • ElevenLabs API access for generating high-quality voiceovers from scripts.
  • OpenAI Whisper API access to transcribe voiceovers into clean text.
  • Leonardo API access for both image and video generation tasks.
  • Shotstack API access for editing and rendering the final video with enhanced visuals and timing.

How To Customize

  • You can input your requested video topic or description directly in Node Fields – Set Idea.
  • By default, the script length is set to around 60 seconds in Node 60 Second Script Writer. You can easily change this in the prompt to create shorter or longer videos based on your needs.
  • While the default setup uses Google Gemini for script and prompt generation, you can replace it with OpenAI ChatGPT, Claude, or any other compatible chat-based model you prefer.
  • The voiceover is currently created using ElevenLabs, but you’re free to substitute it with other text-to-speech engines like Google Cloud Text-to-Speech, HeyGen, etc.
  • We're using OpenAI Whisper to transcribe the voiceover into text. You can switch to alternatives such as AssemblyAI, Deepgram, or other compatible providers depending on your preference.
  • This workflow uses Leonardo for both image and video generation. You can swap it out for other compatible providers based on availability or style preference.
  • Video editing is handled by Shotstack by default. You can plug in alternatives like Runway, FFmpeg, or other API-based editors depending on your editing needs or desired effects.

If you’d like this workflow customized to fit your tools and platforms availability, or if you’re looking to build a tailored AI Agent for your own business - please feel free to reach out to Agent Circle. We’re always here to support and help you to bring automation ideas to life.

Need Help?

Join our community on different platforms for support, inspiration and tips from others.

Website: https://www.agentcircle.ai/ Etsy: https://www.etsy.com/shop/AgentCircle Gumroad: http://agentcircle.gumroad.com/ Discord Global: https://discord.gg/d8SkCzKwnP FB Page Global: https://www.facebook.com/agentcircle/ FB Group Global: https://www.facebook.com/groups/aiagentcircle/ X: https://x.com/agent_circle YouTube: https://www.youtube.com/@agentcircle LinkedIn: https://www.linkedin.com/company/agentcircle

1.2 Logical Blocks

This catalog entry is organized from the workflow JSON. The node-level section below shows the executable blocks available for review before importing the template.

2. Block-by-Block Analysis

Block 1 - Upload Audio to Drive

Type / Role
n8n-nodes-base.googleDrive - googleDrive
Config choices
Version 3

Block 2 - Merge

Type / Role
n8n-nodes-base.merge - merge
Config choices
Version 3

Block 3 - Structured Output Parser1

Type / Role
@n8n/n8n-nodes-langchain.outputParserStructured - outputParserStructured
Config choices
Version 1.2

Block 4 - Sticky Note22

Type / Role
n8n-nodes-base.stickyNote - stickyNote
Config choices
Version 1

Block 5 - Sticky Note23

Type / Role
n8n-nodes-base.stickyNote - stickyNote
Config choices
Version 1

Block 6 - Sticky Note24

Type / Role
n8n-nodes-base.stickyNote - stickyNote
Config choices
Version 1

Block 7 - Sticky Note25

Type / Role
n8n-nodes-base.stickyNote - stickyNote
Config choices
Version 1

Block 8 - Sticky Note29

Type / Role
n8n-nodes-base.stickyNote - stickyNote
Config choices
Version 1

Block 9 - Sticky Note30

Type / Role
n8n-nodes-base.stickyNote - stickyNote
Config choices
Version 1

Block 10 - When clicking ‘Test workflow’

Type / Role
n8n-nodes-base.manualTrigger - manualTrigger
Config choices
Version 1

Block 11 - Generate Image Prompts

Type / Role
@n8n/n8n-nodes-langchain.chainLlm - chainLlm
Config choices
Version 1.5

Block 12 - Auto-fixing Output Parse

Type / Role
@n8n/n8n-nodes-langchain.outputParserAutofixing - outputParserAutofixing
Config choices
Version 1

Block 13 - Split Prompts

Type / Role
n8n-nodes-base.splitOut - splitOut
Config choices
Version 1

Block 14 - Wait 5 mins

Type / Role
n8n-nodes-base.wait - wait
Config choices
Version 1.1

Block 15 - Aggregate

Type / Role
n8n-nodes-base.aggregate - aggregate
Config choices
Version 1

Block 16 - Edit with Shotstack

Type / Role
n8n-nodes-base.httpRequest - httpRequest
Config choices
Version 4.2

Block 17 - Wait 1 min

Type / Role
n8n-nodes-base.wait - wait
Config choices
Version 1.1

Block 18 - Fields - Set Idea

Type / Role
n8n-nodes-base.set - set
Config choices
Version 3.4

Block 19 - 60 Second Script Writer

Type / Role
@n8n/n8n-nodes-langchain.chainLlm - chainLlm
Config choices
Version 1.5

Block 20 - OpenAI Chat Mode

Type / Role
@n8n/n8n-nodes-langchain.lmChatOpenAi - lmChatOpenAi
Config choices
Version 1

Block 21 - OpenAI Chat Model

Type / Role
@n8n/n8n-nodes-langchain.lmChatOpenAi - lmChatOpenAi
Config choices
Version 1

Block 22 - Fields - Script Format

Type / Role
n8n-nodes-base.set - set
Config choices
Version 3.4

Block 23 - Generate Voice

Type / Role
n8n-nodes-base.httpRequest - httpRequest
Config choices
Version 4.2

Block 24 - Generate Images

Type / Role
n8n-nodes-base.httpRequest - httpRequest
Config choices
Version 4.2

Showing the first 24 of 37 workflow blocks. Download the JSON for the full node graph.

3. Summary Table

Workflow Create faceless videos with Gemini, ElevenLabs, Leonardo AI & Shotstack
Complexity advanced
Nodes 37
Categories Content Creation, Multimodal AI
Author Agent Circle
Published 15 Jul 2025

4. Reproducing the Workflow from Scratch

  1. 1. Download the workflow JSON

    Use the JSON export at /data/workflows/6014/6014.json as the source template for this automation.

  2. 2. Import the template into n8n

    Open n8n, import the downloaded JSON, and review each node before activating the workflow.

  3. 3. Configure credentials and variables

    Replace placeholder credentials, API keys, webhook URLs, account IDs, and environment-specific values with your own settings.

  4. 4. Test with sample data

    Run the workflow manually or in a staging workspace, inspect node output, and confirm downstream systems receive the expected data.

  5. 5. Activate and monitor

    Enable the workflow only after testing, then monitor executions, errors, and rate limits during the first production runs.

5. General Notes & Resources

Review imported nodes carefully before activation. This catalog entry is intended to help you inspect the workflow structure, understand required services, and find related templates faster.

Node names, credentials, schedules, webhook paths, and external service limits may need adjustment for your workspace.

Frequently asked questions

What does Create faceless videos with Gemini, ElevenLabs, Leonardo AI & Shotstack do?

This n8n template demonstrates walks you through a fully automated process to generate faceless videos from script creation to final download using AI generated voice, images, and smart video editi...

What do I need before importing this workflow?

Review the workflow JSON, configure any required credentials in n8n, and test the automation in a safe workspace before using it in production.

Can I customize this workflow?

Yes. Use the block-by-block analysis and the downloadable JSON to inspect each node, then adjust credentials, prompts, schedules, filters, or destinations for your Content Creation, Multimodal AI use case.