Source profileQuality 92/100

minhnv0807/ai-business-skills/modules/personal-branding/en/24-ai-avatar-production-global/SKILL.md

24-ai-avatar-production-global

Use when a PERSONAL brand needs AI avatar video at scale — three tool tiers, four workflows for single avatar, translation, batch, and hybrid, reference image intake, face, style, logo, and palette replacement, voice clone pairing, anti-detection, and a QA score, with disclosure-law variants for US FTC, EU AI Act, SEA, and LATAM, covering HeyGen and Synthesia. Trigger on 'AI avatar', 'HeyGen video', 'Synthesia', 'talking head AI video', 'translate my videos with AI', 'I cannot be on camera every

Source repository stars
553
Declared platforms
0
Static risk flags
0
Last source update
2026-08-17
Source checked
2026-08-25

Decision brief

What it does: where it fits

Flagship skill of the AI Content cluster. Covers the full pipeline from zero to publish, voice clone, anti-detection, and region-specific disclosure law.

Best for

  • Use when a PERSONAL brand needs AI avatar video at scale — three tool tiers, four workflows for single avatar, translation, batch, and hybrid, reference image intake, face, style, logo, and palette replacement, voice cl…

Not for

  • Tasks that require unconfirmed production actions or broad system permissions.
  • Environments where the pinned source and install steps cannot be inspected.

Compatibility matrix

Platform support, with evidence labels

PlatformStatusEvidenceWhat to check
CodexNot declaredNo explicit evidencePortability before use
Claude CodeNot declaredNo explicit evidencePortability before use
CursorNot declaredNo explicit evidencePortability before use
Gemini CLINot declaredNo explicit evidencePortability before use
Open the compatibility checker

Installation

Inspect first. Install second.

The source command is displayed only when detected. A safe inspection prompt is always available so your agent can explain every action before execution.

Source-detected install commandSource
npx skills add https://github.com/minhnv0807/ai-business-skills --skill "modules/personal-branding/en/24-ai-avatar-production-global"
Safe inspection promptEditorial

Inspect the Agent Skill "24-ai-avatar-production-global" from https://github.com/minhnv0807/ai-business-skills/blob/958bd43b03afbc5afc42dfdb3fb2c087b23309e4/modules/personal-branding/en/24-ai-avatar-production-global/SKILL.md at commit 958bd43b03afbc5afc42dfdb3fb2c087b23309e4. List every install step, command, network request, credential, file read/write, external action, and rollback step. Explain whether it fits my task. Do not install or execute anything until I approve.

Workflow

What the source asks the agent to do

  1. 01

    Workflow 1: Single Avatar Production

    One video, end-to-end in 30-60 minutes.

    One video, end-to-end in 30-60 minutes.
  2. 02

    6-step process

    Review the “6-step process” section in the pinned source before continuing.

    Review and apply the “6-step process” source section.
  3. 03

    Workflow 2: Multi-language translate

    One source video - many languages for global rollout. Use cases: DTC brand expanding markets, multi-language courses, multi-country agency work.

    Create source video (Workflow 1)Upload to translate tool (Rask AI recommended)Pick target language — tool auto-translates and lipsyncs
  4. 04

    Process

    1. Create source video (Workflow 1) 2. Upload to translate tool (Rask AI recommended) 3. Pick target language — tool auto-translates and lipsyncs 4. Review with a native speaker 5. Export and publish per market

    Create source video (Workflow 1)Upload to translate tool (Rask AI recommended)Pick target language — tool auto-translates and lipsyncs
  5. 05

    Workflow 3: Batch Production

    30 videos in 5 days — assembly-line process.

    Templated scripts: 3-5 frameworks, swap the core contentVoice consistency: One voice clone for the entire seriesOff-peak rendering: Queue overnight to skip the queue

Permission review

Static risk signals and limitations

No configured static risk pattern was detected

This is not proof of safety. Runtime behavior, indirect dependencies, and hidden external systems are outside the static scan.

Evidence record

Why each signal appears

EvidenceSourceComputedTestedEditorial
SignalValueEvidence typeMeaning
Quality score92/100ComputedDocumentation, specificity, maintenance, and trust rules
Repository stars553SourceRepository attention, not individual Skill quality
Compatibility0 platformsSourceDeclared in the catalog source record
Usage guideautomated source guideEditorialGenerated or reviewed according to the visible evidence level

Pinned source

Provenance and original SKILL.md

Repository
minhnv0807/ai-business-skills
Skill path
modules/personal-branding/en/24-ai-avatar-production-global/SKILL.md
Commit
958bd43b03afbc5afc42dfdb3fb2c087b23309e4
License
MIT
Collected
2026-08-25
Default branch
master
View the original SKILL.md

AI Avatar Production (Global) — Pipeline 3-Tier, 4 Workflows, QA Score 100

Flagship skill of the AI Content cluster. Covers the full pipeline from zero to publish, voice clone, anti-detection, and region-specific disclosure law.


For newbies

What is an AI Avatar?

An AI Avatar is a video that shows your face (or a stand-in) but uses AI-generated voice and motion. You provide one photo or a short selfie video; the AI produces a final video with natural-looking speech, gestures, and expressions. No filming crew, no studio, no actor required.

What do you need to start?

MethodRequirementQuality
Portrait photo1 forward-facing photo, clean background, 1024x1024+Medium — mouth less natural
Selfie video30s video, looking at the lens, speaking naturallyGood — better lipsync
Custom avatar2-5 min recording with teleprompter + lavalier micExcellent — near photo-real

Minimum gear: Phone with HD front camera + lavalier mic (or headset mic).

How long does it take?

  • One single video (60s): 30-60 min (script + render)
  • Batch of 10: 1-2 days
  • Batch of 30: 4-5 days (with optimized process)

What does it cost?

TierUSD/monthOutput
Free$01-3 videos, watermark
Pro$30-10010-30 videos, no watermark
Enterprise$200-500+30+ videos, custom avatar, API

5 common newbie mistakes

  1. Lipsync drift: Script too fast or voice mismatch -> slow speech 10-15%, use voice clone instead of default voice.
  2. Voice doesn't sound like you: Sample too short or noisy -> re-record 3-5 minutes in a quiet room with phonetically varied script.
  3. Video flagged as "AI content": Platform pattern detection -> see Anti-detection section below.
  4. Blurry / pixelated output: Low-quality input -> use 1024x1024+ photo, natural lighting, no filters.
  5. Slow render: Free tier queue -> render off-peak (early morning in your timezone = US night) or upgrade to Pro.

Information collection (4 questions max)

Ask up to 4 questions before starting:

  1. Primary use case? Brand awareness / Sales / Education / Internal training?
  2. Primary platform? TikTok / YouTube / Facebook / Instagram / LinkedIn / X / Threads?
  3. Budget tier? Free ($0) / Pro ($30-100/mo) / Enterprise ($200+/mo)?
  4. Videos per month target? 1-5 / 10-30 / 30+?

Based on the 4 answers, auto-select Tier + Workflow. If the user has already uploaded reference images, do not ask a long intake form first; classify the images, create the setup/prompt, then ask only for missing assets.


Tier decision — Tools and pricing

TierSuggested toolPrice/monthQualityLimitFits
FreeCaptions Free, HeyGen Trial, D-ID Trial$06/10 — watermark, limited duration1-5 videos, max 60s/videoPersonal test, new freelancers
ProHeyGen Creator ($29), Synthesia Starter ($29), ElevenLabs Pro ($22)$30-1008/10 — no watermark, HD10-30 videos, max 5 min/videoSME, small agency, content creator
EnterpriseHeyGen Business ($89+), Synthesia Enterprise (custom)$200-500+9.5/10 — custom avatar, API, priority render30+ videos, unlimitedLarge agency, large brand, e-learning

Quick recommendations:

  • Just starting: HeyGen Trial (1 video free, full experience)
  • Serious but budget-limited: Captions Pro ($10/mo) for lipsync + ElevenLabs Starter ($5) for voice
  • Scale fast: HeyGen Creator + ElevenLabs Pro = best price/quality combo
  • Enterprise: Synthesia Enterprise + ElevenLabs Scale

Workflow 1: Single Avatar Production

One video, end-to-end in 30-60 minutes.

6-step process

StepTaskToolTime
1. Script150-300 words for a 60s videoSkill 04-script-video-global10 min
2. VoiceGenerate or use voice cloneElevenLabs / HeyGen Voice5 min
3. AvatarPick stock avatar or upload your mediaHeyGen / Synthesia / D-ID3 min
4. RenderCombine voice + avatar, choose background, gesturesTool from step 35-15 min (render)
5. QAQA Score 100 review (see section below)Manual review5 min
6. PublishExport MP4 -> post to platformManual / Scheduler2 min

Script template for AI Avatar (60s)

[HOOK — 3s] Curiosity hook, frame the problem
[PROBLEM — 10s] Describe the customer pain
[SOLUTION — 25s] Your solution, 2-3 key points
[PROOF — 12s] Numbers, testimonial, result
[CTA — 10s] Concrete action: "Link in bio for..."

Workflow 2: Multi-language translate

One source video -> many languages for global rollout. Use cases: DTC brand expanding markets, multi-language courses, multi-country agency work.

Tool comparison

ToolLanguagesPriceNotes
Rask AI130+$50/mo (Pro)Best for translate today
HeyGen Translate40+Included Creator+Built-in, convenient
Synthesia Translate35+Included EnterpriseBest for e-learning

Process

  1. Create source video (Workflow 1)
  2. Upload to translate tool (Rask AI recommended)
  3. Pick target language — tool auto-translates and lipsyncs
  4. Review with a native speaker
  5. Export and publish per market

Caveat: Tonal languages (Mandarin, Vietnamese, Thai) have weaker lipsync. Workaround: produce native voice clone + native avatar per language.

See full disclosure law per region in the variant files.


Workflow 3: Batch Production

30 videos in 5 days — assembly-line process.

Detailed timeline

DayTaskOutputTool
Day 1Script batch — write 10 scripts from template10 scripts (.md)Skill 04-script-video-global + AI assist
Day 2Voice batch — render 10 audio files10 audio (.mp3)ElevenLabs API
Day 3Avatar batch — upload audio + avatar, queue render10 videos renderingHeyGen Batch / Synthesia
Day 4QA batch — review 10 videos, fix issues, re-render10 QA'd videosManual + QA Score
Day 5Publish batch — export, add captions, schedule10 videos publishedBuffer / Later / Manual

Repeat 3 weeks = 30 videos. Or scale Days 1-2 to 15 scripts/week.

Cost estimate batch 30 videos/month

TierTool comboMonthly costPer-video cost
FreeHeyGen Trial + Captions Free$0 (limited 3-5 videos)$0 (watermark)
ProHeyGen Creator + ElevenLabs Pro~$51~$1.70
EnterpriseHeyGen Business + ElevenLabs Scale~$189~$6.30

Batch optimization tips

  • Templated scripts: 3-5 frameworks, swap the core content
  • Voice consistency: One voice clone for the entire series
  • Off-peak rendering: Queue overnight to skip the queue
  • QA checklist: Print the QA Score, check videos like an assembly line

Workflow 4: Hybrid Real + AI

Real face for trust + AI body for speed.

Use cases

  • Real face intro 5s + AI body 55s (save filming time)
  • AI video weekdays + Real video weekly (balance quality/effort)
  • Real talking head + AI B-roll (studio-grade output)

Assembly + tools

  1. Film real intro 5-10s (eye contact, natural greeting); use Captions for lipsync fixes
  2. Create AI for the rest with same outfit/background (HeyGen / Synthesia)
  3. Edit in CapCut / Premiere (precise cuts, smooth transitions)
  4. Color match AI to real footage (LUT or DaVinci Resolve free)

Trust gain: Real face up front -> 20-35% more engagement than full-AI.


Voice Clone Protocol

Voice sample requirements

CriterionRequirement
Duration3-5 minutes
QualityWAV/FLAC, 44.1kHz+, mono, quiet room
Script contentPhonetically varied passages (all vowels, hard consonants)
EmotionRead normal, natural, not acted

Tool comparison

ToolPriceQualityNotes
ElevenLabsFrom $5/mo9/10Best overall, 30+ languages
HeyGen VoiceIncluded Creator+6/10Convenient if using HeyGen
Resemble AIFrom $99/mo7/10Strong API
PlayHTFrom $39/mo7/10Good for narration

Consent form template

MANDATORY before cloning anyone's voice.

VOICE USAGE CONSENT

I, [FULL NAME], consent to [COMPANY] using my voice for: [SPECIFIC PURPOSE].
Term: [X months / Until revoked]
Date: [YYYY-MM-DD]
Signature: _______________

Reference: See references/voice-clone-prompts-global.md


Avatar Setup Checklist

Before recording / uploading photo or video for an AI avatar:

  • Lighting: Natural light or softbox; no harsh shadows on the face
  • Background: Solid (white / gray) or real environment (office, store)
  • Wardrobe: On-brand; avoid small busy patterns (AI moire)
  • Framing: Chest up; eyes on the upper-third line
  • Eye contact: Look directly at the lens (not the screen)
  • Gestures: Natural; hands can rest or do light gestures
  • Resolution: Minimum 1080p (1920x1080); 4K preferred
  • Aspect ratio: 9:16 (TikTok / Reels), 16:9 (YouTube), 1:1 (Feed)
  • File format: MP4 (H.264) for video, PNG / JPG for photo
  • Backup: Keep originals on cloud (Google Drive / OneDrive) before uploading to the tool

Reference Image -> Avatar Prompt Director

Use this when the user drops one or more reference images and wants to create an avatar, replace a face, adapt brand colors, add a logo, or create the prompt before uploading assets into a tool.

Classify Input Images

Image typeRoleRequirement
Style refMood, lighting, background, outfit, camera angleDo not use as identity unless requested
Face refIdentity preservation / face replacement1-3 clear face images, no filter, front + 3/4 angle
Selfie videoBetter custom avatar / natural lipsync30s-2 min, looking at camera, speaking naturally
Logo/palettePersonal/company brand adaptationPNG/SVG logo + 2-4 hex colors
Product/locationProp or avatar environmentClear product label or location/background image

Multiple Images = Multiple Flows

## Avatar Flows

| Flow | Input image | Role | Suggested tool | Missing assets |
|------|-------------|------|----------------|----------------|
| A | style-01 | style/background | Design Master -> HeyGen | face ref, logo |
| B | face-01 | identity | HeyGen custom avatar | script, voice sample |
  • If every image is a different style direction, create a separate prompt for each flow.
  • If images support one avatar, group by role: style + face + logo + palette + product.
  • Ask for each next asset explicitly: face image, selfie video, logo, hex colors, script, voice sample.

Prompt Setup Output

## Avatar Prompt Setup — Flow A

- Style ref:
- Face ref:
- Brand assets:
- Target platform:
- Tool route:

## Copy-Paste Visual Prompt
[English prompt for avatar/source image generation]

## Upload Next
- Face/selfie video:
- Logo:
- Brand colors:
- Voice sample:
- Script:

For a static personal avatar only, route to 30-design-master-global personal-brand mode. For talking-head video, continue this workflow.


Anti-detection for FB / IG / TikTok / YouTube

5 detection signals and fixes

SignalPlatforms flaggingFix
Stiff face, no natural blinkingFB, IGUse selfie video over photo; pick avatars with micro-expressions
Monotone voice, no natural pausesTikTok, FBUse voice clone (natural pacing) over default TTS
Fully static backgroundFB, IGAdd slight noise/grain, or use real-world background
Isolated motion (only mouth moves)TikTokPick avatars with gesture (hands, head); use HeyGen v3+
Metadata flagged as AI toolYouTube (monetize)Re-export through CapCut (strips metadata); add color grade

Techniques to add "human feel"

  1. Add film grain / noise: 2-5% in CapCut or Premiere
  2. Zoom and crop: 5-10% crop with subtle motion (Ken Burns)
  3. Color grade: Apply film LUT or manually grade — avoid "too clean"
  4. Text overlay: Add subtitles, callouts, stickers to cover AI weak spots
  5. B-roll insert: Drop 2-3 b-roll clips (product, lifestyle) every 15-20s
  6. Sound design: Background music + light SFX (immersion + masks AI voice)

Per platform

  • TikTok: Most lenient — content quality wins over AI checks
  • Facebook / Instagram: Moderate scrutiny — anti-detection matters
  • LinkedIn: Practically no detection — best fit for AI avatars
  • YouTube: Strict for monetized videos — must disclose per YPP policy

CRITICAL: NEVER use AI avatars to impersonate real people without consent. This is illegal in most jurisdictions and grounds for permanent platform bans.


Ethics and Disclosure — Region selector

Disclosure laws differ dramatically by region. Pick the matching variant:

RegionVariant fileKey law
US / Canadavariants/01-us.mdFTC Endorsement Guides (16 CFR Part 255), 2023 update
EU / EEA / UKvariants/02-eu.mdEU AI Act Article 50 (always disclose) + UCPD + GDPR
Southeast Asiavariants/03-sea.mdPer-country: ASAS (SG), AKARI (ID), DTI (PH), MCMC (MY), TH
Latin Americavariants/04-latam.mdCONAR + LGPD (BR), PROFECO (MX), AAIP (AR), per-country

ALWAYS read the matching variant BEFORE publishing AI avatar content in that region. Penalties range from warning to multi-thousand-USD fines per influencer (US) and can stack under EU AI Act + GDPR.

Universal disclosure rule of thumb

When in doubt, disclose. Disclosure is rarely penalized; non-disclosure can be.

"This video uses AI Avatar technology for visuals and voice."

Placement: video description, first 3 seconds on-screen text, OR platform "AI-generated" tag (where available — Meta, TikTok, YouTube all now support this).


QA Score — 100 points

Scorecard

#CriterionPointsDescription
1Lipsync/10Mouth tracks speech within 0.2s
2Voice match/10Voice sounds like the speaker (if clone) or natural (if TTS)
3Visual quality/10Sharp image, no artifacts, no blur
4Background/10Background suits context, no render glitches
5Lighting/10Even light, no harsh shadows, matches background
6Gesture/10Natural, no jitters, hand/head movement present
7Script flow/10Hook -> Problem -> Solution -> CTA
8Disclosure/10AI disclosure compliant with region (see variant)
9Platform fit/10Correct aspect ratio, duration, format for platform
10CTA/10Clear call-to-action, easy to execute

Action thresholds

TierScoreAction
Excellent90-100Publish now
Good70-89Publish, note improvements for next round
Needs fix50-69Fix items scoring under 7, then re-render
Redo<50Rebuild from script + voice + avatar

Output template

# AI Avatar Video — [Title] | [Region variant] | [Date]

1. Workflow used: [Single / Translate / Batch / Hybrid]
2. Script: [Content, 150-300 words]
3. Voice: [Tool] — [Voice ID / clone name] — Consent: [Yes / N/A]
4. Avatar: [Tool] — [Avatar ID / custom]
5. QA Score: [X]/100 (10 criteria)
6. Disclosure (per region variant): [Text + placement]
7. Publish: [Platform] — [Aspect ratio] — [Link]

Quality checklist

  • Information collection completed (4 questions)
  • Tier picked (Free / Pro / Enterprise) and aligns with budget + volume
  • Workflow picked (Single / Translate / Batch / Hybrid)
  • Voice clone consent recorded (if cloning a real person)
  • Avatar setup checklist completed before recording
  • Anti-detection techniques applied for the target platform
  • Region variant read and disclosure compliant
  • QA Score >= 70 before publishing

Related skills

  • 25-voice-clone-podcast-global — voice clone deep-dive + podcast pipeline
  • 04-script-video-global — script writing for AI avatar
  • 26-thought-leadership-content-global — content strategy for personal brand
  • references/ai-video-disclosure-global — full legal reference
  • references/voice-clone-prompts-global — voice clone training prompts

Global Skill 24 (AI Avatar Production) | Over Powers Agency | v1.1.0

Frequently asked questions

What to verify before installation and use

What does the 24-ai-avatar-production-global source document cover?

Flagship skill of the AI Content cluster. Covers the full pipeline from zero to publish, voice clone, anti-detection, and region-specific disclosure law.

How do I install 24-ai-avatar-production-global?

The source record exposes this install command: npx skills add https://github.com/minhnv0807/ai-business-skills --skill "modules/personal-branding/en/24-ai-avatar-production-global". Inspect the command and pinned source before running it.

Alternatives

Compare before choosing

Computed 9439,098

wshobson/agents

brand-landingpage

Brand-first landing page designer — runs a brand-identity interview (colors, typography, shape language), then generates and iterates on a polished landing page via Stitch with deployment-ready HTML. Use when the user asks to create, design, or build a landing page, homepage, or marketing page and has no established visual direction. Skip when they have a design mockup, need a dashboard or app UI, are working at component level, building a multi-page app, or restyling with known design tokens —

Computed 94584

nexscope-ai/Amazon-Skills

amazon-global-selling

Evaluate and plan Amazon marketplace expansion across countries and regions. Use when a seller asks which Amazon marketplace to enter, how to compare international demand and economics, what tax, product-compliance, logistics, localization, account, or launch workstreams to investigate, or how to build a gated global-selling roadmap. Do not use as legal, tax, customs, or certification advice.

Computed 931,248

first-fluke/oh-my-agent

oma-translation

Context-aware translation that preserves tone, style, and natural word order. Use when translating UI strings, documentation, marketing copy, or any multilingual content. Infers register, domain, and style from the source text and surrounding codebase context.

Computed 921,248

first-fluke/oh-my-agent

oma-translation

Context-aware translation that preserves tone, style, and natural word order. Use when translating UI strings, documentation, marketing copy, or any multilingual content. Infers register, domain, and style from the source text and surrounding codebase context.