AT

albedo-tabai/video-copy-analyzer

Developer tools
205ย stars Quality 41 Trend 41

๐ŸŽฌ One-stop video content extraction and copywriting analysis tool. Download videos, smart subtitle extraction (embedded/burned/audio), and analyze scripts using three AI frameworks.

Overview

๐ŸŽฌ One-stop video content extraction and copywriting analysis tool. Download videos, smart subtitle extraction (embedded/burned/audio), and analyze scripts using three AI frameworks.

README

Video Copy Analyzer

ไธญๆ–‡ๆ–‡ๆกฃ | English

๐ŸŽฌ One-stop video content extraction and copywriting analysis tool. Download videos, smart subtitle extraction (embedded/burned/audio), and analyze scripts using three AI frameworks.

โœจ Features

Stage Function Description
1๏ธโƒฃ Video Download Download from Bilibili/YouTube/Douyin (yt-dlp + custom downloader)
2๏ธโƒฃ Smart Subtitle Extraction Three-tier priority: Embedded โ†’ OCR (RapidOCR) โ†’ ASR (FunASR/Whisper)
3๏ธโƒฃ Smart Correction Context-based auto-correction of transcription errors
4๏ธโƒฃ Three-Dimensional Analysis TextContent + Viral + Brainstorming

๐Ÿš€ Quick Start

Prerequisites

# 1. yt-dlp (video downloader)
pip install yt-dlp

# 2. FFmpeg (must be installed and in PATH)
ffmpeg -version

# 3. Python dependencies
pip install pysrt python-dotenv

# 4. FunASR (Recommended for Chinese, lightweight & accurate)
pip install funasr modelscope

# 5. RapidOCR (ONNX lightweight, for burned subtitle detection)
pip install rapidocr-onnxruntime

# 6. Whisper (Alternative for English/multilingual)
pip install openai-whisper

# 7. requests (for Douyin download)
pip install requests

Usage

This is a Claude Skill designed for AI agents. Install it in your .agent/skills/ directory:

git clone https://github.com/ALBEDO-TABAI/video-copy-analyzer.git .agent/skills/video-copy-analyzer

Then use it with Claude:

โ€œAnalyze this video: https://www.bilibili.com/video/BV1xxxxxโ€

๐ŸŽฏ Smart Subtitle Extraction (3-Tier Priority)

The skill automatically selects the best extraction method:

Video Input
    โ†“
[1๏ธโƒฃ Embedded Subtitle] โ”€โ”€โ†’ Detected โ”€โ”€โ†’ Direct Extract (Highest Accuracy)
    โ†“ Not detected
[2๏ธโƒฃ Burned Subtitle OCR] โ”€โ”€โ†’ RapidOCR Frame Sampling โ”€โ”€โ†’ Detected โ”€โ”€โ†’ Full Video OCR
    โ†“ Not detected
[3๏ธโƒฃ Audio Transcription] โ”€โ”€โ†’ FunASR (Chinese optimized) / Whisper (Multilingual)
    โ†“
Output SRT Subtitles

Extraction Methods Comparison

Tier Method Use Case Accuracy Speed
L1 Embedded Extract Video has subtitle stream โญโญโญโญโญ โšก Fastest
L2 RapidOCR Subtitles burned into video โญโญโญโญ ๐Ÿš€ Fast
L3 FunASR Nano Chinese audio transcription โญโญโญโญ ๏ฟฝ Medium
L3 Whisper English/multilingual audio โญโญโญ ๐Ÿข Medium

Tech Stack

  • RapidOCR (ONNX): Lightweight OCR for burned subtitle detection

    • ๐Ÿš€ Lightweight: ONNX Runtime, no GPU required
    • ๐ŸŽฏ Cross-platform: Windows/Linux/Mac
    • ๐Ÿ“ฆ Easy deploy: Single pip install
    • โœจ High accuracy: Based on PaddleOCR
  • FunASR Nano: Alibaba open-source Chinese ASR model

    • ๐Ÿš€ Lightweight: ~100MB vs Whisper Large ~1.5GB
    • ๐ŸŽฏ Chinese optimized: Better than Whisper for Chinese
    • โฑ๏ธ Timestamp: Word-level timestamps
    • ๐Ÿ’จ Fast: Runs well on CPU

๏ฟฝ๐Ÿ“Š Three-Dimensional Analysis Framework

1. TextContent Analysis

  • Narrative structure breakdown
  • Rhetorical device identification
  • Keyword extraction

2. Viral-Abstract-Script Framework

  • Viral-5D Diagnosis: Hook / Emotion / Peaks / CTA / Social Currency
  • Style positioning
  • Optimization suggestions

3. Brainstorming Framework

  • Core value decomposition
  • 2-3 creative direction exploration
  • Incremental verification points

๐Ÿ“ Project Structure

video-copy-analyzer/
โ”œโ”€โ”€ SKILL.md                          # Core skill instructions
โ”œโ”€โ”€ scripts/
โ”‚   โ”œโ”€โ”€ download_douyin.py            # Douyin video downloader (watermark-free)
โ”‚   โ”œโ”€โ”€ extract_subtitle_funasr.py    # Smart subtitle extraction (FunASR + RapidOCR)
โ”‚   โ”œโ”€โ”€ extract_subtitle.py           # Whisper-based extraction
โ”‚   โ”œโ”€โ”€ transcribe_audio.py           # Audio transcription script
โ”‚   โ””โ”€โ”€ check_environment.py          # Environment verification
โ””โ”€โ”€ references/
    โ””โ”€โ”€ analysis-frameworks.md        # Analysis framework details

๐Ÿ”ง Configuration

On first use, the skill will prompt you to set a default output directory:

  • Option A: Use default ~/video-analysis/
  • Option B: Specify each time
  • Option C: Set a fixed custom directory

๐Ÿ“„ Output Files

After analysis, youโ€™ll receive:

File Content
{video_id}.mp4 Original video
{video_id}.srt Raw subtitles
{video_id}_transcript.md / {video_id}_ๆ–‡ๅญ—็จฟ.md Corrected transcript
{video_id}_analysis.md / {video_id}_ๅˆ†ๆžๆŠฅๅ‘Š.md Three-dimensional analysis report

๐ŸŽฏ Supported Environments

This is a Claude Skill that works with AI coding assistants:

Environment Model Status
Antigravity Gemini 3 Pro โœ… Supported
Cursor Claude 4.5 Opus โœ… Tested & Recommended
Claude Code Claude 4.5 Opus โœ… Supported
Windsurf Any Claude model โœ… Supported
Trae Claude 3.5/4 โœ… Supported

๐Ÿ’ก Best Performance: Tested with Claude 4.5 Opus, achieving optimal results in transcription correction and three-dimensional analysis.

๐Ÿ“ License

MIT License

View this README on GitHub

Recommended Tools

Try a different keyword or remove a filter.

Install

npx skillfish add albedo-tabai/video-copy-analyzer