Files
personal_development/video_transcription/ai_transcriber_v1/README.md
T

1.3 KiB

AI Video Transcriber & Translator

This tool extracts audio from videos, transcribes it using OpenAI's Whisper model, and translates the transcript using Google's Gemini API.

Setup

  1. Install Dependencies:

    pip install -r requirements.txt
    

    Note: You need ffmpeg installed on your system.

  2. API Key: Set your Gemini API key as an environment variable:

    export GEMINI_API_KEY="your_api_key_here"
    

Usage

Quick Start (Wizard)

For a user-friendly, interactive experience, run the wizard script in the root directory:

./run_wizard.py

This will guide you through selecting files, languages, and enabling features like cleanup and embedding.

Advanced (CLI)

Run the main.py script directly:

python ai_transcriber/main.py <path_to_video_or_folder> [options]

Options:

  • --model: Whisper model size (tiny, base, small, medium, large). Default: base.
  • --lang: Target language for translation. Default: English.
  • --force: Overwrite existing transcript/translation files.

Examples:

Single File:

python main.py ../my_video.mp4

Entire Directory (Recursive):

python main.py ../videos_folder/ --lang "Spanish" --model small