Files
personal_development/video_transcription/ai_transcriber_v2
2026-01-11 15:10:56 -05:00
..
2026-01-11 15:10:56 -05:00
2026-01-11 15:10:56 -05:00
2026-01-11 15:10:56 -05:00
2026-01-11 15:10:56 -05:00
2026-01-11 15:10:56 -05:00
2026-01-11 15:10:56 -05:00
2026-01-11 15:10:56 -05:00
2026-01-11 15:10:56 -05:00
2026-01-11 15:10:56 -05:00
2026-01-11 15:10:56 -05:00
2026-01-11 15:10:56 -05:00

AI Video Transcriber & Translator

This tool extracts audio from videos, transcribes it using OpenAI's Whisper model, and translates the transcript using Google's Gemini API.

Setup

  1. Install Dependencies:

    pip install -r requirements.txt
    

    Note: You need ffmpeg installed on your system.

  2. API Key: Set your Gemini API key as an environment variable:

    export GEMINI_API_KEY="your_api_key_here"
    

Usage

Linux / Mac

Run the wizard script:

./run_v2.py

Windows

  1. Install FFmpeg: Download from ffmpeg.org and add the bin folder to your System PATH.
  2. Run: Double-click run_v2.bat.
    • It will automatically create the virtual environment, install dependencies, and launch the tool.

Manual CLI

python ai_transcriber_v2/main.py <path_to_video> [options]

Options:

  • --model: Whisper model size (tiny, base, small, medium, large). Default: base.
  • --lang: Target language for translation. Default: English.
  • --force: Overwrite existing transcript/translation files.

Examples:

Single File:

python main.py ../my_video.mp4

Entire Directory (Recursive):

python main.py ../videos_folder/ --lang "Spanish" --model small