1.3 KiB
1.3 KiB
AI Video Transcriber & Translator
This tool extracts audio from videos, transcribes it using OpenAI's Whisper model, and translates the transcript using Google's Gemini API.
Setup
-
Install Dependencies:
pip install -r requirements.txtNote: You need
ffmpeginstalled on your system. -
API Key: Set your Gemini API key as an environment variable:
export GEMINI_API_KEY="your_api_key_here"
Usage
Quick Start (Wizard)
For a user-friendly, interactive experience, run the wizard script in the root directory:
./run_wizard.py
This will guide you through selecting files, languages, and enabling features like cleanup and embedding.
Advanced (CLI)
Run the main.py script directly:
python ai_transcriber/main.py <path_to_video_or_folder> [options]
Options:
--model: Whisper model size (tiny,base,small,medium,large). Default:base.--lang: Target language for translation. Default:English.--force: Overwrite existing transcript/translation files.
Examples:
Single File:
python main.py ../my_video.mp4
Entire Directory (Recursive):
python main.py ../videos_folder/ --lang "Spanish" --model small