Initial release — OpenMontage: the first open-source agentic video production system
11 production pipelines, 47 tools, 124 agent skills.
Supports cloud APIs (fal.ai, OpenAI, ElevenLabs, Suno, HeyGen, Runway) and
free local providers (diffusers, Piper TTS, WAN 2.1, Hunyuan, CogVideo).
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-03-29 08:25:17 -07:00
|
|
|
# OpenMontage - Environment Variables
|
|
|
|
|
# Copy this to .env and fill in your keys
|
|
|
|
|
|
|
|
|
|
# --- fal.ai (one key unlocks the most tools) ---
|
|
|
|
|
FAL_KEY= # FLUX images, Google Veo video, Kling video, MiniMax video, Recraft images
|
|
|
|
|
# Get one at https://fal.ai/dashboard/keys
|
|
|
|
|
|
2026-03-29 09:06:26 -07:00
|
|
|
# --- Google (one key unlocks image gen + TTS) ---
|
|
|
|
|
GOOGLE_API_KEY= # Google Imagen images, Google Cloud TTS (700+ voices, 50+ languages)
|
|
|
|
|
# Get one at https://aistudio.google.com/apikey
|
|
|
|
|
|
Initial release — OpenMontage: the first open-source agentic video production system
11 production pipelines, 47 tools, 124 agent skills.
Supports cloud APIs (fal.ai, OpenAI, ElevenLabs, Suno, HeyGen, Runway) and
free local providers (diffusers, Piper TTS, WAN 2.1, Hunyuan, CogVideo).
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-03-29 08:25:17 -07:00
|
|
|
# --- Voice ---
|
|
|
|
|
ELEVENLABS_API_KEY= # TTS narration, music generation, sound effects
|
|
|
|
|
OPENAI_API_KEY= # OpenAI TTS fallback and DALL-E image generation
|
|
|
|
|
# Piper local voices do not require env vars; install `piper-tts` via pip
|
|
|
|
|
|
|
|
|
|
# --- Music ---
|
|
|
|
|
SUNO_API_KEY= # Suno AI music generation (full songs, instrumentals, any genre)
|
|
|
|
|
|
|
|
|
|
# --- Video Generation ---
|
|
|
|
|
HEYGEN_API_KEY= # HeyGen API (VEO, Sora, Runway, Kling, Seedance via single key)
|
|
|
|
|
RUNWAY_API_KEY= # Runway Gen-4 (direct API, alternative to fal.ai routing)
|
|
|
|
|
VIDEO_GEN_LOCAL_ENABLED= # Set to "true" for local video gen (needs GPU + diffusers)
|
|
|
|
|
VIDEO_GEN_LOCAL_MODEL= # Local model: wan2.1-1.3b, wan2.1-14b, hunyuan-1.5, ltx2-local, cogvideo-5b
|
|
|
|
|
MODAL_LTX2_ENDPOINT_URL= # Modal self-hosted LTX-2 endpoint (optional)
|
|
|
|
|
|
|
|
|
|
# --- Stock Media ---
|
|
|
|
|
PEXELS_API_KEY= # Pexels stock footage/images (free)
|
|
|
|
|
PIXABAY_API_KEY= # Pixabay stock footage/images (free)
|
|
|
|
|
|
|
|
|
|
# --- Analysis ---
|
|
|
|
|
HF_TOKEN= # HuggingFace token — enables speaker diarization in transcriber
|
|
|
|
|
|
|
|
|
|
# --- Avatar (local installs) ---
|
|
|
|
|
# WAV2LIP_PATH= # Path to cloned Wav2Lip repo (for lip sync)
|
|
|
|
|
# SADTALKER_PATH= # Path to cloned SadTalker repo (for talking head avatars)
|