Initial release — OpenMontage: the first open-source agentic video production system
11 production pipelines, 47 tools, 124 agent skills.
Supports cloud APIs (fal.ai, OpenAI, ElevenLabs, Suno, HeyGen, Runway) and
free local providers (diffusers, Piper TTS, WAN 2.1, Hunyuan, CogVideo).
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-03-29 08:25:17 -07:00
|
|
|
# OpenMontage - Environment Variables
|
|
|
|
|
# Copy this to .env and fill in your keys
|
|
|
|
|
|
2026-04-08 13:04:32 -07:00
|
|
|
# --- Image + video gateway ---
|
Initial release — OpenMontage: the first open-source agentic video production system
11 production pipelines, 47 tools, 124 agent skills.
Supports cloud APIs (fal.ai, OpenAI, ElevenLabs, Suno, HeyGen, Runway) and
free local providers (diffusers, Piper TTS, WAN 2.1, Hunyuan, CogVideo).
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-03-29 08:25:17 -07:00
|
|
|
FAL_KEY= # FLUX images, Google Veo video, Kling video, MiniMax video, Recraft images
|
|
|
|
|
# Get one at https://fal.ai/dashboard/keys
|
|
|
|
|
|
2026-03-29 09:06:26 -07:00
|
|
|
# --- Google (one key unlocks image gen + TTS) ---
|
|
|
|
|
GOOGLE_API_KEY= # Google Imagen images, Google Cloud TTS (700+ voices, 50+ languages)
|
|
|
|
|
# Get one at https://aistudio.google.com/apikey
|
2026-06-22 17:53:10 +05:30
|
|
|
# Alternative to the API key: service-account JSON auth.
|
|
|
|
|
# TTS uses Cloud Text-to-Speech; Imagen routes to Vertex AI.
|
|
|
|
|
GOOGLE_APPLICATION_CREDENTIALS= # path to a service-account JSON key file
|
|
|
|
|
GOOGLE_CLOUD_PROJECT= # GCP project id (required for Imagen via Vertex AI)
|
|
|
|
|
GOOGLE_CLOUD_LOCATION= # Vertex AI region, default us-central1
|
2026-03-29 09:06:26 -07:00
|
|
|
|
Initial release — OpenMontage: the first open-source agentic video production system
11 production pipelines, 47 tools, 124 agent skills.
Supports cloud APIs (fal.ai, OpenAI, ElevenLabs, Suno, HeyGen, Runway) and
free local providers (diffusers, Piper TTS, WAN 2.1, Hunyuan, CogVideo).
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-03-29 08:25:17 -07:00
|
|
|
# --- Voice ---
|
|
|
|
|
ELEVENLABS_API_KEY= # TTS narration, music generation, sound effects
|
|
|
|
|
OPENAI_API_KEY= # OpenAI TTS fallback and DALL-E image generation
|
2026-04-05 15:31:37 -07:00
|
|
|
XAI_API_KEY= # Grok image generation/editing and Grok video generation
|
2026-05-04 21:40:17 +08:00
|
|
|
DOUBAO_SPEECH_API_KEY= # Volcengine Doubao Speech TTS (new console API Key)
|
|
|
|
|
DOUBAO_SPEECH_VOICE_TYPE= # Default Doubao speaker/voice type, e.g. zh_female_vv_uranus_bigtts
|
Initial release — OpenMontage: the first open-source agentic video production system
11 production pipelines, 47 tools, 124 agent skills.
Supports cloud APIs (fal.ai, OpenAI, ElevenLabs, Suno, HeyGen, Runway) and
free local providers (diffusers, Piper TTS, WAN 2.1, Hunyuan, CogVideo).
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-03-29 08:25:17 -07:00
|
|
|
# Piper local voices do not require env vars; install `piper-tts` via pip
|
|
|
|
|
|
|
|
|
|
# --- Music ---
|
|
|
|
|
SUNO_API_KEY= # Suno AI music generation (full songs, instrumentals, any genre)
|
|
|
|
|
|
|
|
|
|
# --- Video Generation ---
|
|
|
|
|
HEYGEN_API_KEY= # HeyGen API (VEO, Sora, Runway, Kling, Seedance via single key)
|
|
|
|
|
RUNWAY_API_KEY= # Runway Gen-4 (direct API, alternative to fal.ai routing)
|
|
|
|
|
VIDEO_GEN_LOCAL_ENABLED= # Set to "true" for local video gen (needs GPU + diffusers)
|
|
|
|
|
VIDEO_GEN_LOCAL_MODEL= # Local model: wan2.1-1.3b, wan2.1-14b, hunyuan-1.5, ltx2-local, cogvideo-5b
|
|
|
|
|
MODAL_LTX2_ENDPOINT_URL= # Modal self-hosted LTX-2 endpoint (optional)
|
|
|
|
|
|
|
|
|
|
# --- Stock Media ---
|
|
|
|
|
PEXELS_API_KEY= # Pexels stock footage/images (free)
|
|
|
|
|
PIXABAY_API_KEY= # Pixabay stock footage/images (free)
|
2026-04-10 16:42:39 -07:00
|
|
|
UNSPLASH_ACCESS_KEY= # Unsplash stock images (free developer key)
|
Initial release — OpenMontage: the first open-source agentic video production system
11 production pipelines, 47 tools, 124 agent skills.
Supports cloud APIs (fal.ai, OpenAI, ElevenLabs, Suno, HeyGen, Runway) and
free local providers (diffusers, Piper TTS, WAN 2.1, Hunyuan, CogVideo).
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-03-29 08:25:17 -07:00
|
|
|
|
|
|
|
|
# --- Analysis ---
|
|
|
|
|
HF_TOKEN= # HuggingFace token — enables speaker diarization in transcriber
|
|
|
|
|
|
|
|
|
|
# --- Avatar (local installs) ---
|
|
|
|
|
# WAV2LIP_PATH= # Path to cloned Wav2Lip repo (for lip sync)
|
|
|
|
|
# SADTALKER_PATH= # Path to cloned SadTalker repo (for talking head avatars)
|