Documentation
Vampro Voice Studio for Adobe Premiere Pro
Vampro Voice Studio is an integrated Premiere Pro panel powered by a local Windows companion service. It enables custom voice cloning, speech-to-speech conversion, text-to-speech synthesis, dialogue cleaning, speaker grouping, and vocal/music stem separation without leaving your editing environment.
Before You Begin
Ensure the following baseline configuration is in place before initiating generation tasks:
- Install and launch the Vampro Voice Studio Windows companion application.
- Open Adobe Premiere Pro and access the Voice Studio extension panel (Window > Extensions > Vampro Voice Studio).
- The panel establishes a secure localhost connection to the companion service on 127.0.0.1.
- Ensure adequate disk storage on your system volume (%LOCALAPPDATA%\Vampro\VoiceStudio) for rendered stems, voice library references, and project cache.
- For 100% offline generation, confirm the English Chatterbox model pack has finished initialization.
Choose How to Process
Voice Studio provides two dedicated execution pipelines selectable directly from the top mode switch:
1. Local Offline Studio
Runs the English Chatterbox GGUF model and native audio workers directly on your workstation. Zero network traffic, zero third-party subscriptions, and complete data isolation. Utilizes GPU acceleration (Vulkan) with automatic CPU fallback.
2. ElevenLabs API (BYOK)
Connect your personal ElevenLabs API key for access to community models and cloud voice clones. Your API key is encrypted locally using Windows DPAPI (CryptProtectData). Cloud jobs are only dispatched when explicitly confirmed.
Add a Voice Reference (Cloning)
Build custom voice clones from source media in three steps:
- Select Source Media: In the Studio tab, select Add Voice and choose from Timeline Clip, Bin Clip, or Upload File (supports both WAV/MP3 audio and video containers).
- Define Sample Range: Set the in and out points to isolate a clean 5–10 second spoken passage without background music or audible overlap. Voice references must be between 1 and 60 seconds.
- Save to Voice Library: Enter an identifiable profile name and choose Save Voice. The reference is indexed in your local library under
%LOCALAPPDATA%\Vampro\VoiceStudio\voices.
Change a Clip's Voice (Speech-to-Speech)
Reshape a spoken performance into a target voice profile while retaining pacing, cadence, and inflection:
- Select a target voice from your Voice Library.
- Choose the source performance from a Timeline Clip, Project Bin Asset, File Upload, or live Microphone Take.
- Click Change Voice to synthesize the converted take.
- Preview the result directly in the integrated audio waveform player.
- Use "Import to Bin", "At Playhead", or "Apply to Original" to bring the file into Premiere Pro. When applying to an original timeline clip, Premiere undo history is preserved.
Create Speech & Takes
Text to Speech
Select Text, pick a voice profile, type or paste up to 5,000 characters of English script, and click Generate. Takes are rendered in mono 24 kHz broadcast-ready PCM WAV format.
Take Management
Generated files appear in the Preview player, job queue, and Takes catalog with duration, timestamp, and model metadata. Completed takes can be queued or deleted at any time.
Audio Repair & Stem Separation
Access production cleanup tools from the Repair tab:
Spectral Noise Reduction
Removes stationary noise profiles, HVAC rumble, and ambient hum from recordings using local spectral gating algorithms.
Speaker Diarization
Estimates speech turns and segments audio into likely speakers (Speaker A, Speaker B) for multi-mic or interview editing. Any speaker segment can be extracted into a Voice Library reference.
Two-Stem Vocal & Accompaniment Separation
Powered by the Sherpa-ONNX runner and Spleeter two-stem model. Isolate vocals and background instrumentation into discrete WAV files for rebalancing in your timeline.
Feedback & Tuning
Rate completed takes with the positive/negative feedback control. Feedback is stored strictly in your local database (studio.db). For offline voice conversion, positive ratings subtly tune the voice-strength parameter for subsequent jobs on that reference. No media or training telemetry is ever transmitted to Vampro servers.
Files, Privacy & Local Storage
Voice Studio stores all operational assets on your local Windows storage under:
%LOCALAPPDATA%\Vampro\VoiceStudio
Subdirectories include audio/, voices/, results/, stems/, and premiere-media/.
To remove all local Voice Studio data, terminate the companion application from the taskbar and delete the %LOCALAPPDATA%\Vampro\VoiceStudio directory.
System Limits & Technical Specifications
| Parameter | Specification | Notes |
|---|---|---|
| Supported OS | Windows 10 / 11 64-bit | macOS support in development |
| Host Application | Adobe Premiere Pro 24.0, 25.0, 26.0+ | Requires UXP plugin architecture |
| Supported Languages | English (Local Model) | Multilingual supported via ElevenLabs API |
| Voice Reference Length | 1 to 60 seconds | First 9 seconds utilized by local acoustic encoder |
| Max Job Duration | Up to 10 minutes per clip | Single continuous audio source per job |
| Hardware Acceleration | NVIDIA / AMD / Intel (Vulkan) | Multi-threaded CPU fallback enabled |
| Output Audio Format | 24 kHz 16-bit Mono PCM WAV | Broadcast compliant |