Skip to main content
SwitchX is available exclusively on Beeble (Cloud app). It is not available in Beeble Studio.
SwitchX is a professional video-to-video AI generation tool designed for filmmakers, VFX artists, and creators. Unlike standard video-to-video models that hallucinate or alter the entire frame, SwitchX uses your original footage’s pixels as direct control signals. The core principle of SwitchX is simple: Switch anything, keep what matters.
  • Masked Areas: Completely generated based on your prompts and reference images.
  • Unmasked Areas: Retained from the original footage, but intelligently relit and restyled to blend seamlessly into the newly generated environment.
To get the best results from SwitchX, master two concepts: the Alpha Mask controls what gets changed, and the Reference Image controls how it looks.

Model Versions

SwitchX ships in two versions, selectable from the Model dropdown at the top of the input panel. SwitchX 2.0 is the default and the recommended model; SwitchX 1.0 remains available and has its own dedicated page. SwitchX Model dropdown
2160p generation and the 10-bit MOV download require a Professional or Team plan.

Source Footage

Source quality carries into the output: compression artifacts in the input show up in the result, so feed SwitchX the highest-quality source you have.
  • Native ProRes ingest: upload your ProRes 422 (Proxy, LT, 422, HQ) .mov master directly, up to 4.5GB. No in-app trimming or playback, so trim before uploading. ProRes 4444 (embedded alpha) and variable-frame-rate exports are rejected, and Continue in Canvas is unavailable for ProRes jobs.
  • High-bitrate H.264 works just as well, visually near-identical to a ProRes master. The quality loss usually blamed on H.264 comes from low-bitrate exports, not the codec.
Recommended transcode settings (Resolve, Compressor, or Shutter Encoder):
  • Codec: H.264
  • Bitrate: 40–60 Mbps at 1080p (scale up proportionally for 4K), or CRF 16 if your tool supports it
  • Keep the original resolution and frame rate
  • Export in Rec.709 (apply your display LUT first if the footage is Log)
The equivalent ffmpeg command:

Alpha Masks

The alpha mask tells SwitchX exactly what to generate and what to preserve. SwitchX offers four masking modes:

Auto Mode

Best for: Standard background replacements, relighting a subject, and virtual production.
The engine automatically detects and isolates the main subject in your first frame, then propagates and tracks that mask throughout the rest of the video. No manual selection is required. SwitchX Mask - Auto mode

Select Mode

Best for: Precision control, changing specific elements (e.g., changing clothes while keeping the face and hands intact), and advanced inpainting.
You manually choose exactly which parts of the frame to mask. Only the selected areas will be generated; everything else stays untouched from your original footage. SwitchX Mask - Select mode result Powered by interactive AI masking tools (SAM3, MatAnyone), choose a frame and click to select specific objects on it. The Alpha Editor lets you fine-tune your selections: SwitchX Mask - Select mode
  • Multiple Selections: Select multiple objects. For the cleanest results, add distinct parts (like a face and a hand) as separate objects rather than grouping them.
  • Deselect: Right-click to remove an unintended selection.
  • Invert: Invert the mask of a specific object. For example, to turn a hand into a robot arm, select the hand, click Invert, and the engine will only generate within that area.

Fill Mode

Best for: Complete scene relighting and restyling.
This mode selects the entire frame. Use Fill when you want to keep the original scene’s geometry and composition intact but alter the overall lighting, mood, or artistic style. SwitchX Mask - Fill mode

Upload Mode

Best for: VFX professionals using external compositing software.
Upload a precise alpha matte created in software like Nuke or After Effects. SwitchX reads the black-and-white alpha video pixel-by-pixel for absolute precision.

Camera Tracking

SwitchX 2.0 supports camera tracking natively: the engine tracks the camera motion of your source footage and applies it to the newly generated environment, masked areas included. Control it under Advanced settings:
  • On (Follow Source): the default. The generated environment follows your source’s camera motion.
  • Off (Ignore Source): the generated environment ignores the source’s camera motion. Use this when you want the new background to stay stable regardless of how the camera moves.
SwitchX 1.0 has no built-in tracking — it infers camera movement from the unmasked pixels: SwitchX 1.0 § Camera Tracking.
Every environment below is fully generated, yet follows the source camera exactly — all with Camera Tracking On.

Truck (Lateral Tracking)

The camera travels sideways with the subjects — the generated street scrolls past with the source’s speed and parallax.

Source

SwitchX 2.0

Pan (Follow)

The camera swings to follow the bike through the corner — the generated night circuit stays locked to the move even at race pace.

Source

SwitchX 2.0

Push-in (Dolly)

The camera pushes forward past the foreground shelf — the generated aquarium keeps the same approach and parallax.

Source

SwitchX 2.0


Fast Motion

SwitchX 2.0 holds detail through fast movement — lip sync, dance, gesture — where video models tend to smear or drift off the performance.

Lip Sync

Every syllable lands after the switch from greenscreen to jungle — wardrobe and world change, the mouth doesn’t.

Source

SwitchX 2.0

Dance (Full-Body Motion)

Subway-platform floorwork relit from flat fluorescents to amber and signal green — the source movement carries through unchanged.

Source

SwitchX 2.0

Gesture (Fast Hands)

A crossover dribble at full speed, switched from a street court to an indoor gym — the ball and both players stay sharp.

Source

SwitchX 2.0


Reference Images

The reference image is your visual blueprint. It should contain all the details you want in your final video: background, lighting, mood, and costumes. SwitchX reads this image and uses it as a guide:
  • Masked areas: SwitchX generates the background and content from the reference image into the masked region.
  • Unmasked areas: SwitchX extracts the style and lighting from the reference image and applies it to your original footage, preserving the original pixels.
Source, mask, and reference combining into the SwitchX output

Designing the Reference Image

This is the single most important thing to get right. The gap between a good reference and a great one is the gap between an okay result and a stunning one.
  • Show the subject and environment together. A background-only reference can’t tell the model how to light your subject — put the person in the frame, relit the way you want.
  • Pick the most representative frame (it doesn’t have to be the first), then edit it into your target look.
  • Think across shots. For multi-shot work, consistency between your reference images matters as much as the beauty of each one.
You can upload any image, or use the built-in Create with AI tool: pick a frame from your video and generate a matching reference with image models like Nano Banana, Seedream, or GPT Image. Create Reference tool For full control, build the reference in Canvas: rebuild the environment as a 3D scene, composite the subject in, and finish with a photoreal integration pass. Canvas node graph building a reference image: source footage into a 3D environment, the foreground subject isolated and composited in, then a photoreal integration pass feeding SwitchX

Imperfect References Are Fine

Your reference image doesn’t need to perfectly match your source footage. For example, if you generate a reference with Nano Banana and the face changes, that’s fine. SwitchX understands what your original pixels are and will only bring the lighting and style from the reference, applying it on top of your actual footage.
Source

Source

Reference Image

Reference Image

SwitchX Result

SwitchX Result

Prompting

Be highly specific: name the environment, lighting, and mood (e.g., “a dramatic cliff in Ireland, soft overcast lighting, highly detailed props”); vague prompts like “take me to heaven” yield poor results. If you’re struggling with art direction, leave Auto pilot on and SwitchX writes the prompt from your reference image automatically. Prompt with Auto pilot enabled

The Professional Workflow (Iteration)

For pixel-perfect results, do not rely solely on the initial AI-generated image:
  1. Generate a reference image using the Create with AI tool.
  2. Download the generated image to your local machine.
  3. Refine in Photoshop or similar: adjust color grading, add props, fix details.
  4. Re-upload the edited image into SwitchX. The final video will strictly follow this tailored reference.

Settings & Export

Each resolution tier is a maximum cap on the shortest side: SwitchX 1.0 only caps down, while SwitchX 2.0 can also enlarge a smaller source (those options are tagged Upscale in the picker).

Iterate then Finish

SwitchX 2.0 introduces a new way to work: explore faster for less in 720p, pick your favorite, then Finish it in 4K. More than an upscale: Finish re-renders your take from the original source footage, recovering source detail to match a native 4K generation. Iterate at 720p/1080p, then Finish reads the source again for the final render