> ## Documentation Index
> Fetch the complete documentation index at: https://docs.beeble.ai/llms.txt
> Use this file to discover all available pages before exploring further.

# SwitchX 1.0

> Professional video-to-video AI generation with SwitchX 1.0

<iframe width="100%" style={{ aspectRatio: "16/9" }} src="https://www.youtube.com/embed/sbqy6-XMGqM" title="SwitchX" frameBorder="0" allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowFullScreen />

<Note>
  SwitchX is available exclusively on **Beeble (Cloud app)**. It is not
  available in Beeble Studio.
</Note>

<Note>
  This page covers **SwitchX 1.0**, selectable from the **Model** dropdown in
  the input panel. The default model is [SwitchX 2.0](/beeble/switchx), which
  adds the 2160p resolution tier, 10-bit MOV output, a Camera Tracking
  control, and the **Iterate then Finish** workflow: explore in 720p, then
  finish your favorite take in 4K.
</Note>

**SwitchX** is a professional video-to-video AI generation tool designed for filmmakers, VFX artists, and creators. Unlike standard video-to-video models that hallucinate or alter the entire frame, SwitchX uses your original footage's pixels as direct control signals.

The core principle of SwitchX is simple: **Switch anything, keep what matters.**

* **Masked Areas:** Completely generated based on your prompts and reference images.
* **Unmasked Areas:** Retained from the original footage, but intelligently relit and restyled to blend seamlessly into the newly generated environment.

To get the best results from SwitchX, master two concepts: the **Alpha Mask** controls *what* gets changed, and the **Reference Image** controls *how* it looks.

***

## Alpha Masks

The alpha mask tells SwitchX exactly what to generate and what to preserve. SwitchX offers four masking modes:

| Mode       | Description                                                         | Best For                                               |
| ---------- | ------------------------------------------------------------------- | ------------------------------------------------------ |
| **Auto**   | Auto-detects and isolates the main subject. No manual input needed. | Background replacement, relighting, virtual production |
| **Select** | Manually select specific elements to keep or generate.              | Targeted changes (e.g., new outfit, inpainting)        |
| **Fill**   | Selects the entire frame. Keeps geometry, alters aesthetics.        | Full-frame relighting or restyling                     |
| **Upload** | Upload a custom alpha matte from external software.                 | Pixel-perfect control from Nuke, After Effects, etc.   |

### <Icon icon="wand-magic-sparkles" /> Auto Mode

<Info>
  **Best for:** Standard background replacements, relighting a subject, and
  virtual production.
</Info>

The engine automatically detects and isolates the main subject in your first frame, then propagates and tracks that mask throughout the rest of the video. No manual selection is required.

<img src="https://mintcdn.com/beebleai/rvysJVa9SnigtOU-/images/switchx-mask-auto.webp?fit=max&auto=format&n=rvysJVa9SnigtOU-&q=85&s=6f9d696580be6fcef03fa731721b2539" alt="SwitchX Mask - Auto mode" width="1024" height="421" data-path="images/switchx-mask-auto.webp" />

### <Icon icon="arrow-pointer" /> Select Mode

<Info>
  **Best for:** Precision control, changing specific elements (e.g., changing
  clothes while keeping the face and hands intact), and advanced inpainting.
</Info>

You manually choose exactly which parts of the frame to mask. Only the selected areas will be generated; everything else stays untouched from your original footage.

<img src="https://mintcdn.com/beebleai/rvysJVa9SnigtOU-/images/switchx-mask-select-result.webp?fit=max&auto=format&n=rvysJVa9SnigtOU-&q=85&s=0e4266888dee85bfc4ad5f379570e5a6" alt="SwitchX Mask - Select mode result" width="1024" height="419" data-path="images/switchx-mask-select-result.webp" />

Powered by an interactive AI masking tool (SAM3), click to select specific objects on the first frame. The Alpha Editor lets you fine-tune your selections:

<img src="https://mintcdn.com/beebleai/rvysJVa9SnigtOU-/images/switchx-mask-select.webp?fit=max&auto=format&n=rvysJVa9SnigtOU-&q=85&s=d768b47c53f4c5339ed007e597ac8d14" alt="SwitchX Mask - Select mode" width="1024" height="760" data-path="images/switchx-mask-select.webp" />

<Tip>
  * **Multiple Selections:** Select multiple objects. For the cleanest results,
    add distinct parts (like a face and a hand) as separate objects rather than
    grouping them. - **Deselect:** Right-click to remove an unintended selection.
  * **Invert:** Invert the mask of a specific object. For example, to turn a
    hand into a robot arm, select the hand, click **Invert**, and the engine will
    only generate within that area.
</Tip>

### <Icon icon="grid" /> Fill Mode

<Info>**Best for:** Complete scene relighting and restyling.</Info>

This mode selects the entire frame. Use Fill when you want to keep the original scene's geometry and composition intact but alter the overall lighting, mood, or artistic style.

<img src="https://mintcdn.com/beebleai/rvysJVa9SnigtOU-/images/switchx-mask-fill.webp?fit=max&auto=format&n=rvysJVa9SnigtOU-&q=85&s=90d0ddc0c3ec5d62b41c03d380323a22" alt="SwitchX Mask - Fill mode" width="1024" height="426" data-path="images/switchx-mask-fill.webp" />

### <Icon icon="arrow-up-from-bracket" /> Upload Mode

<Info>
  **Best for:** VFX professionals using external compositing software.
</Info>

Upload a precise alpha matte created in software like Nuke or After Effects. SwitchX reads the black-and-white alpha video pixel-by-pixel for absolute precision.

***

## Camera Tracking

SwitchX only sees the **unmasked** (foreground) region of your video. The masked area is completely hidden from the model, meaning SwitchX must infer all camera movement solely from the visible foreground pixels.

This has a direct impact on camera tracking quality:

* **Where it excels:** Shots where the foreground contains rich visual data to infer motion, such as a subject with complex movement, organic camera shake, or visible depth changes.
* **Where it struggles:** Simple, linear camera movements (like lateral trucking/panning shots) where the isolated foreground lacks sufficient parallax cues to estimate motion accurately.

<Warning>
  Even if your original footage contains tracking markers in the background (e.g., on a green screen), SwitchX **cannot use them** if those markers fall within the masked region. Once an area is masked, it becomes 100% invisible to the AI model, regardless of the tracking data present in your source footage.
</Warning>

***

## Reference Images

The reference image is your visual blueprint. It should contain all the details you want in your final video: background, lighting, mood, and costumes. SwitchX reads this image and uses it as a guide:

* **Masked areas:** SwitchX generates the background and content from the reference image into the masked region.
* **Unmasked areas:** SwitchX extracts the style and lighting from the reference image and applies it to your original footage, preserving the original pixels.

<img src="https://mintcdn.com/beebleai/CYElWM-DBOhfOQH2/images/switchx-source-mask-reference.webp?fit=max&auto=format&n=CYElWM-DBOhfOQH2&q=85&s=4a16f6b3777ecbb060b3782b783de762" alt="Source, mask, and reference combining into the SwitchX output" width="1920" height="790" data-path="images/switchx-source-mask-reference.webp" />

### Designing the Reference Image

This is the single most important thing to get right. The gap between a good reference and a great one is the gap between an okay result and a stunning one.

* **Show the subject and environment together.** SwitchX learned from references that look like the final shot; a background-only reference can't tell the model how to light your subject. Put the person in the frame, relit the way you want.
* **Any frame works: pick the most representative one.** It doesn't have to be the first frame. Find the frame that best represents the shot, then edit it into your target look.
* **Think across shots.** For multi-shot work, consistency between your reference images matters as much as the beauty of each one.

### Imperfect References Are Fine

Your reference image doesn't need to perfectly match your source footage. For example, if you generate a reference with Nano Banana and the face changes, that's fine. SwitchX understands what your original pixels are and will only bring the lighting and style from the reference, applying it on top of your actual footage.

<Columns cols={3}>
  <Frame caption="Source">
    <img src="https://mintcdn.com/beebleai/rvysJVa9SnigtOU-/images/switchx-ref-source.webp?fit=max&auto=format&n=rvysJVa9SnigtOU-&q=85&s=157c5b785e3cb248816e8dd6da70e9af" alt="Source" width="1024" height="539" data-path="images/switchx-ref-source.webp" />
  </Frame>

  <Frame caption="Reference Image">
    <img src="https://mintcdn.com/beebleai/rvysJVa9SnigtOU-/images/switchx-ref-reference.webp?fit=max&auto=format&n=rvysJVa9SnigtOU-&q=85&s=cd5b2b9f093ab3c0e4868726ee41a68b" alt="Reference Image" width="1024" height="576" data-path="images/switchx-ref-reference.webp" />
  </Frame>

  <Frame caption="SwitchX Result">
    <img src="https://mintcdn.com/beebleai/rvysJVa9SnigtOU-/images/switchx-ref-result.webp?fit=max&auto=format&n=rvysJVa9SnigtOU-&q=85&s=3b3d85ace127f3c0698dec7e1b95b8fa" alt="SwitchX Result" width="1024" height="539" data-path="images/switchx-ref-result.webp" />
  </Frame>
</Columns>

### Creating a Reference Image

You can upload any image, or use the built-in **Create with AI** tool: pick any frame from your video, then generate a matching reference with image models like Nano Banana, Seedream, or GPT Image.

<img src="https://mintcdn.com/beebleai/CYElWM-DBOhfOQH2/images/switchx-create-reference.webp?fit=max&auto=format&n=CYElWM-DBOhfOQH2&q=85&s=ff6609af6979855487b4da501a68d384" alt="Create Reference tool" width="3080" height="2000" data-path="images/switchx-create-reference.webp" />

### Prompting

Be highly specific: name the environment, lighting, and mood (e.g., *"a dramatic cliff in Ireland, soft overcast lighting, highly detailed props"*); vague prompts like *"take me to heaven"* yield poor results. If you're struggling with art direction, leave **Auto pilot** on and SwitchX writes the prompt from your reference image automatically.

<img src="https://mintcdn.com/beebleai/CYElWM-DBOhfOQH2/images/switchx-prompt-autopilot.webp?fit=max&auto=format&n=CYElWM-DBOhfOQH2&q=85&s=87c2af36d171c34bbe8d2bae095bc34b" alt="Prompt with Auto pilot enabled" width="656" height="306" data-path="images/switchx-prompt-autopilot.webp" />

### The Professional Workflow (Iteration)

For pixel-perfect results, do not rely solely on the initial AI-generated image:

1. **Generate** a reference image using the **Create with AI** tool.
2. **Download** the generated image to your local machine.
3. **Refine** in Photoshop or similar: adjust color grading, add props, fix details.
4. **Re-upload** the edited image into SwitchX. The final video will strictly follow this tailored reference.

***

## Settings & Export

### Resolution

| Setting          | Details                   |
| ---------------- | ------------------------- |
| **Resolution**   | 720p or 1080p             |
| **Max Frames**   | 240 frames (10s at 24fps) |
| **Output**       | 8-bit MP4                 |
| **Aspect Ratio** | Always preserved          |
| **Frame Rate**   | Always preserved          |

Resolution is a **maximum cap** on the shortest side: SwitchX 1.0 only caps down, so a source smaller than the selected tier renders at its native resolution. Aspect ratio and frame rate are never altered.

### Source Footage

Source quality carries into the output: compression artifacts in the input show up in the result, so feed SwitchX the highest-quality source you have.

* **Native ProRes ingest:** upload your ProRes 422 (Proxy, LT, 422, HQ) `.mov` master directly, up to **4.5GB**. No in-app trimming or playback, so trim before uploading. ProRes 4444 (embedded alpha) and variable-frame-rate exports are rejected, and Continue in Canvas is unavailable for ProRes jobs.
* **High-bitrate H.264 works just as well**, visually near-identical to a ProRes master. The quality loss usually blamed on H.264 comes from low-bitrate exports, not the codec.

Recommended transcode settings (Resolve, Compressor, or Shutter Encoder):

* Codec: H.264
* Bitrate: 40–60 Mbps at 1080p (scale up proportionally for 4K), or CRF 16 if your tool supports it
* Keep the original resolution and frame rate
* Export in Rec.709 (apply your display LUT first if the footage is Log)

The equivalent ffmpeg command:

```bash theme={null}
ffmpeg -i input.mov -c:v libx264 -preset slow -crf 16 -pix_fmt yuv420p -c:a aac -b:a 256k output.mp4
```
