ComfyUI's engine, without the engine room — AI image and video made easy, free, and open source.
Website · Documentation · Download · Discord · Roadmap · Patreon · Gumroad
Cubric Vision is a desktop workspace for generating images and video on your own machine. It runs ComfyUI as its engine — curated models, tuned workflows, no node graphs to wire up. You type a prompt, pick a model, and refine the result with masking, detailing, upscaling, and video tools. Your prompts, images, videos, and project files stay on your disk.
Free, open source, and made by Mad Pony Interactive. No accounts. No API fees. Runs on your machine — remote GPU optional.

![]() | ![]() |
![]() | ![]() |
Text-to-video and image-to-video from the latest open-source video models — runs locally, staged previews so you spend compute only on shots worth finishing. Some models generate with sound and support first- and last-frame guidance.
https://github.com/user-attachments/assets/dd4c46e8-6941-49a2-b279-949fa1815aae
Grab the portable build for your platform from GitHub Releases — no installer, just extract and launch. On first run the app sets up its ComfyUI engine and asks where to store models.
Windows: extract the zip and run CubricVision.exe. Windows will show
"Windows protected your PC" — click More info, then Run anyway. The
builds are not code-signed, so this warning is expected.
macOS: downloaded builds are quarantined. Clear it, then launch:
xattr -dr com.apple.quarantine "<extracted folder>"
Then double-click start.command.
Linux: extract the tarball and run ./start.sh. Engine setup needs
git — the installer will offer to install it if it's missing.
Step-by-step instructions are in the installation guide.
Every build is free and public on GitHub Releases. If Cubric Vision is useful to you, you can fund its development on Patreon (recurring) or Gumroad (one-off) — support, not a paywall.
Windows 10/11, Linux, or macOS 14 (Sonoma) or later on Apple Silicon — every Mac with an M-series chip can run macOS 14 or newer.
| Workflow | GPU VRAM | System RAM |
|---|---|---|
| Images | 8 GB+ | 16–32 GB |
| Video and the heaviest image models | 12–16 GB+ | 32–64 GB |
Each model shows its own memory needs in the app, and lighter tiers are available for smaller machines. If a model is too heavy for your GPU, you can run it on a rented remote GPU instead.
The full user guide lives at docs.cubric.studio: getting started, projects, prompt box, image tools, video tools, models, gallery, history, and hotkeys.
Vision is the first app in the Cubric Studio family. Audio and Prompt are planned siblings — all local, all open source.
Want to run from source or contribute? Start with docs/DEVELOPMENT.md, then read CONTRIBUTING.md before opening a PR. For security-sensitive reports, see SECURITY.md.
Cubric Vision is licensed under AGPL-3.0-only. Portable builds ship with readable app source — open source isn't a marketing line here, it's how the app is distributed.
3,728 commits
JavaScript
85.7%
Python
7.5%
CSS
4.9%
HTML
1.1%
ComfyUI's engine, without the engine room — AI image and video made easy, free, and open source.
Website · Documentation · Download · Discord · Roadmap · Patreon · Gumroad
Cubric Vision is a desktop workspace for generating images and video on your own machine. It runs ComfyUI as its engine — curated models, tuned workflows, no node graphs to wire up. You type a prompt, pick a model, and refine the result with masking, detailing, upscaling, and video tools. Your prompts, images, videos, and project files stay on your disk.
Free, open source, and made by Mad Pony Interactive. No accounts. No API fees. Runs on your machine — remote GPU optional.

![]() | ![]() |
![]() | ![]() |
Text-to-video and image-to-video from the latest open-source video models — runs locally, staged previews so you spend compute only on shots worth finishing. Some models generate with sound and support first- and last-frame guidance.
https://github.com/user-attachments/assets/dd4c46e8-6941-49a2-b279-949fa1815aae
Grab the portable build for your platform from GitHub Releases — no installer, just extract and launch. On first run the app sets up its ComfyUI engine and asks where to store models.
Windows: extract the zip and run CubricVision.exe. Windows will show
"Windows protected your PC" — click More info, then Run anyway. The
builds are not code-signed, so this warning is expected.
macOS: downloaded builds are quarantined. Clear it, then launch:
xattr -dr com.apple.quarantine "<extracted folder>"
Then double-click start.command.
Linux: extract the tarball and run ./start.sh. Engine setup needs
git — the installer will offer to install it if it's missing.
Step-by-step instructions are in the installation guide.
Every build is free and public on GitHub Releases. If Cubric Vision is useful to you, you can fund its development on Patreon (recurring) or Gumroad (one-off) — support, not a paywall.
Windows 10/11, Linux, or macOS 14 (Sonoma) or later on Apple Silicon — every Mac with an M-series chip can run macOS 14 or newer.
| Workflow | GPU VRAM | System RAM |
|---|---|---|
| Images | 8 GB+ | 16–32 GB |
| Video and the heaviest image models | 12–16 GB+ | 32–64 GB |
Each model shows its own memory needs in the app, and lighter tiers are available for smaller machines. If a model is too heavy for your GPU, you can run it on a rented remote GPU instead.
The full user guide lives at docs.cubric.studio: getting started, projects, prompt box, image tools, video tools, models, gallery, history, and hotkeys.
Vision is the first app in the Cubric Studio family. Audio and Prompt are planned siblings — all local, all open source.
Want to run from source or contribute? Start with docs/DEVELOPMENT.md, then read CONTRIBUTING.md before opening a PR. For security-sensitive reports, see SECURITY.md.
Cubric Vision is licensed under AGPL-3.0-only. Portable builds ship with readable app source — open source isn't a marketing line here, it's how the app is distributed.
3,728 commits
JavaScript
85.7%
Python
7.5%
CSS
4.9%
HTML
1.1%