perfectgf/lora-dataset-studio

Self-hosted, one-tab workbench for the whole training lifecycle: build Character/Concept/Style datasets (generate, scrape, or triage a big dump in the Image Bank), curate, caption, clean watermarks, then train LoRAs on five model families or a full Krea 2 model on a rented GPU, and rank checkpoints in a Test Studio. Free & local.

Python

295

28 commits

updated Sep 29, 2026

See the code

See what people are saying

SourceMessageScoreDate

ai slop game based on batlle video in real time with only open model (r/StableDiffusion)

Check out [Chimera Arena (https://chimeraarena.com/static/showcase/index.html?lang=en)](https://chimeraarena.com/static/showcase/index.html?lang=en)! Here’s my slightly scrappy AI-generated teaser for Chimera Arena, my card game combining AI-generated video battles, real-time 2D combat, and…

16

Oct 3, 2026

README

Support us on Patreon Join our Discord

[!WARNING] No affiliation with loradataset.com. That website is not operated, endorsed, or supported by the LoRA Dataset Studio team. Any payments made to that service do not support this project. Use this repository and the links provided here to find our official downloads and community channels.

LoRA Dataset Studio V2

Build, curate, caption and train image datasets from one browser interface. LDS runs on your machine and uses ai-toolkit for training and ComfyUI for local generation and checkpoint testing.

Install · Documentation · Plugins · Releases · Discord

The core and public plugins are free under the PolyForm Noncommercial license. The core needs no account; optional usage statistics are off by default. External APIs and rented GPUs have their own charges. Additional paid plugins may be offered later.

Generation engines
Generation engines
Dataset workspace
Dataset workspace

A complete Character LoRA workflow in seven minutes:

https://github.com/user-attachments/assets/d51ff89c-34e9-41a9-b47d-08939a8c867b

People shown in the demo and screenshots are AI-generated.

What it does

AreaCapabilities
DatasetsCharacter, Concept and Style workflows; image and ZIP/folder import; subject-aware shot catalogs; local Klein and Krea 2 Edit generation; reference editing and exact retries
Image BankReview folders in place, score quality, group duplicates and people, search by text or similarity, and build balanced shortlists to promote into datasets
CurationKeep/reject, crop, mirror, rotate, face similarity, composition checks, editable watermark masks, text detection and recoverable cleaning
CaptioningLocal vision models or JoyCaption, family-appropriate prose or tags, appearance policies, bulk editing, targeted re-captioning and external image/.txt round trips
Local trainingGuided ai-toolkit recipes for Z-Image, SDXL, Krea 2, FLUX.1, FLUX.2 Klein, Anima and Qwen-Image 2.1; queues, advanced settings, checkpoint continuation and experimental slider LoRAs
Review and testingGenerate with all seven image training families; fixed-seed checkpoint/strength comparisons, prompt batches, runs, logs, lineage, votes, rankings and a generated-image Gallery
FilesStandard training ZIPs and sidecars, backup/restore without API keys, ComfyUI deployment, configurable storage and Trash

Dependencies vary by feature. Bank search ranks matches; it does not guarantee exclusions. Undo covers specific actions, and Delete rejected can remove source files after confirmation. Video and slider workflows are experimental. See the feature reference, requirements and known limitations for the detailed behavior.

Plugins

Install optional features from Plugins → Store, configure and prepare them in their own settings, and update them through Plugins → Updates. Several plugins can be installed together with one LDS restart. Each is independently installable; model downloads, hardware and provider credentials depend on the feature.

PluginWhat it doesMain requirements
API image enginesGenerate images with Gemini/Nano Banana, ChatGPT or OpenRouter; experimental ChatGPT subscription connectionProvider credentials; API usage is billed by the provider
Camera anglesCreate new viewpoints of Gallery or dataset images with controls for direction, height and distanceComfyUI, Qwen-Image-Edit and Multiple-Angles LoRA
CanvasArrange runs and checkpoints on a visual board, compare or blend LoRAs, generate images and export layoutsComfyUI for generation; one model family per generation batch
Cloud trainingTrain image or video models on rented GPUs, monitor runs, recover checkpoints and deliver full Krea 2 modelsvast.ai account and API key; Hugging Face access for applicable runs
DLSS 5 Neural RenderingImprove a finished video's lighting and materials, compare with the original and export the resultWindows x86-64, NVIDIA driver, compatible model supplied by you and the prepared bridge
Klein ImproveRe-render image detail with instructions, LoRA presets and finishing controlsComfyUI and Klein models; appearance and color can change
Live channelsContinuously render scripted H3 scenes and watch them in a browser or VLCCompatible local ComfyUI, H3 weights and stream encoder; experimental
Model toolsQuantize full models to fp8 or merge weighted LoRAs into a base checkpointPython with PyTorch and output disk space; CPU processing, no GPU required
Publish to CivitaiUpload checkpoints and generated images to a model pageCivitai API key; checkpoints default to drafts, image posts default to publication
Publish to Hugging FaceExport kept images, captions, metadata and a dataset card to the HubWrite-enabled Hugging Face token; private repository by default and rights confirmation
Resource monitorShow CPU, GPU, RAM, VRAM and temperatures, with guarded memory releaseNo model download
SeedVR2Restore and upscale images with high-resolution tiling and finishing controlsComfyUI, SeedVR2 nodes and models; optional TTP nodes for tiling
Video laneBuild video datasets from local files or web imports, curate shots, train LoRAs and test H3 generation, continuation and interpolationDependencies vary by task: PyAV/ffmpeg, ai-toolkit, ComfyUI and model weights; Video Bank remains beta
Web scrapingImport selected images or clips from searches and supported gallery URLs into datasets or banksSource-dependent credentials and permissions; Pexels requires explicit dataset/ML authorization

API providers apply their own billing and content policies. Model licenses also apply, including MiniMax H3's territory restrictions; check the video limits before using it. Plugin authors can start with the SDK and package guide.

Screenshots

Click a screenshot to view it at full size.

Image Bank
Image Bank
Analysis and coverage
Analysis and coverage
Batch analysis
Batch analysis
Dataset curation
Dataset curation
Text detection
Text detection
Training runs
Training runs
Checkpoint comparisons
Checkpoint comparisons
LoRA Canvas
LoRA Canvas
Civitai prompts
Civitai prompts
Camera angles
Camera angles
Video Test Studio
Video Test Studio
DLSS 5 comparison
DLSS 5 comparison

Setup & install

Windows: download LoRA-Dataset-Studio-windows.zip from the latest release, extract it into a new folder and run start.bat. The launcher prepares Python and opens LDS in your browser.

Complete Setup, then create a dataset or install the plugins you need. Importing, organizing and manually captioning images require no GPU or API key.

InstallationInstructions
Git checkout or manual Python environmentNative installation
Docker with an existing or fresh ComfyUIDocker guide
Docker without a GPUAPI-only setup
PinokioOne-click installation
Rented RunPod GPURunPod guide

Updates: ZIP installations use Update & restart and retain their datasets, media, settings and history. Git installations follow their configured branch; v2 is maintained, while v1 is frozen. Old main installations must migrate to V2. Pinokio and Docker use their own update procedures.

Minimum requirements

The core runs without a GPU. Python 3.10–3.12 supports the local ML extras; the Windows launcher downloads Python 3.12 if needed. Local generation typically needs about 16 GB NVIDIA VRAM; training requirements depend on the family and settings. See the hardware and dependency tables before downloading models or renting a GPU.

Documentation

The documentation index includes development and plugin references.

Configuration & network access

Use Settings for normal configuration. Native installs bind to 127.0.0.1 by default. Read the security policy before enabling network access; a protected connection also lets you use LDS from a phone or tablet.

Optional usage statistics are off by default. Update checks, requested model downloads, configured providers and scraping can contact external services. The network and privacy guide describes these connections, the data shared and public-access settings.

Support the project

GitHub Sponsors supports development, API testing and rented test GPUs. Bug reports, contributions and sharing the project also help. For support, generate a diagnostic report under Guide → Getting help, then use Discord or GitHub issues.

Affiliate disclosure: the project's vast.ai links pay the project 3% of referred users' spending for the lifetime of their account, at no extra cost. Rentals use your own API key and are billed directly by vast.ai. This is separate from optional usage statistics.

Use material you have the rights and consent to train on. Non-consensual likeness use, impersonation, fraud, and sexual or exploitative content involving minors are prohibited. You remain responsible for datasets and outputs. See responsible use for consent, privacy, copyright, platform terms and warranty details.

Contributing

See CONTRIBUTING.md for development and pull requests, and the Code of Conduct for community rules. Report vulnerabilities privately as described in SECURITY.md.

License

PolyForm Noncommercial 1.0.0. Noncommercial use is permitted; commercial use requires separate permission from the licensor.

ai-toolkit
comfyui
dataset
diffusion-models
fine-tuning
flask
full-model-training
lora
react
stable-diffusion

perfectgf/lora-dataset-studio

Self-hosted, one-tab workbench for the whole training lifecycle: build Character/Concept/Style datasets (generate, scrape, or triage a big dump in the Image Bank), curate, caption, clean watermarks, then train LoRAs on five model families or a full Krea 2 model on a rented GPU, and rank checkpoints in a Test Studio. Free & local.

Python

295

28 commits

updated Sep 29, 2026

See the code

See what people are saying

SourceMessageScoreDate

ai slop game based on batlle video in real time with only open model (r/StableDiffusion)

Check out [Chimera Arena (https://chimeraarena.com/static/showcase/index.html?lang=en)](https://chimeraarena.com/static/showcase/index.html?lang=en)! Here’s my slightly scrappy AI-generated teaser for Chimera Arena, my card game combining AI-generated video battles, real-time 2D combat, and…

16

Oct 3, 2026

README

Support us on Patreon Join our Discord

[!WARNING] No affiliation with loradataset.com. That website is not operated, endorsed, or supported by the LoRA Dataset Studio team. Any payments made to that service do not support this project. Use this repository and the links provided here to find our official downloads and community channels.

LoRA Dataset Studio V2

Build, curate, caption and train image datasets from one browser interface. LDS runs on your machine and uses ai-toolkit for training and ComfyUI for local generation and checkpoint testing.

Install · Documentation · Plugins · Releases · Discord

The core and public plugins are free under the PolyForm Noncommercial license. The core needs no account; optional usage statistics are off by default. External APIs and rented GPUs have their own charges. Additional paid plugins may be offered later.

Generation engines
Generation engines
Dataset workspace
Dataset workspace

A complete Character LoRA workflow in seven minutes:

https://github.com/user-attachments/assets/d51ff89c-34e9-41a9-b47d-08939a8c867b

People shown in the demo and screenshots are AI-generated.

What it does

AreaCapabilities
DatasetsCharacter, Concept and Style workflows; image and ZIP/folder import; subject-aware shot catalogs; local Klein and Krea 2 Edit generation; reference editing and exact retries
Image BankReview folders in place, score quality, group duplicates and people, search by text or similarity, and build balanced shortlists to promote into datasets
CurationKeep/reject, crop, mirror, rotate, face similarity, composition checks, editable watermark masks, text detection and recoverable cleaning
CaptioningLocal vision models or JoyCaption, family-appropriate prose or tags, appearance policies, bulk editing, targeted re-captioning and external image/.txt round trips
Local trainingGuided ai-toolkit recipes for Z-Image, SDXL, Krea 2, FLUX.1, FLUX.2 Klein, Anima and Qwen-Image 2.1; queues, advanced settings, checkpoint continuation and experimental slider LoRAs
Review and testingGenerate with all seven image training families; fixed-seed checkpoint/strength comparisons, prompt batches, runs, logs, lineage, votes, rankings and a generated-image Gallery
FilesStandard training ZIPs and sidecars, backup/restore without API keys, ComfyUI deployment, configurable storage and Trash

Dependencies vary by feature. Bank search ranks matches; it does not guarantee exclusions. Undo covers specific actions, and Delete rejected can remove source files after confirmation. Video and slider workflows are experimental. See the feature reference, requirements and known limitations for the detailed behavior.

Plugins

Install optional features from Plugins → Store, configure and prepare them in their own settings, and update them through Plugins → Updates. Several plugins can be installed together with one LDS restart. Each is independently installable; model downloads, hardware and provider credentials depend on the feature.

PluginWhat it doesMain requirements
API image enginesGenerate images with Gemini/Nano Banana, ChatGPT or OpenRouter; experimental ChatGPT subscription connectionProvider credentials; API usage is billed by the provider
Camera anglesCreate new viewpoints of Gallery or dataset images with controls for direction, height and distanceComfyUI, Qwen-Image-Edit and Multiple-Angles LoRA
CanvasArrange runs and checkpoints on a visual board, compare or blend LoRAs, generate images and export layoutsComfyUI for generation; one model family per generation batch
Cloud trainingTrain image or video models on rented GPUs, monitor runs, recover checkpoints and deliver full Krea 2 modelsvast.ai account and API key; Hugging Face access for applicable runs
DLSS 5 Neural RenderingImprove a finished video's lighting and materials, compare with the original and export the resultWindows x86-64, NVIDIA driver, compatible model supplied by you and the prepared bridge
Klein ImproveRe-render image detail with instructions, LoRA presets and finishing controlsComfyUI and Klein models; appearance and color can change
Live channelsContinuously render scripted H3 scenes and watch them in a browser or VLCCompatible local ComfyUI, H3 weights and stream encoder; experimental
Model toolsQuantize full models to fp8 or merge weighted LoRAs into a base checkpointPython with PyTorch and output disk space; CPU processing, no GPU required
Publish to CivitaiUpload checkpoints and generated images to a model pageCivitai API key; checkpoints default to drafts, image posts default to publication
Publish to Hugging FaceExport kept images, captions, metadata and a dataset card to the HubWrite-enabled Hugging Face token; private repository by default and rights confirmation
Resource monitorShow CPU, GPU, RAM, VRAM and temperatures, with guarded memory releaseNo model download
SeedVR2Restore and upscale images with high-resolution tiling and finishing controlsComfyUI, SeedVR2 nodes and models; optional TTP nodes for tiling
Video laneBuild video datasets from local files or web imports, curate shots, train LoRAs and test H3 generation, continuation and interpolationDependencies vary by task: PyAV/ffmpeg, ai-toolkit, ComfyUI and model weights; Video Bank remains beta
Web scrapingImport selected images or clips from searches and supported gallery URLs into datasets or banksSource-dependent credentials and permissions; Pexels requires explicit dataset/ML authorization

API providers apply their own billing and content policies. Model licenses also apply, including MiniMax H3's territory restrictions; check the video limits before using it. Plugin authors can start with the SDK and package guide.

Screenshots

Click a screenshot to view it at full size.

Image Bank
Image Bank
Analysis and coverage
Analysis and coverage
Batch analysis
Batch analysis
Dataset curation
Dataset curation
Text detection
Text detection
Training runs
Training runs
Checkpoint comparisons
Checkpoint comparisons
LoRA Canvas
LoRA Canvas
Civitai prompts
Civitai prompts
Camera angles
Camera angles
Video Test Studio
Video Test Studio
DLSS 5 comparison
DLSS 5 comparison

Setup & install

Windows: download LoRA-Dataset-Studio-windows.zip from the latest release, extract it into a new folder and run start.bat. The launcher prepares Python and opens LDS in your browser.

Complete Setup, then create a dataset or install the plugins you need. Importing, organizing and manually captioning images require no GPU or API key.

InstallationInstructions
Git checkout or manual Python environmentNative installation
Docker with an existing or fresh ComfyUIDocker guide
Docker without a GPUAPI-only setup
PinokioOne-click installation
Rented RunPod GPURunPod guide

Updates: ZIP installations use Update & restart and retain their datasets, media, settings and history. Git installations follow their configured branch; v2 is maintained, while v1 is frozen. Old main installations must migrate to V2. Pinokio and Docker use their own update procedures.

Minimum requirements

The core runs without a GPU. Python 3.10–3.12 supports the local ML extras; the Windows launcher downloads Python 3.12 if needed. Local generation typically needs about 16 GB NVIDIA VRAM; training requirements depend on the family and settings. See the hardware and dependency tables before downloading models or renting a GPU.

Documentation

The documentation index includes development and plugin references.

Configuration & network access

Use Settings for normal configuration. Native installs bind to 127.0.0.1 by default. Read the security policy before enabling network access; a protected connection also lets you use LDS from a phone or tablet.

Optional usage statistics are off by default. Update checks, requested model downloads, configured providers and scraping can contact external services. The network and privacy guide describes these connections, the data shared and public-access settings.

Support the project

GitHub Sponsors supports development, API testing and rented test GPUs. Bug reports, contributions and sharing the project also help. For support, generate a diagnostic report under Guide → Getting help, then use Discord or GitHub issues.

Affiliate disclosure: the project's vast.ai links pay the project 3% of referred users' spending for the lifetime of their account, at no extra cost. Rentals use your own API key and are billed directly by vast.ai. This is separate from optional usage statistics.

Use material you have the rights and consent to train on. Non-consensual likeness use, impersonation, fraud, and sexual or exploitative content involving minors are prohibited. You remain responsible for datasets and outputs. See responsible use for consent, privacy, copyright, platform terms and warranty details.

Contributing

See CONTRIBUTING.md for development and pull requests, and the Code of Conduct for community rules. Report vulnerabilities privately as described in SECURITY.md.

License

PolyForm Noncommercial 1.0.0. Noncommercial use is permitted; commercial use requires separate permission from the licensor.

ai-toolkit
comfyui
dataset
diffusion-models
fine-tuning
flask
full-model-training
lora
react
stable-diffusion

Languages

Python

64.4%

JavaScript

35.0%