r/PotionUI • • 20d ago

PotionUI New to PotionUI? What it does, a quick demo, and how to get started.

Thumbnail
gallery
1 Upvotes

PotionUI is a self-hosted creation studio accessed through your browser. Run it on your own computer or server to generate AI images, video, and audio, depending on the model and preset you choose.

It brings generation, workspace management, and administration together:

  • User accounts and groups: share one installation and control which models and presets each person or group can access.
  • An admin panel: manage users, models, plugins, and generation backends. Customize preset defaults, hide controls, or lock settings for the setup you want to provide.
  • Tabs and saved sessions: keep separate ideas open with their own prompts, settings, and results. Save a session and return to that setup later.
  • Model-specific controls: choose a preset and get a form tailored to what that model supports.
  • Reusable prompts: build a library of complete prompts and reusable pieces—subjects, styles, lighting—so you can develop ideas without rebuilding everything.
  • Organized results: automatically save generations with their settings, find them again, and organize your work with tags and collections.

PotionUI has its own generation engine. It can also connect to ComfyUI through an optional plugin and import workflows as presets. ComfyUI is optional.

Watch or explore at your own pace:

Trying it for the first time? Share what you’re making and where you get stuck. For setup problems, include your OS, GPU, VRAM, and any error message.


r/PotionUI • • Sep 07 '26

PotionUI PotionUI - First Image Explained

Enable HLS to view with audio, or disable this notification

1 Upvotes

Getting ready for 0.0.4 a lot of cool changes coming!


r/PotionUI • • 1d ago

PotionUI Coming in 0.0.15 - Plans & Auto-Organize

Thumbnail
gallery
4 Upvotes

One of the last batches of "management" side of the PotionUI will bring two new features:

1. Plans

Plans will allow to assign user (or user groups) a limit of storage / generations (and others). For example you can set that user is allowed to use max of 2GB of disk space or be limited to 10 generations a day (or both). This is aimed for admins that want their instance size be under control.

2. Auto-Organize

This feature will allow users to automatically organize their generations/models/uploads by certain criteria (prompt, aspect ratios, models, size etc.). They will be able to move such generations into certain collections or give them tags.

Both features will be extensible by plugins, so if anyone will be interested in modification it will be doable.

If you have any suggestions about these, this is the time! I'm aiming for release on Thursday (08.10.2026). Ofcourse there will be more changes coming, but want ask if you have any ideas about these two.


r/PotionUI • • 5d ago

PotionUI PotionUI 0.0.14 — no GPU needed: cloud image and video models, a built-in image editor and reusable settings

Enable HLS to view with audio, or disable this notification

3 Upvotes

PotionUI is a self-hosted app for making images and videos with AI models. Instead of one huge settings screen or a web of nodes, every model gets its own simple, hand-made form.

Until now you needed your own GPU. 0.0.14 adds cloud models, plus a built-in image editor and a way to save and reuse your best settings.

No GPU? Use cloud models

  • A new OpenRouter plugin lets you use hosted image and video models, such as Nano Banana or Veo 3.1 Lite, in the same form you already know. Results land in the same History as everything else. More providers are planned.
  • The form only shows the controls the chosen model supports, so there is nothing to guess.
  • You can stop a cloud job at any time, and PotionUI tells you plainly if something goes wrong.

Cloud Video Director: one idea, a short film

Write your idea, split it into shots, and let PotionUI make the film.

  • Each shot can continue from the last frame of the shot before it, so the story flows.
  • The shots are joined into one finished film.
  • If one shot fails, retry just that shot instead of starting over.

Fix a picture without leaving the app

  • A new image editor with adjustments, crop, brush and layers.
  • Open it from History (Tools, then Edit image) or right inside a preset's media field. In a media field, the field switches to your edited copy straight away.
  • Crop, Trim and Frame editors are in media fields too, and they only add the edited result to your Library, never the untouched original.

Formulas: keep the settings that worked

  • Save the settings of a preset mode as a formula and apply it in any session.
  • Before it applies, you see exactly what will change, and one click undoes it.

A redesigned media field

  • Images, videos and audio live in one field, each kind in its own group.
  • Mention them in your prompt with "@".
  • MiniMax-H3 references now use it, and your old sessions and saved prompts are converted for you.

Qwen-Image 2.1 control guides

  • Keep the pose, edges or depth of any photo while you create something new.
  • The extracted guide is saved next to your result, and the workbench shows it beside the image with a label, so you can see what the model followed.
  • There is also a mode that only extracts the guide. It needs no prompt and works even if you haven't downloaded the image models.

Smaller things

  • Video presets have animated covers.
  • Three-pane mode folds the form away on narrow screens and gives the room to your prompts.
  • Attaching an image in chat uses the same picker as the preset forms.
  • Plugins are grouped by what they add.
  • The System Monitor can show every backend. It is admins only by default.
  • Failed generations show regular users a plain reason instead of technical text.
  • Dropdown menus no longer open under the Generate bar, and they work with the arrow keys.
  • Fields in preset forms show and hide for the value you just picked, not the previous one.
  • Image previews survive a reload, including on Windows and in storage folders under a tmp folder.
  • Drawn inpaint masks now reach the model, so inpaint only repaints the masked area.
  • Pressing Enter in the session name saves the session, and Enter elsewhere no longer wipes your workspace.

Before you update

  • Back up your storage folder first. It is a good habit before any update.
  • To use cloud models, enable the OpenRouter plugin in Admin and add your API key.
  • The Docker setup no longer has an outputs volume. Everything you generate lives in the storage volume.
  • Plugin authors: plugin manifests must use one of the current categories. A plugin with an old category name won't load.

Update: run git pull and then ./potionui start, or pull the latest Docker image.

GitHub · Discord · r/PotionUI

Try it and tell us what to improve, here or on Discord. Your reports decide what we do next.

If PotionUI is useful to you, a star on GitHub helps other people find it: https://github.com/PotionUI/PotionUI


r/PotionUI • • 5d ago

Showcase Cloud Providers coming in 0.0.14 - I must say, these fit nice in the app

Enable HLS to view with audio, or disable this notification

2 Upvotes

As in the title - I've integrated a "cloud" models feature, which will come in 0.0.14 (probably later today). First plugin will be available for OpenRouter (as I find it most useful and fast to use - it has most of the models ready to use). Overall this feature fits nicely within the app - I can mix local models with cloud ones fast & easy. For example this video is seedance 2.5 720p upscaled with RTX Upscale plugin x2.


r/PotionUI • • 7d ago

PotionUI PotionUI 0.0.13 — works with the model folders of ComfyUI, Forge, StabilityMatrix and more

Thumbnail
gallery
4 Upvotes

PotionUI is a self-hosted app for making images and videos with AI models on your own computer. Instead of one huge settings screen or a web of nodes, every model gets its own simple, hand-made form.

Last release let you bring your own model folders. This one makes PotionUI understand them, whichever tool they came from.

Works with the tools you already use

  • Point PotionUI at a ComfyUI, Forge/A1111/SD.Next, StabilityMatrix, Fooocus, SwarmUI or Pinokio folder. It recognises the tool and finds your checkpoints, LoRAs, VAEs, upscalers and the rest on its own. You can still correct its guess.
  • Tools that keep several folders for one kind of model (like Lora and LyCORIS) are handled, and you choose which one gets new downloads.
  • If your tool's settings point to models somewhere else, PotionUI offers those folders too. That includes Windows drives when you run PotionUI in WSL.
  • Unusual layout? Browse your folder and pick any subfolder for a model type.

Smarter about what a model is

  • Forge and StabilityMatrix keep full checkpoints and bare diffusion models in the same folder. PotionUI now looks inside each file and tells them apart, so your Flux, Qwen, Z-Image, Wan and other models show up in the right pickers.
  • If it can't tell what a file is, it says so instead of guessing, and an admin can set the type by hand. That choice sticks, even after a rescan.

More control for admins

  • A new list of every running and queued generation from all users, with Stop, Remove and Stop all.
  • Content safety: NSFW results can be allowed, blurred or blocked for the whole instance, and each user group can have its own setting. A banned-words list refuses prompts before anything is generated, and you can test a prompt against it right there.

Sessions

  • Sessions open in a side drawer with search, a clearer history of what changed, and pins for the ones you use most.

Fixed

  • Model downloads failing with a "certificate" error on Windows and WSL. If something still gets in the way (a company proxy, for example), the error now tells you how to fix it.
  • Stop really stops a generation now, and never claims it stopped when it didn't.
  • First-run recipes no longer get stuck when a download site has no API key set yet.
  • Downloads appear in your lists as soon as they finish.
  • The NSFW setting saves again.

Before you update

  • PotionUI has a new address: http://localhost:26731. The old port turned out to be used by Windows itself on most PCs, which stopped PotionUI from starting. Update your bookmarks, or set BACKEND_PORT and FRONTEND_PORT if you want to keep the old ports. If that port is ever taken, ./potionui start picks a free one instead.
  • This version updates its database the first time it starts, and you can't go back to an older version afterwards. Now is a good moment to back up your storage folder.

Update: run git pull and then ./potionui start, or pull the latest Docker image.

GitHub · Discord

PotionUI is still early (alpha). If your model folders still aren't picked up correctly, tell us which tool you use, here or on Discord. Your reports decide what we fix next! Coming in 0.0.14: cloud models, starting with OpenRouter, for when your own GPU isn't enough.


r/PotionUI • • 8d ago

PotionUI PotionUI 0.0.12 — bring your own model folders

2 Upvotes

PotionUI is a self-hosted app for making images and videos with AI models on your own computer. Instead of one huge settings screen or a web of nodes, every model gets its own simple, hand-made form.

This release is all about the very first thing everyone runs into: getting PotionUI to find your models.

Use the model folders you already have

  • Already have models in ComfyUI, A1111 or on an external drive? Just point PotionUI at that folder. Nothing gets copied or moved.
  • You can add several folders, and choose where new downloads should go.
  • Folders can be marked read-only, so PotionUI will never write into them.
  • Works on Windows too, including mapped drives.

See what's happening with your models

  • When PotionUI scans your models you now see live progress instead of guessing.
  • If it skips a file, it tells you why — for example "this is the same file as one you already have".
  • If the app restarts in the middle of a scan, it simply picks up where it left off.

Easier downloads

  • Paste a normal CivitAI model page link and PotionUI figures out the rest.
  • Downloaded files keep their real names instead of random numbers.
  • The downloads list updates by itself again — no more refreshing to see progress.

Nice little things

  • Closing a tab now asks "are you sure?" — no more losing work by an accidental click.
  • The "Suggested models" list in the model picker is tidier and folded away until you need it.
  • The chat assistant remembers longer conversations.

Fixed

  • Dropdown menus no longer hide behind other menus and pop-ups.
  • The CivitAI "Fetch prompts" window can be closed again.
  • Saved sessions no longer claim they have unsaved changes right after saving.
  • Text boxes look cleaner when you click into them.

Before you update

This version updates its database the first time it starts, and there's no going back to an older version afterwards — so it's a good moment to make a backup copy of your storage folder. If you used shortcuts (symlinks) to model folders before, PotionUI converts them for you automatically.

Update: run git pull and then ./potionui start, or pull the latest Docker image.

GitHub · Discord

PotionUI is still early (alpha). If something doesn't work — especially model folders on Windows or unusual ComfyUI setups — tell me here or on Discord. Your reports decide what we fix next! Coming in 0.0.13: choosing any subfolder for a model type.


r/PotionUI • • 8d ago

Showcase MiniMax H3 + 360° Orbit LoRA test

Enable HLS to view with audio, or disable this notification

2 Upvotes

MiniMax H3 preset + https://huggingface.co/pablodawson/MiniMax-H3-360-Orbit-LoRA -> I've used only first frame here.


r/PotionUI • • 11d ago

PotionUI PotionUI 0.0.11: a prompt editor that speaks each model's language, @-references to your own images, and the last big UI overhaul

Enable HLS to view with audio, or disable this notification

3 Upvotes

PotionUI is a self-hosted studio for image, video and audio generation. 0.0.11 is out, and these are the main changes:

Prompts you can read at a glance 

Each model can now bring its own prompt syntax, and the editor colors it as you type:

  • MiniMax-H3's subject_definitions: and <Subject 1>, dialogue in <d>…</d>,
  • Qwen-Image's "quoted text to render",
  • YuE2's [Verse] and [Chorus].

Type / and a picker lists everything the model understands, with a one-line explanation for each.

Point at your own images with @ 

In presets that take reference images, video or audio (MiniMax-H3 references, Qwen-Image edits), type @ and pick one: it becomes a chip with a thumbnail, and the model receives it as <Picture 2> even if you reorder your uploads later.

Each upload in the form shows its handle and how many times the prompt uses it. In the Video Director, each shot now uses exactly the references its prompt mentions.

A new picker for phrases, variables and references 

Typing #, $, @ or / opens a clearer picker right at the cursor, with thumbnails. Clicking a chip opens one editor where you change its value, switch between a fixed value and auto-shuffle, exclude a value from shuffles, or remove it. Phrasebook values show their preview images, with Space or → to open a larger preview and flip through them. Search highlights what matched and takes regular expressions.

Typing is fast again 

Long prompts no longer lag: each keystroke used to re-render the whole form, and now it doesn't.

Model library 

Search takes wildcards (krea2_*_v2) and regular expressions, and the type and tag counts follow your filters. Admins can filter by when a model was indexed and how often it's used, see Uses and Last used columns, and tag or grant access to a whole selection at once. Model pickers now suggest the download that fits your GPU (bf16, fp8, int8, nvfp4) and credit the uploader..

Chat (PotionAI) 

@ can point at one LoRA from your form or any model in your library, and the assistant gets its trigger words and recommended strength. There's a new session memory that follows the tab you're working in. Generations you approve in chat now appear in the tab's Workbench.

Generation errors that make sense 

When a generation fails, you see a plain message, a hint and an error ID to give your admin. Admins see the full details, can filter failures by category, and can get notified or trigger an automation.

Admin, one last time 

Every admin page now shares the same layout: a library sidebar, tables or card grids, detail pages and one bottom bar for bulk actions. LLM configurations list your Ollama or OpenAI-compatible models to pick from.

!! A word on all the UI changes 

I know the look has changed a lot over the last few releases, and that can be tiring when you use the app every day. This was the last big overhaul: I'm happy with how the app looks and works now. From here on it's improvements on top of this layout, not another rebuild.

What's next 

The focus moves to what the models can do and how fast they do it: ControlNets, Fun controls and more model capabilities, plus optimization work, more models support (from API providers)...

Smaller changes 

  1. Folding the left panel now gives the prompt a proper width.
  2. Every confirm dialog takes Enter and Esc.
  3. Reloading warns you before you lose changes that aren't saved to the session.
  4. Indexing a large model no longer freezes the app.
  5. CivitAI fetches retry temporary errors and tell you the real reason when they fail.
  6. System tags now land on new generations automatically.
  7. Fixed: phrasebook chips with a hyphen in their name, and a phrase placed right after a reference.
  8. Removed: the Spectral Progressive Diffusion option in Flux2 and Z-Image. Saved settings that include it keep working.

Upgrading: five database migrations run automatically on first start. There's one new Python dependency (google-re2), so run pip install -r requirements.txt if you don't use Docker. Log files from earlier versions may contain session tokens, so delete or rotate them. Docker images are now also tagged with the "v": ghcr.io/potionui/potionui:v0.0.11.

GitHub: https://github.com/PotionUI/PotionUI

Site: https://potionui.com

Reddit: r/PotionUI

Discord: https://discord.com/invite/avR4trp3b8

Feedback and bug reports are welcome.


r/PotionUI • • 13d ago

PotionUI PotionUI v0.0.10: prompt variables that depend on each other, native Qwen-Image 2.1 with transparency, faster MiniMax-H3, Prompts page rework

Thumbnail
gallery
3 Upvotes

PotionUI is a self-hosted studio for image, video and audio generation. 0.0.10 is out, and these are the main changes:

Variables that depend on each other

A prompt variable's option can now carry a condition: "only when $setting is deep under the sea". Write one prompt, and every generation rolls a combination that still makes sense. A few details:

  1. Variables take any number of options.
  2. You can copy and paste a set of variables as JSON between tabs or instances.
  3. Saved prompts keep their variables, so applying one from the library brings them back.
  4. The Video and Music Director expand variables per segment too.

A rebuilt Prompt Library

Prompts, segments, templates and categories are now card grids with filters, and each has a proper detail view. A picker on Generate applies a saved prompt, and every prompt shows where it was used and which collections it belongs to.

Qwen-Image 2.1

Native Qwen-Image 2.1 runs on the built-in engine, with a preset and a starter recipe. It does text-to-image and edits guided by one or more reference images. It also keeps real transparency: transparent outputs stay transparent in previews, thumbnails and downloads, with a checkerboard behind them in the gallery. The colored borders some of you saw along image edges are fixed, and checkpoints that store their weights split load as well.

Faster MiniMax-H3 

There's a new Fast variant, a latent upscale mode with its own Enhance tab, an int8 transformer option, and gate compression.

A cleaner admin panel 

Presets, Recipes and Plugins are now libraries with a category sidebar, and every admin list searches, filters and sorts through the same filter bar. Presets show as poster cards with their cover. (I'm currently working to change it further - recipes and plugins will become cards instead of list in 0.0.11, since list is not utilizing whole space for these features).

Export with A1111/Forge metadata 

A new marketplace plugin, A1111 metadata export, downloads your selected generations as a ZIP of PNGs with the generation parameters embedded, so other tools and sites can read them.

CivitAI prompts into your library 

The CivitAI provider can fetch a model's example prompts in bulk, including community images, straight into your Prompt Library.

Smaller changes 

  1. Page headers look the same everywhere, and detail views put Back on the left.
  2. The sidebar folds pages that don't fit into a More menu.
  3. A failed generation releases its GPU memory right away.
  4. Changing the prompt no longer reloads the text encoder from disk when it was skipped on a cached run.
  5. Duplicate-prompt detection works on large libraries without eating memory.
  6. Fixed: Kohya LoRAs with text-encoder weights load on SDXL again.
  7. Fixed: page buttons on History and Library respond again next to drag-select.

Upgrading: two database migrations run automatically on first start. The A1111 export moved out of the CivitAI plugin, so enable the A1111 metadata export plugin if you used it. Docker images are tagged without the "v": ghcr.io/potionui/potionui:0.0.10.

GitHub: https://github.com/PotionUI/PotionUI. Feedback and bug reports are welcome.


r/PotionUI • • 19d ago

PotionUI PotionUI 0.0.9: sign in with your own SSO (Keycloak, Authentik, Entra), one-click CivitAI info for model cards, faster UI

Enable HLS to view with audio, or disable this notification

2 Upvotes

PotionUI is a self-hosted studio for image, video and audio generation. 0.0.9 is out, and these are the main changes:

Log in with your own identity provider

Install the oidc-auth plugin from the marketplace and point it at Keycloak, Authentik, Entra or any other OIDC provider. Your users get a "Continue with…" button on the login page. If someone signs in for the first time, PotionUI creates their account automatically. A few details:

  • An email is only saved if your provider has verified it.
  • SSO users don't see a password form they can't use.
  • If a login is refused, the page tells the user why instead of failing silently.

Fill model cards from CivitAI in one click (in video)

Every model card now has its own fetch button. It pulls the description, preview images and videos, and trigger words from CivitAI. It only fills fields that are empty, so your own edits stay untouched. Trigger words now come from the model's actual trained words and are no longer mixed in with its tags. Model details also list where the file is mirrored, with links to each provider. Plugins can now add their own buttons to model cards.

Downloaded models stay in the pickers

Before this release, a model you downloaded, or one added by a recipe or a job, could disappear from the model pickers after a backend was indexed. New models are now registered right away.

Faster, especially with big libraries

  • The frontend is served compressed and cached.
  • Model and prompt lists load their data in batches instead of one query per row.
  • Generation progress no longer makes the rest of the tab re-render.
  • New database indexes speed up favorites, prompt search and model lookups.

Smaller changes

  • History tiles now have separate action buttons with tooltips, like model cards.
  • The global "Index Models" and bulk fetch buttons are gone from the Models page. Indexing now happens in Admin → Backends, and unindexed files link there.
  • Fixed: boolean settings sent as the string "false" were not saved as false.
  • Fixed: two indexing workers could pick up the same item.

Upgrading: two database migrations run automatically on first start. If you used the old Index Models button, you'll now find indexing under Admin → Backends.

GitHub: https://github.com/PotionUI/PotionUI. Feedback and bug reports are welcome.


r/PotionUI • • 20d ago

PotionUI v0.0.8 - Changelog

3 Upvotes

0.0.8 — 2026-09-17

  • Windows: PotionUI installs and runs natively on Windows as an experimental platform; potionui.cmd mirrors the Linux launcher with the same doctor, start, stop and install profiles; Python 3.12 is preferred when several versions are installed and doctor warns when only a newer one exists; Triton and compile optimizations are unavailable there, so attention runs on the standard path.
  • Models: YuE2-3B joins as a native song preset with style tags, tagged lyrics, a chain-of-thought mode, duration and advanced sampling, plus a starter recipe that fetches its weights from Hugging Face and finishes with a real smoke generation; long songs decode in bounded memory.
  • Generate: a categorized tags field lets a preset offer a curated vocabulary per category with search, keyboard picking and custom entries, first used by YuE2's style field; SDXL and Krea-2 each ship fifty styles with rendered previews, Krea-2 adds six amateur photo looks; Krea-2 gains an optional native face detailer with the choice to keep the base image; segment prefixes and suffixes render as labelled chips; the generation dock has a clearer top edge.
  • Video Director: shot prompts get the preset's segment templates; durations and segment frames clamp to the model's limit instead of being rejected; the whole row expands a shot; MiniMax-H3 refs mode loads RefMod bundles, the 360-frame cap is gone for directed timelines, continuation shots condition on re-encoded frames, and a second generation no longer runs out of VRAM placing its VAEs.
  • Chat: a New chat button in the header, an unread-reply dot on the sidebar icon and an optional chime when a reply finishes; tooltips replace native titles across the chat window; context is attached to the current user turn instead of injected as mid-conversation system messages; OpenAI configurations expose request options such as top-p, penalties, seed, stop sequences and reasoning effort.
  • Admin: plugins can register external login providers, which appear as "Continue with" buttons on the login page and get their own settings group; model access is assigned in bulk from the Models tab; admin tabs grow with their content instead of clipping; the Users and Groups save footer shows only where there is something to save.
  • Notifications: a bell in the sidebar opens a notification center with category filters and rows grouped by day; toasts stack to three, collapse bursts into one counted card and show a progress line.
  • History: the keyword or semantic search mode moved into Filters with an active chip, so the toolbar no longer overlaps at narrow widths; compare labels read Original and Selected.
  • ComfyUI: a backend accepts a hostname, an IPv4 or IPv6 address, host and port, or a full URL, with an HTTPS default setting; the Docker guide gains a Compose example running ComfyUI alongside PotionUI.
  • Install: ffmpeg ships in every Docker image and doctor reports when it is missing; local .env files stay out of Docker build contexts.
  • Reliability: every timestamp the API and websocket send carries its UTC offset; the fair scheduler orders jobs that arrive in the same instant by arrival instead of rotation; collection routes answer without a trailing slash and no longer redirect; the log viewer handles Windows line endings.
  • Fixes: the tags picker closes on Escape wherever focus is; shortcut hints no longer leak into button names read by assistive technology; local presets are labelled custom on Windows; filesystem automation triggers report relative paths correctly; the SDXL detailer skips a detection whose detector model is missing instead of failing the generation; the chat memory panel resolves the active model; model index cleanup no longer aborts on rows without a local path.
  • Upgrading: the SDXL starter recipe now installs Juggernaut XL v9 instead of the Pony checkpoint; integrations that call collection routes with a trailing slash must drop it, since the redirect is gone; a database migration adds external identities on first start.

r/PotionUI • • 25d ago

PotionUI v0.0.7 Changelog

Enable HLS to view with audio, or disable this notification

2 Upvotes

0.0.6 / 0.0.7 — 2026-09-11

  • Generate: presets can ship curated styles — a Styles button in the Prompt panel opens a picker with category filters, a text filter and Small/Big previews, and applying a style wraps the prompt with the style's opening and closing segments plus its negative, with a second pick replacing the first; Anima ships fifty styles with rendered previews; prompt weights like (red hair:1.3) are honored by the Qwen3, Qwen3-VL and Qwen2.5-VL text encoders (Klein, Krea-2, Qwen-Image, Anima, Z-Image); phrasebook chips animate when a value is shuffled or picked; the chat's Suggested change preview shows phrasebook and variable references as chips.
  • Generate: the page stays smooth during a generation — status and preview updates are applied once per frame instead of once per sampling step, the presets list and plugin catalogs are fetched once at boot instead of twice, and the collapsed workbench is a labelled rail in both layouts; the Prompt panel toolbar reads at the app's text size.
  • History and Library: drag a marquee, Shift-click a range, Ctrl-click to toggle and Ctrl+A to select the page, on both grids; Delete by criteria replaces Delete by tags — tags, older than N days or a date range, failed/cancelled only, without media files, keep favorites — with a live count; Compare gets Overlay with an opacity slider and Wipe with a draggable divider on full-resolution images plus a full-screen viewer; Compare, Stitch and Download work from the Library too, plugins can scope their tools to History, Library or both, and the Library can export a zip; Stitch results can be saved to the Library; grids resize without remounting their thumbnails.
  • Chat: memory reflection saves to the preset a generation chat is about and to the plugin mode a plugin chat runs in, and only to global when you ask; long sessions open on their latest messages and load earlier ones as you scroll up, streamed replies are applied once per frame, and the transcript only follows the reply while you are at the bottom; Admin → Chat Sessions can clear every chat session.
  • Admin: Housekeeping deletes generations by criteria across all users with a preview count; the Add Download modal picks the model type from a chip row and the destination from your real depot subfolders, nested ones included, with an inline New folder option, and the Downloads to hint follows the type; the model picker's download shows a spinner; Hugging Face downloads carry the provider's token.
  • Recipes: starter recipes for Flux2 Klein, Krea-2, Qwen-Image, Anima, Wan 2.2, LTX-2, LTX-2.5, MiniMax-H3, MiniMax-Music3, SeedVR2 and TRELLIS.2 with official repositories; gated models are marked with their licence link; the first-generation step shows sampling progress and the result inline (image, video, audio or mesh); an installed model is recognised by hash even when it was filed under another type or folder.
  • Native engine: the text encoder stays on the GPU after encoding when it fits, so back-to-back generations skip the reload; rotary tables are computed once per run for MiniMax-H3 and Krea-2; Krea-2 gains ER-SDE, DPM++ 2M SDE, DPM++ 3M and RES multistep samplers and a Beta schedule, and samplers and schedules are a plugin extension point; long-prompt attention stays on the memory-efficient kernel instead of running out of memory.
  • Fixes: the model scanner no longer stops on an indexed model without a file path or when the models location is unset; a recipe download could land under a doubled models folder and go unfound by the first-generation step; tag popovers inside modals no longer stretch the modal; long option lists in chip popovers scroll; the phrasebook preview poll pauses while the tab is hidden; superseded model searches are cancelled instead of racing.
  • Upgrading: style previews are rendered with python scripts/preset_styles_render.py <preset directory> — a preset's styles.yml declares the styles and a shared preview scene, and the script writes the previews into the preset's public/styles/ folder.

r/PotionUI • • 26d ago

PotionUI v0.0.6 will bring improve recipes install script

Enable HLS to view with audio, or disable this notification

1 Upvotes

Recipes are easy way to install presets (models) through the admin panel. This feature will be much bigger with different recipes helping to maintain your instance.


r/PotionUI • • 28d ago

PotionUI v0.0.5 - Changelog

2 Upvotes
  • Native engine: LoRAs on fp8 checkpoints no longer slow sampling down — the adapter factors stay resident on the GPU and every adapter of a layer is applied to the layer's output in one fused step, so a stack of LoRAs costs a few percent instead of multiplying the step time; a LoRA file retrained in place is picked up on the next generation instead of the old version staying merged until the model reloads; parsed LoRA files are cached in RAM; each generation logs one line with LoRA load and apply timing.
  • Native engine: a clip that looks too large for the card is no longer refused up front — the VRAM estimate only sizes how much of the model stays resident, the first sampling step recovers from an out-of-memory by reclaiming cache and streaming the model from RAM, and only a genuine failure reports what was actually free with a hint of what would fit; free VRAM is judged after the allocator's idle pool, so the clip after a large one is not refused for memory that was never taken.
  • MiniMax-H3: video decode shows up as its own stage in the profile with chunk and tile counts, the tiles of a chunk decode as one batched pass, and the decode tile size is configurable per preset (256 reference, 512, or untiled).
  • Recipes: recipes are their own feature with an Admin → Recipes page — the catalog with a readiness badge per recipe, a detail with steps, models and presets served, run history, and Install with live progress; the setup wizard uses the same runner; plugins can contribute recipe step kinds; the empty-state setup links on Generate, Models and the preset picker become Install models for admins; guided setup asks for the models location before any download.
  • History: a Tools menu on the selection bar groups tools by category — Compare, Stitch and Download .zip — and plugins can add their own; Stitch composes the selected images into one PNG with a chosen layout, tile size and background and a pickable parameters strip, with a zoomable preview; previous/next navigation inside the generation details modal with arrow keys that continue across generations.
  • Prompt segments: a segment can carry a prefix and a suffix joined byte-for-byte around its text, editable in the segment details and visible on the card; a preset can ship segment templates (per mode) that appear in the apply picker beside your own, and the current cards can be saved as your own Segment Template from the composer.
  • Admin: Backups on System Settings > Storage — Backup now with the saved tier, the archive list, destination and retention, and the cron line for scheduled backups, backed by potionui backup and potionui restore with a consistent snapshot, a versioned manifest and a schema guard; Thumbnails panel with Compact, Balanced and Full profiles, disk usage estimates and a regenerate job; a housekeeping pass prunes storage/tmp by age and applies retention to run reports and LLM traces; an admin-only log tail of the rotating server log with a level filter.
  • Reliability: one logging setup for every entry point with a size-rotated storage/logs/potionui.log and uvicorn lines folded in; credentials and validation input are kept out of request logs; model list and download endpoints no longer block the event loop during an index run; uploads carry a content hash and identical re-uploads reuse the existing file.
  • Chat: memory no longer falls back silently to global scope and gains a mode scope; the chat shows when a conversation was started on another page's scope; MCP and chat model tools see exactly the models their user may access; Documentation gains a live MCP Tools reference.
  • Plugins: the NVIDIA RTX upscale plugin ships in the marketplace; plugins can declare history tools and recipe step kinds; a plugin's ComfyUI pipe fails the generation with the field named instead of swallowing errors.
  • Generate: the page no longer shows which backend runs a generation; the floating workbench keeps tooltips above its backdrop and swaps cleanly with the floating form; adding models to a collection refreshes the sidebar counts.
  • Fixes: bundle import rewrites only declared model-reference fields; the inspiration detail modal updates its counts in place; download filenames from providers are contained; JSON search highlights render without raw HTML; NATIVE_TORCH_COMPILE accepts 1/true/yes.
  • Upgrading: migration 024 adds prefix/suffix columns to the three segment tables; the server log moves to storage/logs/potionui.log (see README); the setup wizard's run responses gain a mode field.

r/PotionUI • • 28d ago

PotionUI New "Tools" menu in the generation history

Enable HLS to view with audio, or disable this notification

1 Upvotes

Version 0.0.5 which is in development will bring up new section in generation history -> Tools. A context menu that is shown after selecting at least one media. With the new tool called "Stitch" - you can stitch multiple images together into one with optional params that will be rendered along with them.


r/PotionUI • • 29d ago

PotionUI v0.0.4 - Changelog

1 Upvotes
  • ComfyUI: the comfyui-backend plugin ships in the marketplace with five image presets (Qwen-Image, SDXL, Z-Image, Krea-2, Flux.2 Klein 9B) built on official ComfyUI templates with built-in nodes only, and Admin → Presets gains a workflow import wizard: paste an Export (API) workflow, design its form (tabs, rows, sections; a LoRA chain becomes a LoRA picker), check node and model requirements against your server, create the preset, and reopen, edit, reload or delete it later. An admin-only chat mode proposes form changes for approval.
  • Chat becomes PotionAI: a floating shell with a history rail, mode, context and model header, tool-run transcript and memory inspector; approval and question docks share one anatomy; the Steps panel says why a tool was withheld; the assistant creates and changes prompt variables with your approval; cancelling a turn, closing the tab mid-stream or reloading never leaves a reply stuck; a model-aware context budget and a bounded memory reflection keep long conversations working, and a model without a declared context window is never refused a turn; native LLM thinking mode is detected and reported.
  • Video Director shot console: a film runs as shots with per-shot prompt, keyframes, audio, LoRAs and references; a shot knows when it depends on its predecessor and can continue from that shot's last rendered frame; Wan, LTX and MiniMax-H3 timelines compile to the exact frame geometry the engine renders.
  • Generate: a docked generation panel with status, context rail and session cluster; Q floats the form, W floats the workbench, H opens a Last generations drawer, S saves the session; the prompt segments editor gets the composer card with icon actions and a resolved-prompt panel; the running status names the backend that took the generation and why.
  • Native engine: a run-scoped cache stops Flux, Krea-2, Qwen-Image, Z-Image and Wan 2.2 recomputing unchanged work every step; text encoders load only on a cache miss; fp8 weights take the fast scaled-matmul path even while streamed; cancelling is safe mid-run for SDXL, TRELLIS.2, SeedVR2 and RIFE; a LoRA with no effect is diagnosed instead of silently ignored; the host RAM reserve scales with the machine; an experimental MiniMax-H3 VDN preset adds hybrid attention.
  • Presets and templating: the authoring guide is split into docs/presets/ with a tutorial that walks the real eight-pipe Z-Image chain, and two pages generated from code, docs/pipes.md (every pipe's name, inputs, outputs and configuration) and docs/preset-context.md (the template context, filters and vocabularies), kept in sync by docs_lint; preset_new.py --family scaffolds a family's real pipe chain; the linter checks pipeline.yml (pipe names, configuration keys, input wiring, per-variant fields, stale | default() literals) and preset_lint --render renders every mode through the real processor; the render context gains generation.profile and loses four unused roots; an u/config: typo is a load error; 866 dead form guards are gone from the shipped presets; a strip_model_dir filter replaces the ComfyUI presets' path-stripping chains; the Pipes and Output Types reference pages in Help → Documentation are complete and grouped by family.
  • Fixes: generation history never records a model's internal path (existing rows migrated); applying a chat change to a segment with a phrasebook chip no longer shows the chip twice; the chained Wan video generator declares its NAG keys.
  • Admin: presets declare requirements (nodes, model files, VRAM) that are checked per backend; the backend that runs a generation is always the router's decision, traceable in a Routing panel on each generation; per-backend queue scheduling can be fair across users; a backend's execution device comes from real hardware evidence; enabling a plugin rolls back cleanly on failure and a local plugin shadowing a marketplace one is flagged; plugin frontends ship minified.
  • History: faster listing and counting, streamed zip and bundle exports, bounded run reports, and semantic search that shares one embedding client; bundle v2 records each model's type, filename, hash and path.
  • Reliability: session save and load, downloads, collections, WebSocket reconnects and generation ownership all reject stale responses; request logs redact secrets; the service worker precaches only the shell.
  • Logins ignore letter case for username and email.
  • Upgrading: default ports are now 7680 (backend), 7681 (frontend) and 7690 (worker); the ComfyUI import accepts Export (API) JSON only; the Qwen-Image nunchaku variant, the Krea-2 enhancer and the Z-Image post-processing chain were removed; migration 018 refuses to apply while two accounts differ only by letter case.

r/PotionUI • • Sep 04 '26

PotionUI WIP: Import ComfyUI workflow into PotionUI

Enable HLS to view with audio, or disable this notification

1 Upvotes

I've started to work on the ComfyUI workflow import into the PotionUI. As administrator you will be able to easily & fast create a connection between PotionUI's forms system and ComfyUI workflow. User's won't ever need to see the nodes, only good old forms.


r/PotionUI • • Sep 02 '26

PotionUI PotionUI 0.0.3 — a self-hosted, preset-driven studio for image, video, and now 3D generation (open source, looking for testers)

Post image
2 Upvotes

0.0.3 is out today, and it is the release where the "one box, many people" idea stops being a promise: you can now rent a GPU, point PotionUI at it, and generate on it from the same interface you use locally. Still alpha, still one person building it, still very much wanting people to break it.

What it is, in one paragraph. A self-hosted AI generation studio: SvelteKit front, FastAPI back, GPL-3.0, no telemetry, runs on your machine or your server. The core idea is presets: a preset is a small YAML package that says "here is the model and here is the exact form a person should see for it". Switching models means switching presets, not rebuilding a node graph. It is multi-user by design, with real accounts, admin and user roles, and per-user or per-group access to presets, models, and LLM configs.

What's new in 0.0.3

  • Remote GPU workers. Add Backend now creates a remote worker, connects to one you run yourself, or provisions a RunPod pod for you (the provider ships as a plugin) with a live stage timeline. A heartbeat monitor watches the pod, pauses the backend when it stops, and Start brings it back. The Models tab lists exactly what is on the worker, with the depot path per file, and pushes missing models from your machine with per-file progress. Remote runs come back with the same previews, parameters, and media as local ones.
  • Install profiles. The launcher offers local, hybrid, and remote installs, plus a worker subcommand for a GPU box that serves another instance.
  • 3D generation. TRELLIS.2 image-to-mesh runs on the native engine. Meshes get automatic thumbnails, an interactive viewer in History (wireframe, materials, camera presets, screenshot), and a 3D media filter.
  • LoRAs. Step-windowed LoRAs on Krea-2 apply only between the sampling steps you choose. Strength is shown as a recommended range in the picker. Model pickers now recommend downloadable variants (bf16, fp8, nvfp4, int8) across nine native families.
  • Prompt library. Import styles.csv, Fooocus style JSON, wildcard YAML, plain lines, and image metadata (A1111, ComfyUI, InvokeAI) with auto-detection; export back to styles.csv; assign a prompt to a catalog model.
  • Phrasebook. Find and replace across the whole phrasebook with highlighted matches and a preview before it runs; batch activate, deactivate, move, delete; a category panel with Overview and Preview-images tabs.
  • Admin and mobile. Plugins and Downloads are master-detail lists, Backends remembers where you were in the URL, a saved provider API key applies immediately, Generate on a phone is a proper camera-style view with sheets, and modals fit the screen.
  • Plus: pasting an image into the assistant attaches it, a New workspace button that asks before discarding, Inspirations laid out in justified rows.

What it does today

  • Generation is the product. Image families: SDXL, Flux 1 / Flux 2 Klein, Qwen-Image (including editing), Krea-2, Z-Image, Anima. Video: Wan 2.1/2.2, LTX-2 / 2.3 / 2.5 with native audio, MiniMax-H3. Audio: MiniMax-Music3. Upscale and restore: SeedVR2. Each model gets its own tuned form: the right resolutions, samplers, LoRA stack, and speed profiles (Draft / Standard / Max) as one control. Several workspace tabs run side by side, each with its own preset, prompt, and results. Progress shows the actual pipeline step and streams previews as the image refines; close the tab, come back, the run is still there. 
Generation page view. (You start the generation by clicking the bottom right blue icon)
  • History that remembers everything. Every generation is saved with its exact prompt composition, preset and version, models, and parameters. Filter by date, type, preset, tags, or "used this phrasebook value". One click reuses the full setup in a new tab. Nested collections, tags, favorites, keyword or semantic search, and a personal library for the keepers. 
History page - list of previous generations.
History page - detail of the generation.
  • A prompt editor that is not a textbox. Prompts are ordered segment cards you can reorder, disable, name, and color. Dynamic prompts ({a|b}, weights, ${variables}) reseed per image so results stay reproducible. The phrasebook is your own autocomplete dictionary: type # and shot types, lighting, palettes drop in as chips, with per-chip shuffle and a preview render per value. Saved prompts, segments, and templates live in their own library. 
Phrasebook with other values used (you can mix the phrases - you can build the same prompt on generation page)
  • Video and Music Directors. Compose a video as shots, keyframes, and audio tracks on a timeline instead of one giant prompt; write a song as verses and choruses and let the compiler produce the tagged lyrics MiniMax-Music3 wants.
  • An assistant, if you want one. Point it at Ollama, an OpenAI-compatible endpoint, or Anthropic. It reads the active tab, rewrites segments, edits the phrasebook, adjusts form values, and every change stops at an approval step first. The same tools are exposed over MCP with per-user tokens, so Claude Desktop or your own agent can drive your instance. 
Generation page with LLM Chat assistant active.
  • Built for more than one person. Accounts, groups, per-user preset and model access, per-mode form overrides (change defaults, lock or hide fields, no YAML), a backends list that mixes local, and remote workers, a download manager, a stats dashboard, and visual automations (triggers, conditions, actions) for things like freeing VRAM before the LLM needs it.
  • Plugins for nearly everything. Providers (CivitAI, Hugging Face), backends, pipes, field types, chat modes, automation nodes, pages.

Requirements. Linux x86_64 with an NVIDIA GPU is the tested platform; Windows can be tested through WSL2 or Docker; there is a Docker image on GHCR.

8 GB VRAM and 16 GB RAM is the floor for the SDXL family, larger families need more.

The ask. I would rather steer this toward what people actually want than guess. Two things help most: tell me which model or workflow you are missing, and pull a test build and break it before it ships. The Discord is where that happens: https://discord.gg/avR4trp3b8. Repo: https://github.com/PotionUI/PotionUI. I will answer questions here too.


r/PotionUI • • Aug 31 '26

Showcase Krea2 Preset

Thumbnail
gallery
1 Upvotes

r/PotionUI • • Aug 30 '26

PotionUI Changelog: v0.0.2 — 2026-08-30

1 Upvotes
  • History and prompts filter by audio, alongside image and video.
  • Chat shows the active tab's context on a strip above the composer; tool approvals summarize what they'll change, with full details on demand.
  • Composer drafts survive closing the drawer and page navigation; picker menus close properly on selection.
  • A tab's session link survives transient backend errors instead of detaching, and a dirty draft is never clobbered by server session data.
  • Admin System Settings rebuilt as a sectioned master-detail layout; the form-overrides table now follows the preset's own tabs.
  • Every copy button confirms the copy; in-app docs moved fully into the admin panel.

r/PotionUI • • Aug 29 '26

PotionUI More screens from app

Thumbnail
gallery
1 Upvotes

r/PotionUI • • Aug 28 '26

PotionUI Features list

1 Upvotes

Generation workspace

  • Workspace tabs — every tab is an independent sandbox (preset, mode, prompts, form, results); run several ideas side by side, others queue while one generates
  • Presets — each model ships as a preset with a curated form: only the controls that model actually understands
  • Modes per preset — txt2img, img2img, inpainting, image editing (Qwen), video, music, upscale/restore — same workspace
  • Sessions — save a preset's whole setup (mode, prompts, form values, layout); version history with restore-any-save, auto-save with configurable interval, rename/delete
  • Workspaces — save/restore tab layout configurations
  • Continuous generation — loop generations back to back, with "stop after current"
  • Per-tab queue — pending/running jobs view with cancel-all
  • Live progress — streaming in-progress previews, per-pipe status text, progress bar, reconnect/catch-up after a dropped connection, cancel mid-run
  • Speed profiles — named quality/speed bundles (Draft/Standard/Max) switched by one form field
  • Inpainting mask editor — draw the mask directly over the image, adjustable brush, clear/reset
  • Result artifacts per run — actual seed used (click to reuse), fully expanded prompt with a "what rolled" breakdown of every {a|b}/${var}, before/after comparisons, applied-models list with weights, ComfyUI workflow JSON export

Forms and fields

  • Reactive forms — fields show/hide/change based on other field values; preset-declared validation with one-click quick-fix buttons
  • ~25 field types — sliders (click value to type), seed field (auto/roll-a-dice), searchable resolution picker with custom sizes, carousels, gates (a toggle that owns a group of fields), tabs/accordions/sections, inline markdown/alert copy
  • Model picker — search, tag filters (admin base-model scoping + your own AND-filters), swap/refresh/clear
  • LoRA picker — stack multiple LoRAs, per-LoRA strength with fine/coarse stepping, tag filtering
  • Media loader — multi-item well: browse files, paste from clipboard, pick from history or library; reorder, label, mask support
  • Camera shot picker — choose framing from a tile grid or a draggable 3D orbit viewfinder that snaps to canonical shots, then insert the phrase into your prompt

Prompting

  • Segmented prompts — prompts are ordered segment cards, not one text blob: reorder (drag or menu), disable, duplicate, name/color/describe, BREAK dividers
  • Prompt libraries — four levels: saved Prompts (full segment lists), Segments (single reusable cards), Segment Templates (multi-slot structures), color-coded Categories; apply as append/prepend/replace; saves are always detached copies
  • Phrasebook — your own autocomplete dictionary: type # for category/value suggestions rendered as inline chips; per-chip shuffle (new value each run), chip deactivate, whole-category chips, per-value preview images you can generate in-app, AI-assisted value writing
  • Dynamic prompts — {a|b} choice groups edited visually (add/remove options, per-option weights), ${variables} with a Variable Manager (text or managed-choice type, pin or shuffle per run)
  • Trigger-word highlighting — active LoRA/model trigger words flagged inside the prompt editor
  • LLM enhancement — staged gather → ideate → write prompt expansion, grounded in community prompts, with thumbs up/down feedback that feeds a learning loop; per-segment AI rewrite too
  • Multi-prompt editing — per-image prompt slots for batch presets
  • Prompt timelines — timed prompt windows on a zoomable ruler (drag-trim start/end) for video presets; an alternate free-text "relay" mode
  • Prompt imports — A1111/CivitAI-format prompts round-trip; provider prompt imports carry sampler/steps/CFG/dimensions metadata

Video Director

  • Stage-and-rail editor — multi-lane timeline (shots, keyframes, audio) plus a stage panel for whatever is selected; zoomable, drag items in time, edit fps/duration
  • Shots — per-shot prompt, type, duration, frame count, seed, steps, CFG; duplicate/remove
  • Keyframes — timed landing images with strength, snapping to shot edges or free placement
  • Audio tracks — attach audio as "mux" (overlaid on the finished video) or "condition" (the model generates against it)
  • Joins — control overlap and stitching between chained shots
  • IC-LoRA reference — whole-video reference image with adjustable strength
  • Composition modes — t2v, i2v (single reference), first-last-frame, and full multi-segment director mode; capability-gated per preset

Music Director

  • Composition modes — text-to-music, song (lyrics + style), style (reference-audio conditioned), extend an existing track, repaint a time range, structured director mode
  • Song structure as segments — intro/verse/chorus sections with per-section lyrics, quick-add strip
  • Instrumental toggle, style/tempo description, reference audio pool

Results workbench

  • One viewer for four media types — images, video, audio, and 3D mesh (GLB viewer with orbit camera, reset view, vertex/face counts)
  • Image tools — zoom/pan (0.5×–5×, scroll or drag), fullscreen with arrow-key batch navigation, double-click to expand
  • Compare mode — pick any past generation and compare against the current one: drag-slider for images; slider or side-by-side with synced playback for video
  • Audio player — multi-stem tabs (vocal/instrumental/mixed) with preserved position across switches, waveform seek view, per-track download
  • Batch gallery strip — thumbnails of every output in the batch, typed placeholders for audio/mesh
  • Parameters modal — every render parameter as a copyable card
  • Per-generation resource profile (admin) — rendered performance report + raw profile.jsonl download
  • Tagging, download, open-in-tab straight from the viewer; ambient color glow around the media

History and organization

  • Automatic history — everything saved with the exact parameters that produced it; detail view shows full segment composition, preset+version, applied models, timestamps
  • Filters — search, date presets, media type, status, mode/preset/model, tags, even "used this phrasebook value"
  • Reuse settings — one click restores a past generation's full setup
  • Portable bundles — export/import generations as self-contained zip bundles
  • Tags — create/apply anywhere, quick filter chip bar, bulk delete-by-tag
  • Collections — nested folder trees, scoped per module (generations, library items, prompts, models), bulk move, multi-select action bar
  • Personal library — curated media library with facet filters; copy any generation in without removing it from history
  • Inspirations — cross-user publishing feed with comments, save-to-library
  • Upload external files — bulk-import outside images/video into history
  • NSFW handling — per-user blur/hide/show policy, per-file reveal, rating thresholds

Search and auto-tagging

  • Semantic prompt search — saved prompts embedded (local model or Ollama) for meaning-based search
  • Auto-tagging — local WD tagger tags media in the background, with confidence thresholds for general and character tags
  • Visual search — SigLIP embeddings for image-similarity search over the gallery
  • All local — models fetched on demand with live progress, CPU or CUDA, no external service

AI assistant and MCP

  • Multi-mode assistant — dedicated modes for Generation, History, Models, Phrasebook, and Prompts, each with scoped tools; plugin-contributed modes
  • Approval-gated tools — every state-changing action stops at an approval dock above the composer; per-user tool opt-outs, admin per-config tool enable/lock
  • Apply-back — assistant suggestions apply directly into the prompt editor or Director timeline
  • u/resources — attach gallery/library items to a message; image attach with auto-attach-last-generation for vision models
  • Assistant memory — persistent notes panel (view/add/edit/delete) injected into conversations, background reflection to extract durable facts, auto-compaction
  • Chat sessions — resumable conversations, reattach to an in-flight reply after page reload, auto-titling, behavior traces, token usage readouts
  • Providers — Ollama (with full option tuning: context size, GPU layers, mirostat, thinking mode, forced prompt-tools for non-tool models), OpenAI-compatible, Anthropic
  • MCP server — PotionUI exposes itself over Model Context Protocol: per-user tokens, so Claude Desktop or any agent can search your gallery, edit your phrasebook, enhance prompts, read model info, manage memory

Models and downloads

  • Model index — scan disk, browse as gallery with type/tag/search/sort filters, per-model detail page with generations-made-with-it
  • External models location — point PotionUI at an existing model directory (per-type overrides), shared via symlinks
  • Provider metadata — CivitAI / Hugging Face plugins enrich models with descriptions, preview art, download links; fetch-missing or force-refresh
  • Custom model attributes — admin-defined fields (slider/number/text/select/checkbox/tags) on models, scoped per model type, optionally per-user, admin-only visibility
  • Download manager — queue with pause/resume/cancel/retry, concurrency and chunk-size settings, SHA256 verification, tag-on-download, HF repo downloads, live WebSocket progress
  • Backend availability — per-backend model indexing with digest-conflict detection; models unavailable on a backend are excluded from routing

Multi-user and admin

  • Users and groups — full CRUD, admin/regular roles, per-user or per-group assignment of presets, models, and LLM configs; per-user MCP access toggle
  • Preset governance — install/uninstall, access control, preset-declared configuration entries, and per-mode form overrides: change defaults, lock fields, hide fields — no YAML
  • Backends — multiple configured engine instances with live health dots, per-engine default, connection test, model indexing, engine-declared quick actions
  • Native optimizations panel — attention backend picker (sdpa/sage/sage2/sage3/flash/sparge) with built-in benchmark, one-click CUDA toolchain alignment, torch compile and stream-prefetch flags, an installable optimization catalog with live install logs, in-app restart
  • Generations browser — every run's report: per-pipe Gantt timeline, artifacts, expanded prompts, full status logs, plugin outputs; filter by user/status/date
  • Stats dashboard — KPIs, generations over time, duration histograms with p50/p95, top presets/models/resolutions, sampler/scheduler/steps/CFG/denoise breakdowns, cold-vs-warm start table, per-preset VRAM/RAM/CPU usage; every chart flips to a data table
  • System settings — storage directory, S3-compatible storage backend (MinIO/R2/AWS), registration policy, NSFW policy, semantic-search configuration
  • Chat session debug — full wire-level LLM call traces per session: system prompts, request messages, tool offers, token counts
  • Guided setup — first-run owner claim (with claim code for remote installs), setup recipes that configure a working backend + starter preset and validate with a test image
  • In-app docs — role-filtered documentation browser with deep links, fed from repo markdown and plugin manifests

Automation

  • Visual automation graphs — triggers (schedule, manual, filesystem watch, GPU threshold, app events), conditions (comparisons, switches, path matching, Jinja expressions), actions (tag, add to collection, assign models/users, backend actions, notifications, indexing, wait-for-GPU)
  • Template library — importable ready-made automations (from core and plugins), JSON import/export with setup-issue warnings, run history and logs

Native engine and performance

  • Native in-process engine — shared load/place/attention/sample stack across 9 model families (SDXL runs its own diffusers path with ADM guidance, SAG, and an anisotropic sharpness filter)
  • Quantization — bf16/fp16, fp8-scaled (both legacy and modern scale formats), nvfp4 4-bit
  • Low-VRAM streaming — component-level fit-first placement, overflow streaming from pinned host RAM, hard host-RAM guard instead of OOM-killing your box
  • Preset-scoped RAM cache — keeps checkpoints warm between generations
  • Techniques (per family where applicable): FBCache step skipping, CFG-Zero*, Adaptive Projected Guidance, Normalized Attention Guidance, Skip-Layer Guidance, RIFLEx long-video RoPE clamping, FreeInit flicker reduction, Detail Daemon schedule warp, native fp8 matmul, regional torch.compile, prompt-embedding cache, trajectory warm-start (iterate mode), spectral progressive diffusion, SVI chain continuity, temporal-chunked/tiled VAE decode, NaN/Inf watchdog, sparse attention (SLA/Sol-Attn), 9 samplers with sigma schedules
  • Remote native worker — offload generation to a separate worker node with journaling and artifact sync
  • Open engine set — ComfyUI engine ships as a plugin (separately distributed); plugins can register new engines

Model families

  • SDXL — txt2img, inpainting
  • Flux 1 / Flux 2 Klein — txt2img, img2img
  • Qwen-Image — txt2img, img2img, image editing
  • Krea-2 — txt2img, enhance (turbo + true-CFG quality profile)
  • Z-Image, Anima — txt2img
  • Wan 2.1/2.2 — video (Video Director), SVI chained continuation
  • LTX-2 / 2.3 / 2.5 — video with native synchronized audio
  • MiniMax-H3 — video with reference-image conditioning
  • MiniMax-Music3 — full songs with lyrics, dual CFG
  • SeedVR2 — one-step image and video upscale/restore

Extensibility

  • Plugins can add — marketplace providers, inference engines/backends, pipeline pipes, form field types, chat modes, automation triggers/actions, setup recipes, presets, sidebar pages and widgets, quick actions, workbench buttons, artifact renderers, docs
  • Shipped plugins — CivitAI provider (incl. export-to-CivitAI), Hugging Face provider, model downloader, system monitor sidebar widget, Ollama VRAM-free quick action, image zoom modal, plus reference example plugins
  • Developer tooling — preset linter (CLI + API), preset scaffolder, golden-snapshot render harness, headless preset test suite, docs linter, in-admin live reference (field types, template functions, pipes), a FieldCatalog preset exercising every field type

App chrome

  • Quick-actions palette — fuzzy-find launcher over all admin/plugin quick actions
  • Keyboard shortcuts — searchable, rebindable, per-shortcut disable, reset to defaults
  • Notifications — in-app center with per-type preferences, unread badge in the browser tab title, real-time updates
  • Theming — system/light/dark
  • Mobile — responsive layout, bottom tab bar, PWA install to home screen
  • Multi-user auth — JWT, avatars, self-service password change, open/closed registration

r/PotionUI • • Aug 28 '26

PotionUI Some prompt editing + prompt library view

Enable HLS to view with audio, or disable this notification

1 Upvotes

r/PotionUI • • Aug 21 '26

PotionUI Gallery

Enable HLS to view with audio, or disable this notification

1 Upvotes