Lore

AI image generation

2 INSANE AI TOOLS: X-Portrait 2 + Detail Daemon

The video showcases two separate AI tools for visual creators: X-Portrait 2, an unreleased face-animation model that transfers a real face's tracked motion (expression, eyes, mouth) onto a stylized input image to produce highly expressive animated video, and Detail Daemon, a ComfyUI custom node that lets users tune how much fine detail is added to Flux/SDXL/SD1.5 generations by adjusting the noise schedule during sampling.

Olivio Sarikas · 2024-11-08 · English

Key ideas

  1. X-Portrait 2 is not yet released; it works by tracking a real input video's facial motion and applying it to a separate input image

  2. The tool tracks facial features in detail — eyes (where they're looking), mouth, and overall expression — producing what the video calls massively expressive results

  3. The same input image can be rendered in completely different visual styles while retaining the tracked expressiveness, opening possibilities like independent films, series, or becoming a virtual YouTuber

  4. Olivio frames X-Portrait 2 as evidence that AI is not replacing artists but working alongside them, especially empowering independent artists

  5. Detail Daemon installs via ComfyUI Manager: copy the GitHub URL, open the Manager, use 'Install from git URL', paste the link, and restart ComfyUI

  6. The Detail Daemon package ships with example workflows; the 'comparing detailers' workflow is recommended as the starting point

  7. The GitHub page's notes section documents all Detail Daemon Sampler settings plus suggested value ranges, which differ by model (Flux vs. SDXL vs. SD1.5)

  8. A duplicate-looking note in the workflow exists only to show a graph visualizing how detail/noise is adjusted across sampling steps — it is not required to use the node

  9. Detail can be weighted toward the start of generation (bigger, rougher details) or the end (smaller details) by adjusting this schedule

  10. At low strength (0.1) changes are subtle, e.g. improved eye/feather detail; strength increases progressively add more background detail, which matters because Flux otherwise produces strong background bokeh with little detail

  11. Pushing detail strength too high overexposes the image, and pushing further eventually converts the output into a drawing/illustration rather than a photo

  12. A second test example (a creature image) shows Detail Daemon can fix anatomical issues (e.g., restoring a missing arm) and add glowing/fur detail before ultimately degrading into a drawing-like 'error' result

  13. Beyond the Detail Daemon Sampler, two other methods exist: the Lying Sigma Sampler (three values: dishonesty factor, start, stop) and Multiply Sigma (three values: factor, start, end)

  14. The Lying Sigma Sampler is described as easier to understand with fewer values but potentially harder to control precisely

  15. Olivio characterizes Flux as a 'completely different Beast' requiring more tweaking and tools than other models, but notes Flux LoRA/model training is comparatively simple and approachable despite being more limited in scope

  16. X-Portrait 2 — An unreleased face-animation model that tracks a real input video's facial motion (eyes, mouth, expressions) and transfers it onto a separate input image to produce an expressive animated video. Apply: Supply an input image and an input video of a real face performing desired expressions to generate a stylized, expressive animated output, per the demos shown on the tool's website (not yet publicly released as of the video).

  17. Detail Daemon Sampler — A ComfyUI custom node/sampler that adjusts the amount of fine detail in Flux/SDXL/SD1.5 generations by modifying the noise schedule across sampling steps, with a strength setting and other adjustments documented with suggested ranges (which vary per model) on its GitHub page. Apply: Install via ComfyUI Manager's 'Install from git URL,' open the included 'comparing detailers' example workflow, and increase the detail strength gradually from a low value like 0.1, watching for the point where results overexpose or convert into a drawing-like image.

  18. Lying Sigma Sampler — An alternate Detail Daemon method with three values — a 'dishonesty factor' plus a start and stop point — described as simpler to understand than the Detail Daemon Sampler but harder to control precisely. Apply: Use as a lighter-weight alternative to the Detail Daemon Sampler; adjust the dishonesty factor and its start/stop range and experiment with the resulting detail changes.

  19. Multiply Sigma — A third Detail Daemon method with three values — a factor plus a start and an end point — for adjusting the noise/detail schedule during sampling. Apply: Tune the factor and start/end values as another way to control detail strength within the same Detail Daemon workflow, alongside or instead of the other two methods.

Insights

The duplicate note in the Detail Daemon workflow is purely a visualization wrapper around a noise-schedule graph, which Olivio explicitly says should have been merged into the actual settings note — a UI/documentation gap he calls out directly

Detail Daemon's value is partly a workaround for a specific Flux weakness (strong background bokeh with little detail) rather than a universal enhancer applicable identically to every model

Pushing the detail strength parameter follows a predictable degradation arc — subtle change, then more detail, then overexposure, then conversion into a drawing — suggesting the setting has a defined, repeatable curve rather than random behavior at extremes

Olivio treats the 'broken' extreme outputs (drawing-style conversion, over-brightened backgrounds) as aesthetically valuable results in their own right, not simply failure states to avoid, explicitly saying he's 'pretty much in love' with an erroneous output

«this is not released yet but what you can see on the screen is how this can animate faces and you can see how massively expressive this is»

— 00:12

«one of the biggest details you have with stable fusion with a eye image generation and especially video generation is the emotional expression of the faes»

— 00:26

«that AI is not placing artists it is working alongside artists and especially empowering Independent Artists to do more with their skills and their creative ideas»

— 02:01

«the detail demon sampler this gives you a lot of different adjustments it looks a little bit overwhelming but don't worry it is pretty easy to understand»

— 03:10

«at a certain point it turns into a drawing which I find actually wonderful because you can get some pretty cool results»

— 07:34

«flux is a completely different Beast that needs more tweaking at more tools but the other hand also has more Simplicity in that»

— 09:40

Reception

Strong enthusiasm for the technical capabilities, but practical concerns about hardware requirements, workflow complexity, and accessibility dampen otherwise positive reception.

The video functions as a practical two-part tool tutorial rather than a deep technical explainer, leaning on visual before/after comparisons for Detail Daemon and promotional framing for the unreleased X-Portrait 2; its value is mainly as a getting-started pointer for ComfyUI users, not a rigorous breakdown of either tool's underlying method.

10:24

↳ Olivio Sarikas · YouTube

Watch original