'They Met Again' demo shows audio reference in Stable Diffusion
Reddit user GTManiK shared a Stable Diffusion demo titled 'They Met Again', commenting that audio reference works.
Tagged
Checkpoints, samplers, LoRAs and the techniques the community is sharing. Curated and summarized from dozens of sources by AIBriefs.
Reddit user GTManiK shared a Stable Diffusion demo titled 'They Met Again', commenting that audio reference works.
Users report that the H3 model often generates anatomical errors, suggesting the use of reference images to improve output accuracy. This method is presented as a workaround for current limitations in early LoRA fine-tunes.
Reddit user Hoppss posted a quick Stable Diffusion generation riffing on the 'Jerry, we must cook!' meme to r/StableDiffusion, collecting 37 upvotes and 8 comments.
AI-generated crossover images place Dwight from The Office in Central Perk, enthusiastically explaining AI threats to a sarcastic Chandler from Friends, set in the nineties before smartphones existed.
In r/StableDiffusion, a user praised developers for 'pulling all-nighters' to ship a release, linking to a tweet from RyanLeeMiniMax. The post adds 'It helps everyone! Thank you!'
A Reddit post relays that Comfyanon says, as far as they know, H3 will still release — and tells the community not to trust the timer.
Style LoRA focused on split waterline / over-under compositions, crystal-clear water, and pools and lakes. It began as a few-hour personal project to create a single image; the author released it for others to use.
An r/StableDiffusion user ran the Will Smith benchmark with an AI model and rated the output "Pretty good."
The Forge Neo platform now enables users to apply multiple character LoRAs to specific, distinct regions of a single generated image. This feature allows for custom spatial masking, such as top/bottom or left/right placement, for separate LoRA models.
A Reddit user demonstrates how different LoRA files change the style of the same prompt with 13 examples.
Reddit user MASilverHammer adapted a LoKr config from u/LilBrownBebeShoes, enabling Differential Output Preservation with the class "woman" to successfully train multiple character LoRAs in a single run.
A user trained a LoKr adapter on the Krea2 model using a 43-image dataset and Qwen3 VL 4B instruct for captioning. The training was conducted over 6 epochs with a 0.0001 learning rate.
A detail enhancement LoRA for skin texture in Stable Diffusion, published on Civitai and Hugging Face alongside a dataset and training guide.
Node built with Claude lets users craft prompts, randomize settings, and save/share presets. Author created it after being unsatisfied with existing options.
A workflow for the Anima model enables restyling characters with different artistic styles while preserving identity. Uses a reference image and includes all model links on CivitAI.
A community member shares a first attempt at a Krea 2 style LoRA for Stable Diffusion. The post on Reddit showcases the LoRA and asks for feedback on sharing training data.
A user spent a year developing a free fine-tuning trainer for SDXL and Anima models that runs on a 12 GB GPU. The tool addresses common limitations like forced lower resolution and complex config files required by other trainers.
A refined VAE variant offers crisper edges and stronger micro-detail without altering colors or composition. Released by community member Merserk13.
Workflow for generating videos from storyboard image panels using LTX 2.3. Includes 3×5 loader, panel selector, and automatic processing; author seeks testers.
A set of VFX tools for ComfyUI that allows artists to control AI generation using light, camera, perspective, depth, and 3D placement. The tools are designed to art-direct model outputs with traditional VFX craft rather than compositing nodes.
Guide recommends 20-40 high-quality images with full-body shots and settings to avoid overtraining. Achieves near-perfect likeness in 750 steps.
A 4-stage AI pipeline for transforming, swapping, and restyling characters across images and video. The workflow features character stripping, face swapping, and style transfer, all within ComfyUI.
Custom ComfyUI node suite repairs one or many faces using the connected checkpoint, VAE, and prompts instead of a fixed face-restoration model. Regenerates each detected face in the same visual style as the original image.
Guide to optimizing Krea2 output with specific sampler and scheduler choices. Krea2 uses PDE-based signal processing to prioritize visual feel, texture, and mood over strict prompt adherence.
A Reddit user reports that using 4 Raw steps followed by 4 Turbo steps in Krea2 gives better colors than using the Turbo LoRA alone. The post includes example images.
The updated node addresses flaws identified in the previous version of the Krea2 prompt integration for Stable Diffusion workflows.
Vanilla Krea 2 Turbo's censoring hinders facial expressions, but a Reddit user finds simple facial positioning fixes suffice. Detailed face descriptions are rarely needed.
A Reddit post explains that 'preserve the face' prompts do not work in image edit models like Klein and Qwen. Instead, the post proposes a mental model and three techniques that actually preserve identity during edits.
A Reddit user shared 177 facial expression prompts for Krea2 models. The prompts are designed for consistent character expression with the same seed, using Krea2_turbo_lora and TextFusion Refusal Reduction loras.
A Reddit user trained a VAE for Stable Diffusion 1.5 that renders text better than the original. The model is available on HuggingFace.
A modified workflow for Kandinsky5 Lite I2V optimized for 4GB GPUs generates 5s videos at 675×900 with 8-12 steps. Adapted from the official workflow for lightweight hardware like RTX 3050 Ti mobile.
MemoryWorks VHS v1.1 is now available on Civitai, offering improved VHS-style aesthetics for image generation. The update represents a significant step forward from the original experimental release.
A Reddit user deployed the Waypoint 1.5 world model locally, generating real-time video through a custom UI that feels like a video game. Code is available on GitHub via the worldmodel.c repository.
A community user shares a quantized int8 version of Krea 2 Raw optimized for 12GB VRAM GPUs. Recommended settings include LoRA Turbo at 0.60 strength, 12 steps, and CFG 1.5 at resolutions up to 1024x1536.
User alphama00 asks for tips on generating consistent good images with SDXL Basic and Juggernaut XL models, reporting distorted results. Community discussion provides advice on settings and workflows.
A new Prompt/Style Selector node for ComfyUI, including krea2 presets, is now available on GitHub. The node enables batch prompt processing with style presets, created by community developer berlinbaer.
Reddit user Suspicious_Aide2697 turned the 285 Krea2 style wildcards into a ready-to-use node, crediting the original r/StableDiffusion wildcard collection as the source. The author noted the styles originally shipped as wildcard text files.
A Reddit user shares a PDF guide on making AI videos feel cinematic, emphasizing a filmmaking approach over prompt engineering. The workflow covers techniques to add emotional depth and visual quality.
User AxonkaiLab shares AI-generated 80s-style Star Wars candid street photography on Reddit. Part 2 features more anachronistic scenarios with a vintage 35mm monochrome look. The post includes a humorous reference to Kylo Ren as an 'emo kid'.
LoRA trained with 2220 steps on 37 images using the base/Raw version of the model. Image resolution set to 512x768.
User shares a style LORA trained to blend images while preserving composition. Download from Huggingface with workflow included.
A Reddit user asks if Klein Edit is still the best tool for image editing, noting issues with color preservation and character replacement quality. The community discussion highlights ongoing challenges despite the tool's initial promise.
Two rank-32 functional LoRAs for Krea 2 were released with Diffusers pipelines. They teach image-conditioning behaviors (identity reference and positional outpainting) and include runnable examples.
Trained on 40 low-res stills from 60s-70s shows like Thunderbirds. Uses Ai-Toolkit to generate images in the Supermarionation style.
A Reddit user trained and shared an art style LoRA for Krea2 on Civitai, inspired by an Instagram reel. The model has been well-received, with the user noting heavy usage since Flux1.Dev.
LoRA trained on the artist's style produces black ink, watercolor, and smooth illustrations. Full dataset included on CivitAi page.
The BRKN-PROMPTER-RANDOMIZER is a beta tool that randomizes prompts for Stable Diffusion. It will be released open-source this Friday, as announced by a developer on Reddit.
A Reddit user reports that many issues with Krea2 can be fixed by disabling active LoRAs. The user found that even popular LoRAs can be the culprit, and contradictory prompts are also a common problem.
New Krea2PromptWeight node in KJ nodes pack replaces text prompt encoder and carries through prompt weights. Findings show CFG >1 behavior changes, improving control over generation.
A Reddit user forked AI Toolkit and integrated SAM 3D body scanning to improve body shape learning during LoRA/Lokr training. Training a Lokr with body data takes roughly 60 minutes on an RTX 5090.
Researchers analyzed 6M AI-tagged Pixiv images, covering 22,400 base models and 154,000 LoRAs, to study real-world usage patterns. The paper provides insights into how the community selects and combines models for image generation.
User shares Soviet-themed images generated with Krea 2 Turbo FP8 on an RTX 3070 Ti (8GB VRAM). The images use Realism Engine v2 Lora and require 64GB RAM.
Fastest at 1024x1024 is int8_convrot at 2.424 s median, while int4_convrot is best for speed/VRAM at 20.6 GiB. At 2048x2048, int4_convrot leads at 12.678 s median, 1.25x faster than BF16.
A new LoRA for Stable Diffusion trained specifically for extending image backgrounds while preserving original composition. Designed for background modification rather than character alteration.
A Reddit user shares results showing improved character consistency in text-to-image generation using Krea2's model variant. The post attributes the consistency to reduced variety in the model's training.
Civitai now requires VPN access and has strict rules on celebrity likenesses, prompting users to ask where to share character and celebrity LoRAs. The community discusses alternative platforms and the impact of tightening content policies.
One week after releasing his first Krea 2 analog LoRA, the user retrained it based on community feedback. The updated LoRA addresses issues pointed out by the StableDiffusion subreddit.
The krea2-identity-edit model, available on HuggingFace, now supports outpainting in addition to its identity-consistent editing. A ComfyUI workflow is provided to use the model. Users can extend images while preserving the subject's identity.