Tag
The Qwen Image 2.1 Consistency LoRA is a tool that maintains image consistency during edits, preventing drift and unwanted repaints in AI image generation workflows.
Akatz Labs released an experimental MiniMax H3 character-swap LoRA (1,000 training steps) that replaces a selected character in a source video using an image reference while preserving the scene. The authors note background preservation is decent, but motion timing, facial expressions, and hard cuts remain unreliable.
The user experimented with running Qwen 27B and Qwen Image 2.1 on local GPUs to generate 3D models and renders for printing, using tools like CADQuery and ComfyUI, and plans future projects for creating models from photos.
This article presents GGUF versions of the Qwen-Image-2.1 Text Encoder, including fixes for loading in ComfyUI and recommended sampler settings for image generation.
GGUF quantizations of the Qwen-Image-2.1 model for local image generation using ComfyUI, with recommended quantizations and setup instructions for deployment.
Repackaged model files for Qwen-Image-2.1 optimized for ComfyUI workflows, including text-to-image and image edit capabilities.
This article details the YuE2 model repackaged for ComfyUI, enabling audio and music generation workflows. It provides model files based on MERT-v2-FullSong and SheetSage2 for easy integration into ComfyUI projects.
This repository provides LoRAs for the MiniMax H3 model, designed to run in ComfyUI for video enhancement, such as sharpening videos while maintaining photorealism.
Tests indicate that dual Spark machines can run qwen3.8 and ComfyUI simultaneously, but the devices heat up severely.
Wan 3.0, an AI video generation model by Alibaba Cloud, is now available on ComfyUI, enabling generation of up to 30-second scenes in a single run with resolutions from 480p to 1080p.
A neural latent-space upscaler designed to accelerate high-resolution Minimax H3 video generation by upscaling latent representations directly, avoiding expensive decode-encode round-trips.
An open-source, self-hosted AI music generation radio tool called TAPEDECK that uses MiniMax Music 3 to create endless, personalized radio stations from a user's music library on a local GPU.
Minimax is preparing an open-weight release of Music 3, evidenced by active PRs in Diffusers and ComfyUI, with demos already available and a teaser from Comfy-Org suggesting a major launch within hours.
This article provides instructions for placing repackaged MiniMax-Music-3 model files into ComfyUI directories for music generation.
A Hugging Face repository hosting MiniMax-H3 models converted for ComfyUI usage, along with a Lightx2v distill LoRA for faster inference.
An early-preview LoRA for MiniMax-H3 enables 4-step audio-video generation instead of ~20 steps, yielding roughly 5x faster sampling, though quality is still immature.
Third-party ComfyUI-compatible LoRA conversions for MiniMax-H3 Turbo 4-step audio-video generation, including further-trained checkpoint-500 variants and an example workflow.
An early-preview LoRA for MiniMax-H3 that enables joint video and synchronized audio generation in 4 sampling steps instead of ~20, offering roughly 5x faster sampling, with ComfyUI custom nodes and three bf16 checkpoints.
Scenema Audio, an expressive text-to-speech model with zero-shot voice cloning, is now available as a native ComfyUI custom node, quantized to run on 8GB VRAM. The release adds inline stage direction cues, 12 preset voices, and simplifies the prompt format for ComfyUI.
A viral post highlights how an anonymous developer built ComfyUI, a powerful free open-source AI image tool, contrasting it with Adobe's subscription fees and recent DOJ settlement. It praises ComfyUI's free, local, GPL-3.0 releases.