MiniMax H3 Drains the Last Moat in Open Video
MiniMax released H3, the model behind Hailuo 3.0, on July 31, generating up to 15 seconds of 2K video with native stereo audio in a single pass, with weights on Hugging Face and day-0 ComfyUI support …
MiniMax released H3, the model behind Hailuo 3.0, on July 31, generating up to 15 seconds of 2K video with native stereo audio in a single pass, with weights on Hugging Face and day-0 ComfyUI support …
ComfyUI, the node-based AI creation engine created by Yannik Marek, added native MiniMax H3 workflows on Aug. 2, turning the video model into editable graphs with local-runtime support, claiming it ca…
More than one in 10 of hundreds of Hollywood job postings in late June were likely connected to AI, with top studios like Disney, Netflix, and Amazon recruiting for AI tool development while publicly …
PixLog, an open-source tool by developer Zhao Xuan, adds pixel-aware visual history to Git without creating a second commit graph or staging area. It tracks exact bytes, changed regions, coordinate bl…
A developer built iTube, a pipeline that turns entire novels into narrated, lip-synced motion videos on a single 16GB GPU using FLUX, Wan2.2, PuLID, MuseTalk, and ComfyUI. The key insight is that audi…
A developer argues that major AI companies are deliberately suppressing advanced sampling techniques like min-p, mirostat, and dynamic temperature, offering users only basic controls while hobbyist to…
A workstation with 128GB of unified memory can run ComfyUI locally with models like Qwen Image and LTX Video, rendering images in roughly 20 seconds and enabling unlimited AI content generation withou…
A local ComfyUI pipeline using Qwen Image and LTX Video can batch-generate AI video clips overnight for free, avoiding per-second cloud API costs. The scripted workflow runs on a mid-range unified-mem…
AMD's Ryzen AI Developer Center, bundled with Ryzen AI Max Plus 395 machines, eliminates manual ROCm and driver setup for local AI workloads, offering guided playbooks for ComfyUI, LM Studio, and Unsl…
Researchers propose a knowledge-centric framework for workflow generation in visual creation systems like ComfyUI, achieving richer node diversity, more coherent structures, and higher execution succe…
AMD's ROCm blog explains that choosing the right attention algorithm in ComfyUI can significantly impact performance, memory usage, and stability on AMD GPUs, with PyTorch SDPA recommended as the safe…
ComfyUI's shared Python environment creates dependency conflicts, as exemplified by Qwen3-TTS requiring transformers==4.57.3 while newer models need Transformers 5.x, causing workflows to break. The a…
A user requests an INT8 ConvRot version for FireRed Image Edit v1.1, citing text artifacts and facial drift with standard FP8 variants, and suggests using ComfyUI's native ConvRot support for a quanti…
ComfyUI is a node-based visual programming interface for AI image generation that lets users build custom pipelines by connecting processing nodes instead of writing code, supporting models including …
AMD announced that with ROCm 7.2.1, users can now run PyTorch and ComfyUI natively on Windows on an AMD Ryzen AI Max+ processor, driving the integrated AMD Radeon 8060S GPU directly. The unified memor…
A comprehensive comparison of the top AI image generators in 2026, including Stable Diffusion, Flux, ComfyUI, SDXL, Kandinsky, DALL-E 3, Midjourney, and Leonardo AI, evaluates them on quality, control…
A developer investigating FLUX.1 [dev] for commercial use found its non-commercial license and questioned whether AI model vendors can track self-hosted deployments. The article concludes that technic…
Black Forest Labs released FLUX.2-klein-4B, an image-to-image model under the Apache-2.0 license, now available on Hugging Face with over 470,000 downloads. The model is designed for use with diffuser…
An anonymous developer recounts the four-year history of the AUTOMATIC1111 Stable Diffusion WebUI, from its creation in August 2022 to its decline by 2026. The project, which began as a single Gradio …
Nvidia AI Labs researcher Ziv Ilan presented at GTC 2026 that video diffusion models can achieve real-time performance without 50 denoising steps by using a stack of quantization, caching, and distill…