SDXL consists of an ensemble of experts pipeline for latent diffusion: In a first step, the base model is used to generate (noisy) latents, which are then further processed with a refinement model (available here: https://huggingface.co/stabilityai/stable-diffusion-xl-refiner-1.0/) specialized for the final denoising steps. Note that the base model can be used as a standalone module. Alternatively, we can use a two-stage pipeline as follows: First, the base model is used to generate latents of the desired output size. In the second step, we use a specialized high-resolution model and apply a technique called SDEdit (https://arxiv.org/abs/2108.01073, also known as "img2img") to the latents…
Open weights
openrail++
2.6B parameters
diffusers
Animagine XL 4.0, also stylized as Anim4gine, is the ultimate anime-themed finetuned SDXL model and the latest installment of Animagine XL series. Despite being a continuation, the model was retrained from Stable Diffusion XL 1.0 with a massive dataset of 8.4M diverse anime-style images from various sources with the knowledge cut-off of January 7th 2025 and finetuned for approximately 2650 GPU hours. Similar to the previous version, this model was trained using tag ordering method for the identity and style training. With the release of Animagine XL 4.0 Opt (Optimized), the model has been further refined with an additional dataset, improving stability, anatomy accuracy, noise reduction…
Open weights
openrail++
2.6B parameters
diffusers
This repository contains a model that generates highly aesthetic images of resolution 1024x1024, as well as portrait and landscape aspect ratios. You can use the model with Hugging Face Diffusers. Playground v2.5 is a diffusion-based text-to-image generative model, and a successor to Playground v2. Playground v2.5 is the state-of-the-art open-source model in aesthetic quality. Our user studies demonstrate that our model outperforms SDXL, Playground v2, PixArt-α, DALL-E 3, and Midjourney 5.2. For details on the development and training of our model, please refer to our blog post and technical report. Install diffusers >= 0.27.0 and the relevant dependencies. - The pipeline uses the…
Open weights
other
2.6B parameters
diffusers
/ FIXED: Changed from 50% to 33.33% because there are 3 columns / margin-bottom: 1em; / Added small margin for spacing between stacked images / font-weight: bold; / Corrected 'font-style: bold' to 'font-weight: bold' / } / FIXED: Added missing closing brace here /.overlay, Animagine XL 3.1 is an update in the Animagine XL V3 series, enhancing the previous version, Animagine XL 3.0. This open-source, anime-themed text-to-image model has been improved for generating anime-style images with higher quality. It includes a broader range of characters from well-known anime series, an optimized dataset, and new aesthetic tags for better image creation. Built on Stable Diffusion XL, Animagine XL 3.1…
Open weights
openrail++
2.6B parameters
diffusers
Original model is here. This model created by Crody.
Open weights
other
2.6B parameters
diffusers
This is an image generation model based on training from Illustrious-xl. It utilizes the latest full Danbooru and e621 datasets for training, with native tags caption. For quality tags, we evaluated image popularity through the following process: - Data normalization based on various sources and ratings. - Application of time-based decay coefficients according to date recency. - Ranking of images within the entire dataset based on this processing. Our ultimate goal is to ensure that quality tags effectively track user preferences in recent years. - Latest Danbooru images up to the training date(for v1.0,it mean approximately before 2024-10-23) - E621 images e621-2024-webp-4Mpixel dataset on…
Open weights
other
2.6B parameters
diffusers