SOTA Image Captioning Scripts For Stable Diffusion: CogVLM, LLaVA, BLIP-2, Clip-Interrogator (115 Clip Vision Models + 5 Caption Models)
Patreon exclusive posts index to find our scripts easily, ...
2024-10-02 11:00:00 +0000 UTC View Post
Patreon exclusive posts index to find our scripts easily, ...
2024-10-02 11:00:00 +0000 UTC View PostMimicPC is a cloud service that lets you use AI applications seamlessly in your browser while everything runs on remote cloud servers, providing permanent storage as well
Full tutorial link : will be added hopefully once published
MimicPC Official S...
2024-09-25 09:19:31 +0000 UTC View PostYou can download our APP from here : https://www.patreon.com/posts/110613301
1-Click to install on Windows, RunPod and Massed Compute
Patreon exclusive posts index to find our scripts easily, ...
2024-09-21 15:31:55 +0000 UTC View PostPatreon exclusive posts index to find our scripts easily, ...
2024-09-20 00:49:09 +0000 UTC View PostFull Fine Tuning / DreamBooth of FLUX yields way better results than LoRA training as expected, overfitting and bleeding reduced a lot, check oldest comment for more information, images LoRA vs Fine Tuned full checkpoint
Full configs and grid files shared here : 2024-09-16 10:47:40 +0000 UTC View Post
Medium article : https://medium.com/@furkangozukara/ultimate-flux-lora-training-tutorial-windows-and-cloud-deployment-abb72f21cbf8
... 2024-09-14 22:37:09 +0000 UTC View PostPatreon exclusive posts index to find our scripts easily, ...
2024-09-13 15:33:00 +0000 UTC View PostStarted training my 256 images on FLUX with 8x GPU - Dataset is not ready yet, not very good sharpness and lightning but so many people asking expressions so i am taking a break from research :) - Going up to 200 epochs I also wonder results
2024-09-10 21:54:14 +0000 UTC View PostNOW WE HAVE A MUCH BETTER NEW APP PLEASE USE IT
https://www.patreon.com/posts/112126955
NEW APP : https://www.patreon.com/post...
2024-09-10 20:00:00 +0000 UTC View PostPatreon exclusive posts index to find our scripts easily, ...
2024-09-10 16:49:36 +0000 UTC View PostI have done total 104 different LoRA trainings and compared each one of them to find the very best hyper parameters and the workflow for FLUX LoRA training by using Kohya GUI training script.
You can see all the done experiments’ checkpoint names and their repo links in following public post: 2024-09-09 23:49:10 +0000 UTC View Post
I have been using a machine on RunPod that has 8x RTX A6000 over a week now. I have done all the following experiments with each one being 3000 steps.
Latest configs shared on : https://www.patreon.com/posts/kohy...
2024-09-09 21:00:00 +0000 UTC View Post
CivitAI Link : https://civitai.com/models/731347
Hugging Face Link : https://huggingface.co/MonsterMMORP...
2024-09-08 00:11:27 +0000 UTC View PostRope NEXT Update Has Arrived - both Rope Live and Rope Alucard Versions Are Updated
Both works with CUDA 11.8 and also CUDA 12.4
Both works on Massed Compute perfectly
You can download installer zip files here with instructions : 2024-09-07 01:32:06 +0000 UTC View Post
The installer zip files are provided in this post : https://www.patreon.com/posts/103765029
It took huge time to make FaceFusion NEXT work with onnx
Python 3.10, FFmpeg, Cuda 11.8 2024-09-07 00:41:01 +0000 UTC View Post
Raw FLUX LoRA Output vs SwarmUI Quick 2x Upscale Button vs Our SUPIR APP Upscale (Face Upscale ON) : https://imgsli.com/MjkzODAz/2/1
Our SUPIR APP : https://youtu.be/OYxVEvDf284...
2024-09-03 22:16:06 +0000 UTC View PostI started training a public LoRA style (2 seperate training each on 4x A6000).
Experimenting captions vs non-captions. So we will see which yields best results for style training on FLUX.
Generated captions with multi-GPU batch Joycaption app.
I am showing 5 examples of what Joycaption generates on FLUX dev....
2024-09-02 23:48:09 +0000 UTC View PostI will update this page as I progress. I am just starting so I will add the info. Nothing ready yet
FOLLOW THIS THREAD NOW FOR TUTORIAL : https://www.patreon.com/posts/110879657
Multi-GPU batch caption with JoyCaption. JoyCaption uses Meta-Llama-3.1–8B and google/siglip-so400m-patch14–384 and a fine tuned image captioning neural network.
Link : https://www.patreon.com/posts/110613301
Link for batch caption...
2024-08-26 03:30:13 +0000 UTC View PostHere a Hugging Face space that you can test it yourself : https://huggingface.co/spaces/fancyfeast/joy-caption-pre-alpha - still working
I have been requested to make a Gradio app for this so i mad...
2024-08-24 02:13:07 +0000 UTC View PostResShift: Efficient Diffusion Model for Image Super-resolution by Residual Shifting (NeurIPS 2023, Spotlight)
Official Repo : https://github.com/zsyOAOA/ResShift
I have developed a very advanced Gradio APP.
... 2024-08-19 01:37:52 +0000 UTC View PostPatreon exclusive posts index to find our scripts easily, ...
2024-08-18 21:57:40 +0000 UTC View PostScripts are shared here : https://www.patreon.com/posts/forge-web-ui-and-110323512
Uses official Pytorch 11.8 template of RunPod
The command line arguments are also automatically optimized for 24 GB and above VRAM GPUs...
2024-08-18 12:42:21 +0000 UTC View PostPatreon exclusive posts index to find our scripts easily, Patreon s...
2024-08-16 22:00:00 +0000 UTC View PostI just updated the automatic FLUX models downloader scripts with newest models and features. Therefore I decided to test all models comprehensively with respected to their peak VRAM usage and also their image generation speed.
Automatic downloader scripts : 2024-08-15 03:25:49 +0000 UTC View Post
I have been recently contacted by the original SUPIR developers that they published SUPIR online with better features and they are generated funds from this service to develop SUPIR v2 and SUPIR video.
As you know, SUPIR is the best image upscaler available right now both commercially and open source. It is next level c...
2024-08-15 00:39:17 +0000 UTC View PostPatreon exclusive posts index to find our scripts easily, ...
2024-08-14 21:11:00 +0000 UTC View PostAuraSR is a 600M parameter upsampler model derived from the GigaGAN paper. It works super fast and uses a very limited VRAM below 5 GB. It is deterministic upscaler. It works perfect in some images but fails in some images so it is worth to give it a shot.
GitHub official repo : 2024-08-14 21:10:02 +0000 UTC View Post