ai images

Gaze-LLE: Gaze Target Estimation via Large-Scale Learned Encoders

submitted by /u/hippynox [link] [comments]

1 year ago

Real time video generation is finally real

Introducing Self-Forcing, a new paradigm for training autoregressive diffusion models. The key to high quality? Simulate the inference process during…

1 year ago

I dunno how to call this lora, UltraReal – Flux.dev lora

Who needs a fancy name when the shadows and highlights do all the talking? This experimental LoRA is the scrappy…

1 year ago

Chatterbox TTS fork *HUGE UPDATE*: 3X Speed increase, Whisper Sync audio validation, text replacement, and more

Check out all the new features here: https://github.com/petermg/Chatterbox-TTS-Extended Just over a week ago Chatterbox was released here: https://www.reddit.com/r/StableDiffusion/comments/1kzedue/mod_of_chatterbox_tts_now_accepts_text_files_as/ I made…

1 year ago

The 8 Rules of Open-Source Generative AI Club!

Fully made with open-source tools within ComfyUI: - Image: UltraReal Finetune (Flux 1 Dev) + Redux + Tyler Durden (Brad…

1 year ago

Elevenlabs v3 is sick

This's going to change the face how audiobooks are made. Hope opensource models catch this up soon! submitted by /u/pheonis2…

1 year ago

This sub has SERIOUSLY slept on Chroma. Chroma is basically Flux Pony. It’s not merely “uncensored but lacking knowledge.” It’s the thing many people have been waiting for

I've been active on this sub basically since SD 1.5, and whenever something new comes out that ranges from "doesn't…

1 year ago

World War I Photo Colorization/Restoration with Flux.1 Kontext [pro]

I've got some old photos from a family member that served on the Western front in World War I. I…

1 year ago

Flux Kontext Images — Note how well it keeps the clothes and face and hair

submitted by /u/FitContribution2946 [link] [comments]

1 year ago