qwen 2.1 is very good upscaler
This test used frames taken from H3 generated videos on 0.3MP on my 6GB VRAM. I don’t know if you can get better faces from distance if you upscale over 2K. submitted by /u/dev_ne [link] [comments]
Category Added in a WPeMatico Campaign
This test used frames taken from H3 generated videos on 0.3MP on my 6GB VRAM. I don’t know if you can get better faces from distance if you upscale over 2K. submitted by /u/dev_ne [link] [comments]
In exactly two months, it will be 1 year since the release of Nano Banana Pro, and we still don’t have a comparable local model when it comes to editing. Google really nailed its reference capabilities. You send a good face reference to it and it can generate a totally different photo of the person …
Read more “Still wishing on a local editing model that can compete with NBP”
Hey guys! So, I got early access to Qwen Image 2.1, and I did what any sane person would do, generate 200+ “1girl” images (while testing keywords and prompt structures). Here are some amateur-style images that I generated. If you want me to run any of your prompts, just put them in the comments. My …
submitted by /u/dev_ne [link] [comments]
But unfortunately some of the icons goes missing. Didn’t exactly follow the prompt either. Prompt (10 sec each) – subject_definitions: <Subject 1> is the centered young woman from <Picture 1>, preserving her exact facial identity, shoulder-length dark hair, natural skin appearance, oversized cream sweatshirt with the cat illustration and visible text “Good Night”, plaid pajama …
First, in Minimax, I use a prompt like this: “The character remains completely frozen in place, perfectly still like a statue. The camera smoothly orbits 360 degrees around the character in one continuous shot. No character movement, no pose change, no cuts.” Then I use COLMAP: Create a new database and import the frames extracted …
submitted by /u/Rokkit_man [link] [comments]
YuE2 is an impressive open music model. Give it a style prompt and lyrics, and it can make a complete song. Under the hood, it generates “semantic tokens” that are then turned into audio. The trouble is, the encoder that converts existing recordings into those same tokens was never released. That leaves you without a …
Read more “I trained the missing encoder for YuE2, so we can all bring our own music into it”
The lora itself at 3 steps is nothing to write home about. If the scene isn’t mostly static, you can expect slowmo jerky motion, smearing and straight broken output with butchered sound. HOWEVER, it has VERY good visual quality without obvious overcooking plaguing the turbo loras. You add as much steps of non accelerated generation …
submitted by /u/malcolmrey [link] [comments]