This post summarizes a very important livestream with a WAN engineer. It will at least be partially open (model architecture, training code and inference code). Maybe even fully open weights if the community treats them with respect and gratitude, which is also what one of their engineers basically spelled out on Twitter a few days ago, where he asked us to voice our interest in an open model but in a calm and respectful way, because any hostility makes it less likely that the company releases it openly.
The cost to train this kind of model is millions of dollars. Everyone be on your best behaviors. We’re all excited and hoping for the best! I’m already grateful that we’ve been blessed with WAN 2.2 which is already amazing.
PS: The new 1080p/10 seconds mode will probably be far outside consumer hardware reach, but the improvements in the architecture at 480/720p are exciting enough already. It creates such beautiful videos and really good audio tracks. It would be a dream to see a public release, even if we have to quantize it heavily to fit all that data into our consumer GPUs. 😅
submitted by /u/pilkyton
[link] [comments]
I am putting most of my efforts to achieve more realistic Ai acting with natural…
Photonic devices are hardware systems that can process information using light instead of electricity. These…
Model weights: https://huggingface.co/Comfy-Org/Lens PR: https://github.com/Comfy-Org/ComfyUI/pull/14077 You'll need to git the merge pull request if you're…
Link: https://nju-pcalab.github.io/projects/L2P/ submitted by /u/switch2stock [link] [comments]
Keyword search breaks the moment a user types something a document doesn't literally say.
Welcome to The Blueprint, a regular feature where we highlight how Google Cloud customers are…