| | Base model is definitely SOTA, can even easily compete with closed-source ones in terms of aesthetic. Cinematic quality and color grading is next level. Base model is heavily biased on Asian faces, while it excels on anime/illustration style, while my base model anime/illustration experiments wasn’t that good. Higher CFG is slightly better with anime on base. Generated with RTX6000 Blackwell Pro, Base: 29 sec 1.9it/s, 50 steps | Turbo: 2 sec, 3.9i5/s, 8 steps If you interested seeing them in original size: https://imgur.com/a/75jcjzW ComfyUI models: https://huggingface.co/Comfy-Org/ERNIE-Image/tree/main Turbo: Ernie-Image Turbo submitted by /u/sktksm |
Most current vision-language-action (VLA) models—such as OpenVLA, π0, RT-2, and RDT-1B—are “monolithic.” This means they…
Australian teams working with OpenAI models can now access the latest OpenAI models through Amazon…
AI models have clearly proven their ability to discover and exploit vulnerabilities without much, if…
The company is reducing pressure on workers to use artificial intelligence tools while encouraging them…
Self-driving cars are often controlled by deep learning models that sometimes fail in unexpected situations.…
Today, we’re excited to announce the availability of Claude Fable 5.1 on Amazon Bedrock and…