a vast herd of cute kittens
One thing I’ve noticed with image-generating algorithms is that the more of something they have to put in an image, the worse it is.
I first noticed this with the kitten-generating variant of StyleGAN, which often does okay on one cat:
but is terrible at a crowd of kittens.
A few years later, Dall-E 2 is a LOT more coherent despite having a larger job to do. But it’s still susceptible to the kitten effect.
One kitten:
Two kittens:
Ten kittens:
A vast herd of cute kittens thundering across the plain:
Similarly, there’s a huge difference in quality between a single dog and a herd of them.
And it’s not just animals, of course. Here’s the kitten effect played out in lucky charms marshmallows:
A single lucky charms marshmallow piece on a plate:
a single lucky charms marshmallow piece on a plate, labeled with the name of the shape:
A labeled list of lucky charms marshmallows:
It was not obvious to me that this should be so. If it can make one kitten, why can’t it make ten of equal quality?
My theory is this is may actually be not a numbers thing but a size thing. When I asked Dalle-2 to generate a kitten that takes up less of an image, the kitten gets way, way worse.
It’s not out of pixels to make the detail with – note the sharpness of the wood grain. But at this scale it has run out of something nonetheless.
It does make DALL-E 2’s eye test charts particularly mean.
Bonus post! In which I get DALL-E 2 to look increasingly closely at a giraffe.
ComfyUI-CacheDiT brings 1.4-1.6x speedup to DiT (Diffusion Transformer) models through intelligent residual caching, with zero…
The large language models (LLMs) hype wave shows no sign of fading anytime soon:…
This post was cowritten by Rishi Srivastava and Scott Reynolds from Clarus Care. Many healthcare…
Employee onboarding is rarely a linear process. It’s a complex web of dependencies that vary…
The latest batch of Jeffrey Epstein files shed light on the convicted sex offender’s ties…
A new light-based breakthrough could help quantum computers finally scale up. Stanford researchers created miniature…