Generating an image of a person on a computer using text prompts is easy. Generating one with two people is similarly simple. But creating an image of multiple people actually doing something, and faithfully reproducing not just the people but the thing they’re doing? Not so easy.Generating an image of a person on a computer using text prompts is easy. Generating one with two people is similarly simple. But creating an image of multiple people actually doing something, and faithfully reproducing not just the people but the thing they’re doing? Not so easy.Computer Sciences[#item_full_content]