A new study by Gary Marcus, Ernest Davis, and Scott Aaronson puts DALL-E 2 to the test with difficult prompts to evaluate its common sense and reasoning abilities. The results show that while the AI sometimes succeeds, it struggles to consistently produce accurate images for complex text inputs.