Home / Companies / Surge AI / Blog / Post Details
Content Deep Dive

DALL·E 3 and Midjourney Fail Astral Codex Ten's Image Generation Bet

Blog post from Surge AI

Post Details
Company
Date Published
Author
Edwin Chen
Word Count
2,016
Company Posts That Month
1
Language
English
Hacker News Points
-
Post removed?
No
Summary

In a 2022 bet by Astral Codex Ten (ACT) on the capabilities of generative AI models, the challenge was to create images based on five specific prompts, with success defined as at least one of ten generated images accurately depicting the scene in at least three prompts. Initially, DALL-E 2 failed, but Google’s Imagen model later claimed partial success. However, ACT's victory was contested due to inaccuracies in the generated images, such as missing elements like a key in a raven's mouth or lipstick on a fox. A recent evaluation using DALL-E 3 and Midjourney showed limited progress: DALL-E 3 succeeded in two prompts, partially succeeding in one, while Midjourney failed all. The evaluation highlighted ongoing challenges in image compositionality, suggesting that while AI image generation has advanced, it still struggles with nuanced prompt details, and the bet remains unwon.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
LLM 5 3,629 397 137 -13%
AI Guardrails 1 152 59 36 -22%
Reinforcement learning 1 No monthly metrics for this publish month.
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.