In 2015 a nineteen-year-old taught a machine to draw from sentences — 2,709 blurry 32-pixel squares, the first images ever pulled out of text, minted on Ethereum eight years later.
Alfred, from the DAILY × Fellowship exhibition notes and the DALL-E paper, August 2026Modern machine learning approaches to text to image synthesis started with the work of Mansimov et al. (2015)
— Ilya Sutskever, Mark Chen, Alec Radford et al., 'Zero-Shot Text-to-Image Generation' (DALL-E paper), OpenAI, 2021
By early 2015, neural networks had mastered the art of 'image-to-text' and could create natural language captions for images. Flipping this process, and turning text into image, was a much more complex challenge solved by 19-year old prodigy Elman Mansimov's alignDRAW model. The small 32 by 32 pixel outputs are significant as a proof of concept for a new form of communication between machines and humans, one that involves our own language instead of code.
DAILY × Fellowship, exhibition notes, daily.xyz/exhibition/10002, verified 18 August 2026A 19-year-old researcher at the University of Toronto built the proof of concept for text-to-image AI in 2015. The 32×32-pixel outputs — blurry and small — were the first images a machine ever generated from a sentence. They were minted as 2,709 NFTs eight years later.
Alfred watches the floor, sales and new listings for alignDRAW — and tells you when something moves.
Track Elman Mansimov in Alfred