Skip to main content
Ryan Orban

Ryan Orban

Subject
3 entries

Text to Image

Bookmarks

  1. An Image is Worth One Word: Personalizing Text-to-Image Generation Using Textual Inversion

    Tel Aviv University and NVIDIA paper introducing Textual Inversion — learning a single new text embedding token that represents a user-provided concept, enabling that concept to be composed into any text prompt. Showed that the embedding space of text-to-image models is richly structured and can be expanded with just 3-5 example images.

  2. Craiyon (formerly DALL-E mini)

    Craiyon, formerly DALL-E mini, is a free web-based text-to-image generator that went viral before Stable Diffusion's release. It gave millions their first hands-on experience with AI image generation, despite producing lower-quality images than commercial alternatives.

  3. Deep Daze: Text to Image with CLIP and Siren

    Deep Daze is Phil Wang's early text-to-image tool combining OpenAI's CLIP with Siren (implicit neural representations) — one of the first accessible open-source implementations of text-guided image generation, predating DALL-E and Stable Diffusion by over a year.

All bookmarks