Back to dall-e
OpenAI /

DALL·E 3
DALL·E 3

Closed SourceImage GenimageUpdated October 3, 2023

Model Overview

DALL-E 3 is an advanced text-to-image generation model developed by OpenAI. It represents a significant leap forward in understanding nuanced requests and complex prompts, significantly reducing the need for traditional "prompt engineering." DALL-E 3 is natively integrated into ChatGPT, allowing users to brainstorm, refine, and generate images through natural conversation.

Capabilities

  • Nuanced Understanding: DALL-E 3 excels at adhering strictly to complex instructions, ensuring that specific details, spatial relationships, and object interactions in the prompt are accurately represented in the final image.
  • Text Rendering: A major breakthrough in DALL-E 3 is its ability to render legible and accurate text within generated images (e.g., signs, labels, logos), a historic weakness for diffusion models.
  • ChatGPT Integration: Users can simply describe what they want to ChatGPT, which will automatically expand short prompts into detailed, highly descriptive prompts for DALL-E 3 to generate the best possible result.
  • Safety & Moderation: DALL-E 3 includes robust safety mitigations to decline requests that ask for a public figure by name or that ask for images in the style of a living artist.

Official Examples

DALL-E 3 produces exceptional results when provided with descriptive, conversational prompts:

Example 1: Photography

Prompt: "Photo of a serene park setting. On the left, a golden retriever sits attentively, gazing forward with its tongue out. On the right, a tabby cat lounges lazily, stretching its legs out and looking towards the dog with a curious expression."

Example 2: Text Rendering

Prompt: "An illustration of an avocado sitting in a therapist's chair, saying 'I just feel so empty inside' with a pit-sized hole in its center. The therapist, a spoon, scribbles notes."

Example 3: Artistic & Whimsical

Prompt: "Cartoon drawing of an outer space scene. Amidst floating planets and twinkling stars, a whimsical horse with exaggerated features rides an astronaut, who swims through space with a jetpack, looking a tad overwhelmed."

Example 4: Detailed Illustration

Prompt: "A 2D animation of a folk music band composed of anthropomorphic autumn leaves, playing traditional bluegrass instruments on a rustic wooden stage, surrounded by fireflies."

Intended Use & Limitations

While highly capable, DALL-E 3 has specific limitations:

  • Complex Spatial Relationships: The model may occasionally struggle with very intricate spatial descriptions (e.g., placing one specific object behind another in a crowded scene).
  • Style Restrictions: By design, the model will refuse to generate content that mimics the specific style of living artists.
  • Content Restrictions: The model enforces strict safety filters regarding violent, adult, or hateful content.

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that artificial general intelligence benefits all of humanity.

Key Features

Native ChatGPT integration for prompt refinement

Feature 01

Dramatically improved prompt adherence

Feature 02

Text rendering in images

Feature 03

Built-in safety mitigations

Feature 04

You might also want to compare

Verified Sources

Tags

image-generationcreative

Model Specs

closed-source

Parameters

Undisclosed

Context Window

N/A (image generation)

License

Proprietary

Deployment

api-only

Cost Tiers

chatgpt-image-latest

Resources & Links

Lineage

Model Family

Part of the dall-e family

Only release in this line currently tracked.

Compare Specs

Compare parameters, context windows, modalities, and benchmark scores of this model side-by-side with others.

Compare Model