Enterprise

OpenAI expands access to DALL-E 2, its powerful image-generating AI system

Comment

OpenAI's logo
Image Credits: OpenAI

Today, DALL-E 2, OpenAI’s AI system that can generate images given a prompt or edit and refine existing images, is becoming more widely available. The company announced in a blog post that it will expedite access for customers on the waitlist with the goal of reaching roughly 1 million people within the next few weeks.

With this “beta” launch, DALL-E 2, which had been free to use, will move to a credit-based fee structure. First-time users will get a finite amount of credits that can be put toward generating or editing an image or creating a variation of an image. (Generations return four images, while edits and variations return three.) Credits will refill every month to the tune of 50 in the first month and 15 a month after that, or users can buy additional credits in increments of $15.

Here’s a chart with the specifics:

OpenAI DALL-E 2 pricing
Image Credits: OpenAI

Artists in need of financial assistance will be able to apply for subsidized access, OpenAI says.

The successor to DALL-E, DALL-E 2 was announced in April and became available for a select group of users earlier this year, recently crossing the 100,000-user threshold. OpenAI says that the broader access was made possible by new approaches to mitigate bias and toxicity in DALL-E 2’s generations, as well as evolutions in policy governing images created by the system.

OpenAI DALL-E 2
An example of the types of images DALL-E 2 can generate. Image Credits: OpenAI

For instance, OpenAI said it this week deployed a technique that encourages DALL-E 2 to generate images of people that “more accurately reflect the diversity of the world’s population” when given a prompt describing a person with an unspecified race or gender. The company also said that it’s now rejecting image uploads containing realistic faces and attempts to create the likeness of public figures, including prominent political figures and celebrities, while improving its content filters’ accuracy.

Broadly speaking, OpenAI doesn’t allow DALL-E 2 to be used to create images that aren’t “G-rated” or that could “cause harm” (e.g., images of self-harm, hateful symbols or illegal activity). And it previously disallowed the use of generated images for commercial purposes. Starting today, however, OpenAI is granting users “full usage rights” to commercialize the images they create with DALL-E 2, including the right to reprint, sell and merchandise — including images they generated during the early preview.

As demonstrated by DALL-E 2 derivatives like Craiyon (formerly DALL-E mini) and the unfiltered DALL-E 2 itself, image-generating AI can very easily pick up on the biases and toxicities embedded in the millions of images from the web used to train them. Futurism was able to prompt Craiyon to create images of burning crosses and Ku Klux Klan rallies and found that the system made racist assumptions about identities based on “ethnic-sounding” names. OpenAI researchers noted in an academic paper that an open source implementation of DALL-E could be trained to make stereotypical associations like generating images of white-passing men in business suits for terms like “CEO.”

While the OpenAI-hosted version of DALL-E 2 was trained on a dataset filtered to remove images that contained obvious violent, sexual or hateful content, filtering has its limits. Google recently said it wouldn’t release an AI-generating model it developed, Imagen, due to risks of misuse. Meanwhile, Meta has limited access to Make-A-Scene, its art-focused image-generating system, to “prominent AI artists.”

OpenAI emphasizes that the hosted DALL-E 2 incorporates other safeguards including “automated and human monitoring systems” to prevent things like the model from memorizing faces that often appear on the internet. Still, the company admits that there’s more work to do.

“Expanding access is an important part of our deploying AI systems responsibly because it allows us to learn more about real-world use and continue to iterate on our safety systems,” OpenAI wrote in a blog post. “We are continuing to research how AI systems, like DALL-E, might reflect biases in its training data and different ways we can address them.”

More TechCrunch

It ran 110 minutes, but Google managed to reference AI a whopping 121 times during Google I/O 2024 (by its own count). CEO Sundar Pichai referenced the figure to wrap…

Google mentioned ‘AI’ 120+ times during its I/O keynote

Firebase Genkit is an open source framework that enables developers to quickly build AI into new and existing applications.

Google launches Firebase Genkit, a new open source framework for building AI-powered apps

In the coming months, Google says it will open up the Gemini Nano model to more developers.

Patreon and Grammarly are already experimenting with Gemini Nano, says Google

As part of the update, Reddit also launched a dedicated AMA tab within the web post composer.

Reddit introduces new tools for ‘Ask Me Anything,’ its Q&A feature

Here are quick hits of the biggest news from the keynote as they are announced.

Google I/O 2024: Here’s everything Google just announced

LearnLM is already powering features across Google products, including in YouTube, Google’s Gemini apps, Google Search and Google Classroom.

LearnLM is Google’s new family of AI models for education

The official launch comes almost a year after YouTube began experimenting with AI-generated quizzes on its mobile app. 

Google is bringing AI-generated quizzes to academic videos on YouTube

Around 550 employees across autonomous vehicle company Motional have been laid off, according to information taken from WARN notice filings and sources at the company.  Earlier this week, TechCrunch reported…

Motional cut about 550 employees, around 40%, in recent restructuring, sources say

The keynote kicks off at 10 a.m. PT on Tuesday and will offer glimpses into the latest versions of Android, Wear OS and Android TV.

Google I/O 2024: Watch all of the AI, Android reveals

Google Play has a new discovery feature for apps, new ways to acquire users, updates to Play Points, and other enhancements to developer-facing tools.

Google Play preps a new full-screen app discovery feature and adds more developer tools

Soon, Android users will be able to drag and drop AI-generated images directly into their Gmail, Google Messages and other apps.

Gemini on Android becomes more capable and works with Gmail, Messages, YouTube and more

Veo can capture different visual and cinematic styles, including shots of landscapes and timelapses, and make edits and adjustments to already-generated footage.

Google Veo, a serious swing at AI-generated video, debuts at Google I/O 2024

In addition to the body of the emails themselves, the feature will also be able to analyze attachments, like PDFs.

Gemini comes to Gmail to summarize, draft emails, and more

The summaries are created based on Gemini’s analysis of insights from Google Maps’ community of more than 300 million contributors.

Google is bringing Gemini capabilities to Google Maps Platform

Google says that over 100,000 developers already tried the service.

Project IDX, Google’s next-gen IDE, is now in open beta

The system effectively listens for “conversation patterns commonly associated with scams” in-real time. 

Google will use Gemini to detect scams during calls

The standard Gemma models were only available in 2 billion and 7 billion parameter versions, making this quite a step up.

Google announces Gemma 2, a 27B-parameter version of its open model, launching in June

This is a great example of a company using generative AI to open its software to more users.

Google TalkBack will use Gemini to describe images for blind people

This will enable developers to use the on-device model to power their own AI features.

Google is building its Gemini Nano AI model into Chrome on the desktop

Google’s Circle to Search feature will now be able to solve more complex problems across psychics and math word problems. 

Circle to Search is now a better homework helper

People can now search using a video they upload combined with a text query to get an AI overview of the answers they need.

Google experiments with using video to search, thanks to Gemini AI

A search results page based on generative AI as its ranking mechanism will have wide-reaching consequences for online publishers.

Google will soon start using GenAI to organize some search results pages

Google has built a custom Gemini model for search to combine real-time information, Google’s ranking, long context and multimodal features.

Google is adding more AI to its search results

At its Google I/O developer conference, Google on Tuesday announced the next generation of its Tensor Processing Units (TPU) AI chips.

Google’s next-gen TPUs promise a 4.7x performance boost

Google is upgrading Gemini, its AI-powered chatbot, with features aimed at making the experience more ambient and contextually useful.

Google’s Gemini updates: How Project Astra is powering some of I/O’s big reveals

Veo can generate few-seconds-long 1080p video clips given a text prompt.

Google’s image-generating AI gets an upgrade

At Google I/O, Google announced upgrades to Gemini 1.5 Pro, including a bigger context window. .

Google’s generative AI can now analyze hours of video

The AI upgrade will make finding the right content more intuitive and less of a manual search process.

Google Photos introduces an AI search feature, Ask Photos

Apple released new data about anti-fraud measures related to its operation of the iOS App Store on Tuesday morning, trumpeting a claim that it stopped over $7 billion in “potentially…

Apple touts stopping $1.8B in App Store fraud last year in latest pitch to developers

Online travel agency Expedia is testing an AI assistant that bolsters features like search, itinerary building, trip planning, and real-time travel updates.

Expedia starts testing AI-powered features for search and travel planning