Enterprise

Businesses including Stitch Fix are already experimenting with DALL-E 2

Comment

OpenAI's logo
Image Credits: OpenAI

It’s been just a few weeks since OpenAI began allowing customers to commercially use images created by DALL-E 2, its remarkably powerful AI text-to-image system. But in spite of the current technical limitations and lack of volume licensing, not to mention API, some pioneers say they’re already testing the system for various business use cases — awaiting the day when DALL-E 2 becomes stable enough to deploy into production.

Stitch Fix, the online service that uses recommendation algorithms to personalize apparel, says it has experimented with DALL-E 2 to visualize its products based on specific characteristics like color, fabric and style. For example, if a Stitch Fix customer asked for a “high-rise, red, stretchy, skinny jean” during the pilot, DALL-E 2 was tapped to generate images of that item, which a stylist could use to match with a similar product in Stitch Fix’s inventory.

“DALL-E 2 helps us surface the most informative characteristics of a product in a visual way, ultimately helping stylists find the perfect item that matches what a client has requested in their written feedback,” a spokesperson told TechCrunch via email.

Stitch Fix DALL-E 2
A DALL-E 2 generation from Stitch Fix’s pilot. The prompt was: “soft, olive green, great color, pockets, patterned, cute texture, long, cardigan.” Image Credits: OpenAI

Of course, DALL-E 2 has quirks — some of which are giving early corporate users pause. Eric Silberstein, the VP of data science at e-commerce startup Klaviyo, outlines in a blog post his mixed impressions of the system as a potential marketing tool.

He notes that facial expressions on human models generated by DALL-E 2 tend to be inappropriate and muscles and joints disproportionate, and that the system doesn’t always perfectly understand instructions. When Silberstein asked DALL-E 2 to create an image of a candle on a wooden table against a gray background, DALL-E 2 sometimes erased the candle’s lid and blended it into the desk or added an incongruous rim around the candle.

DALL-E 2 Eric Silberstein
Silberstein’s experiments with DALL-E 2 for product visualization. Image Credits: OpenAI

“For photos with humans and photos of humans modeling products, it could not be used as is,” Silberstein wrote. Still, he said he’d consider using DALL-E 2 for tasks like giving starting points for edits and conveying ideas to graphic artists. “For stock photos without humans and illustrations without specific branding guidelines, DALL·E 2, to my non-expert eye, could reasonably replace the ‘old way’ right now,” Silberstein continued.

Editors at Cosmopolitan came to a similar conclusion when they teamed up with digital artist Karen X. Cheng to create a cover for the magazine using DALL-E 2. Arriving at the final cover took very specific prompting from Cheng, which the editors said is illustrative of DALL-E 2’s limitation as an art generator.

But the AI weirdness works sometimes — as a feature, rather than a bug. For its Draw Ketchup campaign, Heinz had DALL-E 2 generate a series of images of ketchup bottles using natural language terms like “ketchup,” “ketchup art,” “fuzzy ketchup,” “ketchup in space” and “ketchup renaissance.” The company invited fans to send their own prompts, which Heinz curated and shared across its social channels.

Heinz DALL-E 2
Heinz bottles as “imagined” by DALL-E 2, a part of Heinz’ recent ad campaign. Image Credits: OpenAI

“With AI imagery dominating news and social feeds, we saw a natural opportunity to extend our ‘Draw Ketchup’ campaign; rooted in the insight that Heinz is synonymous with the word ketchup — to test this theory in the AI space,” Jacqueline Chao, senior brand manager for Heinz, said in a press release.

Clearly, DALL-E 2-driven campaigns can work when AI is the subject. But several DALL-E 2 business users say they’ve wielded the system to generate assets that don’t bear the telltale signs of AI constraints.

Jacob Martin, a software engineer, used DALL-E 2 to create a logo for OctoSQL, an open source project he’s developing. For around $30 — roughly the cost of logo design services on Fiverr — Martin ended up with a cartoon image of an octopus that looks human-illustrated to the naked eye.

“The end result isn’t ideal, but I’m very happy with it,” Martin wrote in a blog post. “As far as DALL-E 2 goes, I think right now it’s still very much in a “’first iteration’ phase for most bits and purposes — the main exception being pencil sketches; those are mind-blowingly good … I think the real breakthrough will come when DALL-E 2 gets 10x-100x cheaper and faster.”

DALL-E 2 OctoSQL
The OctoSQL logo, generated after several attempts with DALL-E 2. Image Credits: OpenAI

One DALL-E 2 user — Don McKenzie, the head of design at dev startup Deephaven — took the idea a step further. He tested applying the system to generate thumbnails on the company’s blog, motivated by the idea that posts with images get much more engagement than those without.

“As a small team of mostly engineers, we don’t have the time or budget to commission custom artwork for every one of our blog posts,” McKenzie wrote in a blog post. “Our approach so far has been to spend 10 minutes scrolling through tangentially related but ultimately ill-fitting images from stock photo sites, download something not terrible, slap it in the front matter and hit publish.”

After spending a weekend and $45 in credits, McKenzie says he was able to replace 100 or so blog posts with DALL-E 2-generated images. It took finagling with the prompts to get the best results, but McKenzie says it was well worth the effort.

“On average, I would say it took a couple of minutes and about four to five prompts per blog post to get something I was happy with,” he wrote. “We were spending more on money and time on stock images a month, with a worse result.”

For companies without the time to spend on brainstorming prompts, there’s already a startup trying to commercialize DALL-E 2’s asset-generating capabilities. Unstock.ai, built on top of DALL-E 2, promises “high-quality images and illustrations on demand” — for no charge, at the moment. Customers enter a prompt (e.g., “Top view of three goldfish in a bowl”) and then choose a preferred style (vector art, photorealistic, penciled) to create images, which can be cropped and resized.

Unstock.ai essentially automates prompt engineering, a concept in AI that looks to embed a task description in text. The idea is to provide an AI system detailed instructions so that it reliably accomplishes the thing being asked of it; in general, the results for a prompt like “Film still of a woman drinking coffee, walking to work, telephoto” will be much more consistent than “A woman walking.”

It’s likely a harbinger of applications to come. When contacted for comment, OpenAI declined to share numbers around DALL-E 2’s business users. But anecdotally, the demand appears to be there. Unofficial workarounds to DALL-E 2’s lack of API have sprung up across the web, strung together by devs eager to build the system into apps, services, websites and even video games.

More TechCrunch

Apple released new data about anti-fraud measures related to its operation of the iOS App Store on Tuesday morning, trumpeting a claim that it stopped over $7 billion in “potentially…

Apple touts stopping $1.8BN in App Store fraud last year in latest pitch to developers

Online travel agency Expedia is testing an AI assistant that bolsters features like search, itinerary building, trip planning, and real-time travel updates.

Expedia starts testing AI-powered features for search and travel planning

Welcome to TechCrunch Fintech! This week, we look at the drama around TabaPay deciding to not buy Synapse’s assets, as well as stocks dropping for a couple of fintechs, Monzo raising…

Inside TabaPay’s drama-filled decision to abandon its plans to buy Synapse’s assets

The person who claimed to have stolen the physical addresses of 49 million Dell customers appears to have taken more data from a different Dell portal, TechCrunch has learned. The…

Threat actor scraped Dell support tickets, including customer phone numbers

If you write the words “cis” or “cisgender” on X, you might be served this full-screen message: “This post contains language that may be considered a slur by X and…

On Elon’s whim, X now treats ‘cisgender’ as a slur

Facebook once had big ambitions to be a major player in enterprise communication and productivity, but today the social network’s parent company Meta will be closing a very significant chapter…

Meta is shutting down Workplace, its enterprise communications business

The Oversight Board has overturned Meta’s decision to take down a documentary revealing the identities of child abuse victims in Pakistan.

Meta’s Oversight Board overturns takedown decision for Pakistan child abuse documentary

The keynote kicks off at 10 a.m. PT on Tuesday and will offer glimpses into the latest versions of Android, Wear OS and Android TV.

Google I/O 2024: How to watch

Adam Selipsky is stepping down from his role as CEO of Amazon Web Services, Amazon has confirmed to TechCrunch.  In a memo shared internally by Amazon CEO Andy Jassy and…

AWS CEO Adam Selipsky steps down

VC and podcaster David Sacks has revealed a new AI chat app called Glue that fixes “Slack channel fatigue,” he says.

David Sacks reveals Glue, the AI company he’s been teasing on his All In podcast

Harness isn’t founder Jyoti Bansal’s first startup. He sold AppDynamics to Cisco for $3.7 billion in 2017, the week it was supposed to go public. His latest venture has raised…

After surpassing $100M in ARR, Harness grabs a $150M line of credit

You can expect plenty of AI, but probably not a lot of hardware.

Google I/O 2024: What to expect

The company’s autonomous vehicles have had a number of misadventures lately, involving driving into construction sites.

Waymo’s robotaxis under investigation after crashes and traffic mishaps

The company is describing the event as “a chance to demo some ChatGPT and GPT-4 updates.”

OpenAI’s ChatGPT announcement: Watch the GPT-4o reveal and demo here

Sona, a workforce management platform for frontline employees, has raised $27.5 million in a Series A round of funding. More than two-thirds of the U.S. workforce are reportedly in frontline…

Sona, a frontline workforce management platform, raises $27.5M with eyes on US expansion

Uber Technologies announced Tuesday that it will buy the Taiwan unit of Delivery Hero’s Foodpanda for $950 million in cash. The deal is part of Uber Eats’ strategy to expand…

Uber to acquire Foodpanda’s Taiwan unit from Delivery Hero for $950M in cash 

Paris-based Blisce has become the latest VC firm to launch a fund dedicated to climate tech. It plans to raise as much as €150M (about $162M).

Paris-based VC firm Blisce launches climate tech fund with a target of $160M

Maad, a B2B e-commerce startup based in Senegal, has secured $3.2 million debt-equity funding to bolster its growth in the western Africa country and to explore fresh opportunities in the…

Maad raises $3.2M seed amid B2B e-commerce sector turbulence in Africa

The fresh funds were raised from two investors who transferred the capital into a special purpose vehicle, a legal entity associated with the OpenAI Startup Fund.

OpenAI Startup Fund raises additional $5M

Accel has invested in more than 200 startups in the region to date, making it one of the more prolific VCs in this market.

Accel has a fresh $650M to back European early-stage startups

Kyle Vogt, the former founder and CEO of self-driving car company Cruise, has a new VC-backed robotics startup focused on household chores. Vogt announced Monday that the new startup, called…

Cruise founder Kyle Vogt is back with a robot startup

When Keith Rabois announced he was leaving Founders Fund to return to Khosla Ventures in January, it came as a shock to many in the venture capital ecosystem — and…

From Miles Grimshaw to Eva Ho, venture capitalists continue to play musical chairs

On the heels of OpenAI announcing the latest iteration of its GPT large language model, its biggest rival in generative AI in the U.S. announced an expansion of its own.…

Anthropic is expanding to Europe and raising more money

If you’re looking for a Starliner mission recap, you’ll have to wait a little longer, because the mission has officially been delayed.

TechCrunch Space: You rock(et) my world, moms

Apple devoted a full event to iPad last Tuesday, roughly a month out from WWDC. From the invite artwork to the polarizing ad spot, Apple was clear — the event…

Apple iPad Pro M4 vs. iPad Air M2: Reviewing which is right for most

Terri Burns, a former partner at GV, is venturing into a new chapter of her career by launching her own venture firm called Type Capital. 

GV’s youngest partner has launched her own firm

The decision to go monochrome was probably a smart one, considering the candy-colored alternatives that seem to want to dazzle and comfort you.

ChatGPT’s new face is a black hole

Apple and Google announced on Monday that iPhone and Android users will start seeing alerts when it’s possible that an unknown Bluetooth device is being used to track them. The…

Apple and Google agree on standard to alert people when unknown Bluetooth devices may be tracking them

A human safety operator will be behind the wheel during this phase of testing, according to the company.

GM’s Cruise ramps up robotaxi testing in Phoenix

OpenAI announced a new flagship generative AI model on Monday that they call GPT-4o — the “o” stands for “omni,” referring to the model’s ability to handle text, speech, and…

OpenAI debuts GPT-4o ‘omni’ model now powering ChatGPT