AI

Experiment finds AI boosts creativity individually — but lowers it collectively

Comment

Illustration of a robot helping a human to write.
Image Credits: Bryce Durbin / TechCrunch

A new study examines whether AI could be an automated helpmeet in creative tasks, with mixed results: It appeared to help less naturally creative people write more original short stories — but dampened the creativity of the group as a whole. It’s a trade-off that may be increasingly common as AI tools impinge on creative endeavors.

The study is from researchers Anil Doshi and Oliver Hauser at University College London and University of Exeter, respectively, and was published in Science Advances. And while it’s necessarily limited due to its focus on short stories, it seems to confirm the feeling many have expressed: that AI can be helpful but ultimately offers nothing truly new in creative endeavors.

“Our study represents an early view on a very big question on how large language models and generative AI more generally will affect human activities, including creativity,” Hauser told TechCrunch in an email. “While there is huge potential (and, no doubt, huge hype) for this technology to have big impacts in media and creativity more generally, it will be important that AI is actually being evaluated rigorously — rather than just implemented widely, under the assumption that it will have positive outcomes.”

The experiment had hundreds of people write very short stories (eight sentences or so), on any topic but suitable for a broad audience. One group just wrote; a second group was given the opportunity to consult GPT-4 for a single story idea with a few sentences (they could use as much or as little as they liked); a third could get up to five such story starters.

Image Credits: Hauser, Joshi

Once the stories were written, they were evaluated by both their own writers and a second group that knew nothing about the generative AI twist. These people rated the stories on novelty, usefulness (i.e. likelihood of publishing) and emotional enjoyment.

Low creativity, high benefit…high creativity, no benefit

Prior to writing the stories, the participants also completed a word-production task that acts as a proxy for creativity. It’s a concept that can’t be directly measured, but in this case one’s creativity in writing can at least be approximated (without judgment!; not everyone is a born or practiced writer).

“Capturing something so rich and complex as creativity with any measure seems fraught with complications,” wrote Hauser. “There is, however, a rich set of research around human creativity and there is a live debate about how best to capture the idea of creativity in a measure.”

They said their approach was widely used in academia and well documented in other studies.

What the researchers found was that people with lower creativity metrics scored lowest on evaluations of their stories, which arguably validates the approach. They also saw the largest gains when given the opportunity to use a generated story idea (which, it’s worth noting, the vast majority across the experiment did).

Stories by people with a low creativity score who just wrote were reliably rated lower than others on writing quality, enjoyability and novelty. Given one AI-generated idea, they scored higher on every metric. Given the choice of five, they scored even higher.

It really appears that for folks struggling with the creative side of writing (at least within this context and definition), the AI helper is genuinely improving the quality of their work. This probably resonates with many to whom writing does not come naturally, and a language model saying “hey, try this” is the prompt they need to finish a paragraph or start a new chapter.

Image Credits: Hauser, Joshi

But what about the people who scored highly on the creativity metric? Did their writing climb to new heights? Sadly, no. In fact, those participants saw little to no benefit at all, or even (though it’s very close and arguably not significant) worse ratings. It seems that those on the creative side produced their best work when they had no AI help at all.

One can imagine any number of reasons why this might be the case, but the numbers do suggest that, in this situation, AI had a zero to negative effect on writers with innate creativity.

Flattened

But that’s not the part that the researchers were worried about.

Beyond the subjective evaluation of stories by participants, the researchers conducted some analyses of their own. They used OpenAI’s embeddings API to rate how similar each story was to the other stories in its category (i.e. human-only, one AI option, or five AI options).

They found that access to generative AI caused the resulting stories to be closer to the average for their category. In other words, they were more similar and less varied as a group. The total difference was in the 9% to 10% range, so it’s not like the stories were all clones of one another. And who knows, but this similarity might be an artifact of less practiced writers finishing a suggested story versus more creative writers coming up with one from scratch.

The finding was nevertheless enough to warrant a cautionary note in the conclusions, which I could not condense and so quote in full:

While these results point to an increase in individual creativity, there is risk of losing collective novelty. In general equilibrium, an interesting question is whether the stories enhanced and inspired by AI will be able to create sufficient variation in the outputs they lead to. Specifically, if the publishing (and self-publishing) industry were to embrace more generative AI-inspired stories, our findings suggest that the produced stories would become less unique in aggregate and more similar to each other. This downward spiral shows parallels to an emerging social dilemma: If individual writers find out that their generative AI-inspired writing is evaluated as more creative, they have an incentive to use generative AI more in the future, but by doing so, the collective novelty of stories may be reduced further. In short, our results suggest that despite the enhancement effect that generative AI had on individual creativity, there may be a cautionary note if generative AI were adopted more widely for creative tasks.

It echoes the fear in visual art and in web content that if the AI leads to more AI, and what it trains on is just more of itself, it could end up in a self-perpetuating cycle of blandness. As generative AI begins to creep into every medium, it is studies like these that act as counterweights to claims of unbounded creativity or new eras of AI-generated films and songs.

Hauser and Doshi acknowledge that their work is just the beginning — the field is brand new, and every study, including their own, is limited.

“There are a number of paths that we expect future research to pick up on. For instance, implementation of generative AI ‘in the wild’ will look very different than our controlled setting,” Hauser wrote. “Ideally, our study helps guide both the technology and how we interact with it to ensure continued diversity of creative ideas, whether it is in writing, or art, or music.”

More TechCrunch

Tags

Sources with knowledge of the deal said it was an all-stock transaction that valued Good Eggs slightly above its previous valuation of $22 million.

GrubMarket has acquired Good Eggs

Quantum computing may still largely be in the theoretical domain, but the money that it’s attracting is very real.

UK’s Riverlane scores $75M to correct quantum errors

Indian space-tech startup EtherealX aims to test its fully reusable medium-lift vehicle in 20.

India’s EtherealX puts $5M seed toward fully reusable launch vehicles

John Schulman, one of the co-founders of OpenAI, has left the company for rival AI startup Anthropic. In addition, OpenAI president and co-founder Greg Brockman is taking an extended leave…

OpenAI co-founder Schulman leaves for Anthropic, Brockman takes extended leave

Grok — not to be confused with the homophonic AI startup Groq that this morning raised over $600 million — has been spreading false information about Vice President Kamala Harris…

Secretaries of state urge X to stop its Grok chatbot from spreading election misinformation

Saudi Arabia is committing even more money to Lucid Motors as the EV startup struggles to erase its losses. Lucid announced Monday as part of its second-quarter earnings report that…

Lucid pumps $1.5B from Saudi wealth fund after CEO warned relying on its ‘bottomless wealth’ was ‘dangerous’

Listen, I’m tired of talking about Boeing’s Starliner, too, but the spacecraft still isn’t home and questions are mounting about NASA’s transparency.

TechCrunch Space: I’m tired of talking about Starliner, too

Google will appeal a U.S. District Court judge’s opinion Monday that found the technology giant acted illegally to maintain a monopoly in online search. The decision from Judge Amit P.…

Google loses massive antitrust case over search, will appeal ruling

Last year, OpenAI held a splashy press event in San Francisco during which the company announced a bevy of new products and tools, including the ill-fated App Store-like GPT Store.…

OpenAI tempers expectations with less bombastic, GPT-5-less DevDay this fall

Muon Space closed a new tranche of funding for its space-as-a-service business.

Muon Space closes $56M to scale all-in-one satellite platform

We’re so excited to announce that we’ve added a dedicated AI Stage presented by Google Cloud to TechCrunch Disrupt 2024. It joins Fintech, SaaS and Space as the other industry-focused…

Announcing the agenda for the AI Stage at TechCrunch Disrupt 2024

A startup developing AI market research based on location data, and backed by a who’s who, has quietly raised, TechCrunch has learned.

Placer.ai boosts valuation to $1.5B after quietly raising another $75M

Safari’s newest feature, Distraction Control, can remove distracting elements from a website. The feature follows Arc Browser’s addition of Boosts last year, which similarly lets users remove features from a…

Apple’s new Safari feature removes distracting items from websites

By collecting this data, OpenAI “profited significantly” from the creators’ work, the complaint alleges.

YouTuber files class action suit over OpenAI’s scrape of creators’ transcripts

India’s fast-growing quick commerce market is getting a new deep-pocketed entrant: Walmart-owned Flipkart, India’s largest e-commerce firm. Flipkart has started to roll out Flipkart Minutes, its quick commerce service, in…

Flipkart blitzes into India’s 10-minute quick commerce battle

The list includes Elon Musk’s xAI, which is already valued at a staggering $24 billion, as well as a good number of other AI startups.

38 startups have become unicorns so far in 2024: Here’s the full list

When a company is the size of Amazon, a lot of bad actors will come after it and its customers, which makes defending the network a monster job. Over the…

AWS unveils Mithra to identify and mitigate malicious domains across its massive system

The European Commission has closed a Digital Services Act (DSA) investigation of a rewards feature in TikTok Lite by accepting commitments from the social media giant to permanently withdraw the…

TikTok Lite: EU closes addictive design case after TikTok commits to not bring back rewards mechanism

Groq, a startup developing chips to run generative AI models faster than conventional processors, said on Monday that it has raised $640 million in a new funding round led by…

AI chip startup Groq lands $640M to challenge Nvidia

COVID-19 pushed people to take up outdoor activities. Now, startups are helping companies and consumers keep up with demand.

From golf to hunting, a new crop of startups want to make these experiences even better

Despite increasing demand for AI safety and accountability, today’s tests and benchmarks may fall short, according to a new report. Generative AI models — models that can analyze and output…

Many safety evaluations for AI models have significant limitations

OpenAI has built a tool that could potentially catch students who cheat by asking ChatGPT to write their assignments — but according to The Wall Street Journal, the company is…

OpenAI says it’s taking a ‘deliberate approach’ to releasing tools that can detect writing from ChatGPT

Chief Product Officer Craig Saldanha says AI is already transforming the Yelp experience.

Yelp’s chief product officer talks AI and authenticity

Featured Article

Even after $1.6B in VC money, the lab-grown meat industry is facing ‘massive’ issues

Any goal that puts cultivated meat in big box grocery stores or on fast food menus in the 2020s is “unrealistic,” according to experts.

Even after $1.6B in VC money, the lab-grown meat industry is facing ‘massive’ issues

Warren Buffett’s Berkshire Hathaway cut its Apple holding by around half, to $84.2 billion, according to an SEC filing. While Apple remains the firm’s largest stock holding by far, Buffett…

Warren Buffett’s Berkshire Hathaway sells half its Apple stock

A fireside chat between Jensen Huang and Mark Zuckerberg at SIGGRAPH 2024 took some unexpected turns. What started as a conversation about the capabilities of Nvidia GPUs and Zuckerberg’s vision…

Zuckerberg and Jensen show off their friendship, while an AI necklace covets yours

We spoke to Harness CEO and founder Jyoti Bansal about his previous company, which Cisco bought for $3.7 billion in 2017.

When a big company comes after a hot startup, it’s not a slam dunk decision to sell

Dojo is Tesla’s custom-built supercomputer that’s designed to train its “Full Self-Driving” neural networks.

Tesla Dojo: Elon Musk’s big plan to build an AI supercomputer, explained

Featured Article

Trade My Spin is building a business around used Peloton equipment

Trade My Spin has pieced together a logistics network capable of offering same or next day delivery in most major cities in the continental U.S.

Trade My Spin is building a business around used Peloton equipment

Featured Article

Meet the founder who built and sold a $600M enterprise software startup from Sri Lanka

Sanjiva Weerawarana co-founded WSO2 in 2005, recently selling it for more than $600M. He sometimes drives for Uber, too.

Meet the founder who built and sold a $600M enterprise software startup from Sri Lanka