Study Reveals Large Language Model’s Limitations in Humor; ChatGPT Struggles to Be the Life of the Party

“Can machines be funny?” That’s a question that has been asked by many researchers exploring the field of computational humor. A recent study by researchers at Stanford University and Google Research set out to find the answer, taking a look at the humor-generating capabilities of ChatGPT, one of the largest language models in the world.

ChatGPT is part of the GPT family of language models that forms the backbone of the OpenAI language model. This model has learned to generate text that bears remarkable similarities to human writing. The study aimed to explore whether ChatGPT could generate funny content like a human.

The researchers tested ChatGPT’s ability to create humor by presenting it with a joke prompt and analyzing the response. They discovered that more than 90% of the time, ChatGPT’s response was a repetition of one of the 25 different jokes provided.

The top four jokes were recycled in more than half of the responses. This result points to an issue highlighted in the study: while ChatGPT can generate a large amount of text, a significant proportion of that text is repetitive and predictable.

ChatGPT’s contribution

Despite these limitations, the study suggests that ChatGPT is a significant step in the direction of creating “funny” machines. Humor has always been seen as too subjective of a field to develop an algorithm that can effectively create it. However, the study shows that despite the limitations, ChatGPT can generate coherent jokes, demonstrating its potential in various natural language tasks.

The researchers also observed that ChatGPT displayed an understanding of wordplay and double meanings. This result suggests that future iterations of the program may lead to significant advancements in this area.

Difficulty confirming beyond training data

However, the study acknowledges that it is challenging to confirm whether the jokes were hard-coded without access to more extensive language model training data. Without expanding the dataset used in the study, it is difficult to say whether ChatGPT learned these jokes through training or whether these jokes were hard-coded into the algorithm.

ChatGPT’s Limitations

While ChatGPT shows potential for advancements in computational humor, the study concludes that the model cannot confidently create intentionally funny original content. ChatGPT can respond to prompts using its vast dataset of pre-existing jokes, but it still struggles to create original humor.

In other instances, the model struggled to make sense of the joke’s setup, which led to a lack of cohesion and became an obstacle to humor. However, it is worth noting that humor often relies on cultural and contextual understandings, which may be difficult for a language model to grasp.

The study highlights the difficulty for large language models to understand and create humor. As of today, machines are no match for human humor. Yet, the study concludes that even though ChatGPT cannot yet generate intentionally funny, original content, it represents a significant leap forward in the design of machines that have a sense of humor. The potential conveyed by ChatGPT could pave the way for subsequent studies that will analyze humor in more detail, bringing us closer to the day when machines will be able to make us laugh.

Explore more

Why Is Retail the New Frontline of the Cybercrime War?

A single, unsuspecting click on a seemingly routine password reset notification recently managed to dismantle a multi-billion-dollar retail empire in a matter of hours. This spear-phishing incident did not just leak data; it triggered a sophisticated ransomware wave that paralyzed the organization’s online infrastructure for months, resulting in financial hemorrhaging exceeding $400 million. It serves as a stark reminder that

How Is Modular Automation Reshaping E-Commerce Logistics?

The relentless expansion of global shipment volumes has pushed traditional warehouse frameworks to a breaking point, leaving many retailers struggling with rigid systems that cannot adapt to modern order profiles. As consumers demand faster delivery and more sustainable practices, the logistics industry is shifting away from monolithic installations toward “Lego-like” modularity. Innovations currently debuting at LogiMAT, particularly from leaders like

Modern E-commerce Trends and the Digital Payment Revolution

The rhythmic tapping of a smartphone screen has officially replaced the metallic jingle of loose change as the primary soundtrack of global commerce as India’s Unified Payments Interface now processes a staggering seven hundred million transactions every single day. This massive migration to digital rails represents much more than a simple change in consumer habit; it signifies a total overhaul

How Do Staffing Cuts Damage the Customer Experience?

The pursuit of fiscal efficiency often leads organizations to sacrifice their most valuable asset—the human connection that transforms a simple transaction into a lasting relationship. While a leaner payroll might appear advantageous on a quarterly earnings report, the structural damage inflicted on the brand often outweighs the short-term financial gains. When the individuals responsible for the customer journey are stretched

How Can AI Solve the Relevance Problem in Media and Entertainment?

The modern viewer often spends more time navigating through rows of colorful thumbnails than actually watching a film, turning what should be a moment of relaxation into a chore of digital indecision. In a world where premium content is virtually infinite, the psychological weight of choice paralysis has become a silent tax on the consumer experience. When a platform offers