Skip to main content
  1. Home
  2. Computing
  3. News

Study finds AI-generated animal stories rarely have a female lead

Researchers ran six major AI models through 23,800 story completions and found female leads in barely 2 percent of the results.

Add as a preferred source on Google
A kid reading a book.
Catherine Hammond / Unsplash

If you use an AI chatbot to write your kids a bedtime story about talking animals, don’t count on it having a female lead. A recent study has found that major AI models mostly skip gender for animal characters, but when they don’t, the protagonist is almost always male.

Female leads made up just 2 percent of the results

Melanie Walsh, an assistant professor at the University of Washington’s Information School, led the research behind these findings. Her team ran the same prompt through six widely used AI models, including GPT-4o, GPT-5.1, Gemini 2.5, Claude Sonnet 4.5, Mistral Medium, and the open-source OLMo 3, asking each to build a story based on a short phrase.

Recommended Videos

Out of 23,800 completions, the researchers found only 513 with a female lead, roughly one in every 46 stories. Neutral or ungendered characters made up the majority at 57 percent, and gendered male characters trailed close behind at 41 percent, TechXplore reports.

Walsh said her team can “only poke at them from the outside,” since the models are largely proprietary. Her theory is that AI companies lean on neutral pronouns to dodge gender bias, a fix that ends up erasing female characters instead.

GPT-5.1 and Gemini 2.5 defaulted to male most often

GPT-5.1 and Gemini 2.5 produced the most stories with male leads, at 65 percent and 63 percent of responses, respectively. OLMo 3 leaned neutral most, at 85 percent, while Claude Sonnet 4.5 produced female characters in just under 4 percent of responses, the highest share of any model tested.

Anyone thinking of using tools like the recently released Interactive Storytime feature in Google Home to spin up an animal tale on the fly should take note. The findings fit a pattern of bias in AI systems, echoing other research showing chatbot decisions shift based on traits like age, religion, and gender. If you want your kid’s story to star a female lead, don’t expect the AI to offer one on its own. You’ll probably have to ask.

Pranob Mehrotra
Pranob is a seasoned tech journalist with over eight years of experience covering consumer technology. His work has been…
Three hikers trusted Gemini to plan a hike and had to be rescued the next day
Rescuers say Gemini advised the group to carry far less food and water than they needed.
Mount Shasta, United States

AI chatbots can help plan a vacation or suggest what to pack. But there are situations where advice from someone who actually knows what they are doing is the safer choice.

As reported by the Chicago Tribune, three novice hikers from Roseville, California, had to be rescued from Mount Shasta after becoming stranded overnight. According to the Siskiyou County Sheriff’s Office, they had used Google Gemini to help plan the climb. The sheriff’s office said Gemini advised the group to carry far less food and water than they needed.

Read more
Windows 11 is about to turn on a setting that could hurt gaming performance
Microsoft will start enabling Memory Integrity on more Windows 11 PCs in October
A gaming PC with RGB synced lights running Apex Legends.

Microsoft is preparing to enable Memory Integrity on more Windows 11 PCs starting in October. The security feature is meant to protect systems from malicious code, but there is one reason gamers may want to keep an eye on it.

Microsoft has previously acknowledged that Memory Integrity can affect gaming performance on some Windows 11 systems. Back in 2022, the company even published instructions explaining how gamers could temporarily disable Memory Integrity and Virtual Machine Platform if they were causing problems.

Read more
AI chatbots will agree and misinform if you just put a little pressure, warns research
ChatGPT 3.5 proved the most vulnerable when false claims were repeated over and over
Claude AI on an iPhone.

AI chatbots have a well-documented habit of hallucinating information and sometimes agreeing with users even when they are wrong. A new study suggests that simply refusing to take no for an answer can make the problem worse.

Researchers from the University of Arizona tested seven AI models, including GPT-3.5, GPT-4o, GPT-4o-mini, Claude 3.5 Sonnet, Gemini 1.5 Pro, Llama 3 70B, and DeepSeek-R1. Instead of judging them from a single response, researchers kept conversations going while repeatedly feeding the models information they knew was false.

Read more