Skip to main content
  1. Home
  2. Computing
  3. News

Research shows even average users can break past AI safety within Gemini and ChatGPT

Everyday users can reveal what AI testing misses.

Add as a preferred source on Google
average-users-break-past-ai-safety-gemini-chatgpt
Aerps.com / Unsplash

What’s happened? A team at Pennsylvania State University found that you don’t need to be a hacker or prompt-engineering genius to break past AI safety; regular users can do it just as well. Test prompts in the research paper revealed clear patterns of prejudice in responses: from assuming engineers and doctors are men, to portraying women in domestic roles, and even linking Black or Muslim people with crime.

  • 52 participants were invited to craft prompts intended to trigger biased or discriminatory responses in 8 AI chatbots, including Gemini and ChatGPT.
  • They found 53 prompts that worked repeatedly on different models, showing consistent bias among them.
  • The biases exposed fell into several categories: gender, race/ethnicity/religion, age, language, disability, cultural bias, historical bias favouring Western nations, etc.

This is important because: This isn’t a story about elite jailbreakers. Average users armed with intuition and everyday language uncovered biases that slipped past AI safety tests. The study didn’t just ask trick questions; it used natural prompts like asking who was late in a doctor-nurse story or requesting a workplace harassment scenario.

  • The study reveals that AI models still carry deep social biases (like gender, race, age, disability, and cultural) that show up with simple prompts, which means bias may emerge in many unexpected ways in everyday use.
  • Notably, newer model versions weren’t always safer. Some performed worse, showing that progress in capabilities doesn’t automatically mean progress in fairness.
Recommended Videos

Why should I care? Since everyday users can trigger problematic responses in AI systems, the actual number of people who could bypass AI guardrails is much larger.

  • AI tools used in everyday chats, hiring tools, classrooms, customer support systems, and healthcare may subtly reproduce stereotypes.
  • It demonstrates that many AI-bias studies focused on complex technical attacks may miss the real-world user-triggered ones.
  • If regular prompts can unintentionally trigger bias, then bias isn’t an exception; it’s baked into how these tools think.

As generative AI becomes mainstream, improving it will require more than patches and filters; it’ll take real users stress-testing AI.

Manisha Priyadarshini
Manisha Priyadarshini is a tech and entertainment writer with over nine years of editorial experience.
AMD is apparently gearing up to raise GPU prices right after Nvidia’s steep hike
AMD has reportedly told AIB partners about a price hike that goes into effect in. August.
AMD RX 7800

AMD is next in line to raise the asking price of its Radeon GPUs, merely days after the news of a similar hike coming for Nvidia graphics cards started making waves. As per ChannelGate on Weibo (h/t VideoCardz), AMD has informed its board partners that the price of GPU and memory bundles will go up by at least 10% in August.

Is a similar price hike coming for standalone graphics cards? I won't be surprised if that happens. The situation is so bad that AMD is planning to bring back graphics cards with 4GB of onboard memory. AMD just introduced the RX 9050 GPU, which costs $279. Notably, it's just $20 less than the RX 9060 XT that offers double the graphics memory.

Read more
Microsoft will make Windows work well with just 8GB RAM. We desperately need it
The dream of a reliable and affordable Windows laptop hinges on the next milestone at Microsoft.
Surface laptop on wooden table

The year 2026 has marked a huge course correction for Microsoft after years of frustrating users with plenty of confusing processes, unoptimized UI elements, and just the generous bloatware that can bring any Windows system to a crawl. Microsoft has finally shifted into a new phase. From fixing the right-click behavior to speeding up the Search system, the company has fixed plenty of papercuts.

The next major milestone is making Windows 11 run smoothly on Windows PCs with just 8GB of RAM. That directly means entry-level and budget laptops will at least get the basics right, even if that means sacrificing the on-device AI bells and whistles. Let's face it. The increasingly AI-first approach to computing on a Windows 11 machine is a little too taxing on the hardware at hand.

Read more
ChatGPT’s Chrome extension can now read your tabs, YouTube videos, and highlighted text
ChatGPT now understands what's on your screen, letting you ask about tabs, highlighted text, and YouTube videos directly.
chatgpt-chrome-extension-upgrade-youtube-videos

ChatGPT just got a lot better at understanding what you're actually looking at. OpenAI has rolled out updates to its Chrome extension and desktop app that let ChatGPT read your open tabs, react to highlighted text, and answer questions about YouTube videos. The update arrived shortly after OpenAI shelved its standalone Atlas browser to shift its focus toward tools people already use.

https://twitter.com/ChatGPT/status/2082970812584432115

Read more