Sign in

OpenAI adds new safety net to prevent ChatGPT from giving advice on creating viruses, harmful chemicals

OpenAI now actively screens for biological and chemical risk with o3 and o4-mini models, and blocks model responses using a new safety monitor.

Published on: Apr 17, 2025, 13:03:23 IST
Share
Share via
  • facebook
  • twitter
  • linkedin
  • whatsapp
Copy link
  • copy link

Powerful generative artificial intelligence models have a tendency to hallucinate. They can often offer improper advice and stray off track, which can potentially misguide people. This issue has been notably discussed by industry experts, which is why the topic of guardrails has always been a focus in the AI sector. Companies like OpenAI are now actively addressing this problem, continually working to ensure that their powerful new models remain reliable. This is exactly what the company appears to be doing with its latest models, o3 and o4-mini.

OpenAI o3 and o4 have a parallel safety system that prevents these models from potentially handing out dangerous advice on biology and chemicals. (REUTERS)
OpenAI o3 and o4 have a parallel safety system that prevents these models from potentially handing out dangerous advice on biology and chemicals. (REUTERS)
Shaurya Sharma

As first spotted by TechCrunch, the company’s safety report has detailed a new system designed to monitor its AI models. This system screens any prompts submitted by users that relate to biological and chemical dangers.

“We've deployed new monitoring approaches for biological and chemical risk. These use a safety-focused reasoning monitor similar to that used in GPT-4o Image Generation and can block model responses,” OpenAI said, in its OpenAI o3 and o4-mini System Card document.

Also Read: ChatGPT now has a library to save your Ghibli and other AI-generated images

Reasoning Monitor Runs In Parallel To o3 And o4-mini

o3 and o4-ini represent significant improvements over their predecessors. With this increased capability, however, comes an expanded scope of responsibility. OpenAI’s benchmarks indicate that o3 is particularly powerful when responding to queries concerning biological threats. This is precisely where the safety-centric inference monitor plays a critical role.

The safety monitoring system runs in parallel with the o3 and o4-mini models. When a user submits prompts related to biological or chemical warfare, the system intervenes to ensure the model does not respond as per the company’s guidelines.

OpenAI also released some figures. According to their data, with the safety monitor in place, the models refrained from responding to risky prompts 98.7% of the time. “We evaluated this reasoning monitor on the output of a biorisk red-teaming campaign in which 309 unsafe conversations were flagged by red-teamers after approximately one thousand hours of red teaming,” OpenAI added.

Other Mitigations

In addition, OpenAI has implemented other mitigations to address potential risks. These include pre-training measures, such as filtering harmful training data, as well as modified post-training techniques designed to not engage with high-risk biological requests, while still permitting “benign” ones.

The system now actively monitors high-risk cybersecurity threats, including attempts to disrupt high-priority adversaries through methods such as hunting, detection, monitoring, tracking, and intelligence sharing.

Also Read: iPhone 17 Air could launch in September 2025 — Key details revealed

  • Shaurya Sharma
    ABOUT THE AUTHOR
    Shaurya Sharma

    Shaurya Sharma is the Technology Editor at Hindustan Times Digital Streams, where he oversees technology coverage across digital and social platforms. With over eight years of experience across editorial, video production, and digital media, his work focuses on smartphones, AI, consumer gadgets, and shaping audience-first content strategies for modern tech consumers. He began his career in 2018 as a fashion cinematographer before turning his lifelong passion for technology into a profession. From spending his childhood immersed in tech magazines, video games, and the latest gadgets to covering the global consumer tech industry today, technology has remained a constant throughout his journey. Over the years, Shaurya has worked with some of India’s leading media organisations, including CNN-News18, Sportskeeda, and Guiding Tech, where he led video initiatives that combined strong editorial storytelling with engaging visual and social-first execution. A graduate in Journalism and Mass Communication from Manipal University, Shaurya has reviewed hundreds of products across categories including smartphones, laptops, gaming consoles, cameras, and wearables. Beyond work, he is passionate about animal welfare, environmental causes, and automobiles, particularly turbo-petrol carsRead More

Stay updated with the latest Technology News, gadget launches, app updates, artificial intelligence and digital trends. Find reviews, comparisons and useful insights from the world of tech.