Skip to main content
Back to News Hub
🤖OpenAI
January 31, 2024
Funding & Investment

Building an early warning system for LLM-aided biological threat creation

Overview

OpenAI is creating an evaluation blueprint to measure how much large language models might help individuals produce biological threats. In a test involving biology experts and students, the company determined that GPT-4 offers at most a mild uplift in biological threat creation accuracy. This preliminary finding serves as a starting point for ongoing risk research and discussions across the broader community.

Key Takeaways

  • OpenAI has outlined work on a testing framework designed to act as an early warning system for biological risk in artificial intelligence.

    The effort focuses on determining whether models like GPT-4 make it easier for individuals to construct biological threats.

  • During testing, both biological domain experts and students completed tasks to see how much assistance the model provided.

    The evaluation revealed that access to GPT-4 generated "at most a mild uplift in biological threat creation accuracy."

  • Because this minor gain was not large enough to be conclusive, researchers plan to use the results as a foundation for continued research.

    Developing standardized risk benchmarks helps establish clearer safety controls for future model evaluations.

  • OpenAI is creating a blueprint to assess whether large language models can assist people in developing biological threats.

    An evaluation with biology experts and students showed that GPT-4 provides "at most a mild uplift in biological threat creation accuracy."

  • OpenAI views these initial findings as a foundation for further security research and community discussions.

OpenAI has outlined work on a testing framework designed to act as an early warning system for biological risk in artificial intelligence. The effort focuses on determining whether models like GPT-4 make it easier for individuals to construct biological threats. During testing, both biological domain experts and students completed tasks to see how much assistance the model provided.

The evaluation revealed that access to GPT-4 generated "at most a mild uplift in biological threat creation accuracy." Because this minor gain was not large enough to be conclusive, researchers plan to use the results as a foundation for continued research. Developing standardized risk benchmarks helps establish clearer safety controls for future model evaluations.

OpenAI is creating a blueprint to assess whether large language models can assist people in developing biological threats. An evaluation with biology experts and students showed that GPT-4 provides "at most a mild uplift in biological threat creation accuracy." The observed increase in threat creation accuracy was not large enough to draw definitive conclusions.

For more details please read the original article at OpenAI.

Continue Learning

Comments

Comments appear only after moderation. Your email identifies your submission to the moderator and is never displayed here.

No approved comments yet.

Originally published by OpenAI
Read the original