OpenAI has published details regarding its approach to AI safety, stating that ensuring AI systems are built, deployed, and used safely is central to its mission. The organization conducts rigorous testing and engages external experts for feedback before releasing new systems. It also works to improve model behavior using techniques like reinforcement learning with human feedback and builds broad safety and monitoring systems.
Key Points
- OpenAI spent more than 6 months working to make GPT-4 safer and more aligned after its training was complete, prior to public release.
- GPT-4 is 82% less likely to respond to requests for disallowed content compared to GPT-3.5.
- GPT-4 is 40% more likely to produce factual content than GPT-3.5.
- OpenAI requires users to be 18 or older, or 13 or older with parental approval, to use its AI tools.
- The organization does not permit its technology to be used to generate hateful, harassing, violent, or adult content.
- OpenAI uses Thorn’s Safer to detect, review, and report known Child Sexual Abuse Material uploaded to its image tools to the National Center for Missing and Exploited Children.
Context
According to OpenAI, while its AI tools offer benefits like increased productivity and enhanced creativity, they also carry risks. The organization aims to integrate safety at all levels of its systems. OpenAI believes that powerful AI systems require rigorous safety evaluations and advocates for regulation, actively engaging with governments on its form.
Why It Matters
This document outlines OpenAI's commitment to responsible AI development and deployment, providing insights into the measures taken to mitigate risks. Builders and users can understand the safety frameworks and limitations in place for models like GPT-4, influencing how they integrate and interact with these technologies.
What To Do
- Note the stated age requirements for using OpenAI's AI tools.
- Review the categories of content that OpenAI does not permit its technology to generate.
- Observe the improvements in safety and factual accuracy cited for GPT-4 compared to GPT-3.5.
- Watch for further details on features that will allow developers to set stricter standards for model outputs.
Keep Exploring
/atlas/gpt-family
