Skip to content
Longterm Wiki
Back

Credibility Rating

4/5
High(4)

High quality. Established institution or organization with editorial oversight and accountability.

Rating inherited from publication venue: OpenAI

The GPT-4 System Card is a key industry document illustrating how a leading AI lab approaches pre-deployment safety evaluation for a frontier model; useful for understanding practical risk assessment and mitigation processes.

Metadata

Importance: 72/100organizational reportprimary source

Summary

OpenAI's system card for GPT-4 documents safety evaluations, risk assessments, and mitigations conducted prior to deployment. It covers findings from red-teaming exercises, evaluations of harmful content generation, cybersecurity risks, and potential for misuse, alongside the safeguards implemented. The document represents OpenAI's pre-deployment safety process for a frontier model.

Key Points

  • Documents GPT-4's potential risks including harmful content, disinformation, cybersecurity threats, and proliferation of dangerous knowledge.
  • Describes red-teaming efforts with external domain experts to identify safety gaps before public release.
  • Outlines mitigations including RLHF fine-tuning and rule-based reward models (RBRMs) to reduce harmful outputs.
  • Evaluates GPT-4 on uplift potential for chemical, biological, radiological, and nuclear (CBRN) threats.
  • Acknowledges residual risks and limitations of current safety measures, including jailbreaks and prompt injection.

Cited by 2 pages

PageTypeQuality
Alignment Research Center (ARC)Organization57.0
Red TeamingResearch Area65.0

Cached Content Preview

HTTP 200Fetched Apr 9, 20266 KB
-->
 
 
 
 
 
 
 
 
 
 

 
 
 
 
 
 

 
 
 

 
 
 

 

 

 
Ask the publishers to restore access to 500,000+ books.

 
 
 
 
 
 

 

 

 
 

 

 

 
 
 
 
 Hamburger icon
 An icon used to represent a menu that can be
 toggled by interacting with this icon.
 

 
 
 

 

 
 
 Internet Archive logo
 A line drawing of the Internet Archive headquarters
 building façade.
 
 
 

 
 
 
 

 

 
 
 
 
 
 
 
 

 

 

 

 

 

 

 
 

 

 

 

 

 

 

 

 
 

 
 
 

 
 

 

 
 

 
 
 
 
 
 
 
 Web icon
 An illustration of a computer
 application window
 

 

 
 Wayback Machine

 
 
 
 
 
 
 
 
 
 Texts icon
 An illustration of an open book.
 
 

 

 

 
 Texts

 
 
 
 
 
 
 
 
 
 Video icon
 An illustration of two cells of a film
 strip.
 

 

 
 Video

 
 
 
 
 
 
 
 
 
 Audio icon
 An illustration of an audio speaker.
 
 
 
 

 
 
 
 

 
 Audio

 
 
 
 
 
 
 
 
 
 Software icon
 An illustration of a 3.5" floppy
 disk.
 

 

 

 
 Software

 
 
 
 
 
 
 
 
 
 Images icon
 An illustration of two photographs.
 
 

 

 
 Images

 
 
 
 
 
 
 
 
 
 Donate icon
 An illustration of a heart shape
 
 

 

 

 
 Donate

 
 
 
 
 
 
 
 
 
 Ellipses icon
 An illustration of text ellipses.
 
 

 

 
 More

 
 
 
 

 
 

 

 
 

 
 
 
 
 Donate icon
 An illustration of a heart shape
 

 

 

 "Donate to the archive"
 

 
 

 
 
 

 
 
 
 User icon
 An illustration of a person's head and chest.
 
 

 
 
 
 Sign up
 |
 Log in
 
 

 

 

 
 
 
 
 Upload icon
 An illustration of a horizontal line over an up
 pointing arrow.
 

 

 Upload
 
 
 
 
 
 Search icon
 An illustration of a magnifying glass.
 

 

 
 
 

 
 Search the Archive
 
 
 
 
 
 Search icon
 An illustration of a magnifying glass.
 

 

 
 
 

 

 

 
 

 
 

 

 

 

 
 
Internet Archive Audio

 

 
 Live Music
 Archive
 
 Librivox
 Free Audio
 
 

 

 
Featured

 
 
 
All Audio

 
 
Grateful Dead

 
 
Netlabels

 
 
Old Time Radio
 

 
 
78 RPMs
 and Cylinder Recordings

 
 
 

 

 
Top

 
 
 
Audio Books
 & Poetry

 
 
Computers,
 Technology and Science

 
 
Music, Arts
 & Culture

 
 
News &
 Public Affairs

 
 
Spirituality
 & Religion

 
 
Podcasts

 
 
Radio News
 Archive

 
 
 

 
 
 
Images

 

 
 Metropolitan Museum
 
 Cleveland
 Museum of Art
 
 

 

 
Featured

 
 
 
All Images

 
 
Flickr Commons
 

 
 
Occupy Wall
 Street Flickr

 
 
Cover Art

 
 
USGS Maps

 
 
 

 

 
Top

 
 
 
NASA Images

 
 
Solar System
 Collection

 
 
Ames Research
 Center

 
 
 

 
 
 
Software

 

 
 Internet
 Arcade
 
 Console Living Room
 
 

 

 
Featured

 
 
 
All Software
 

 
 
Old School
 Emulation

 
 
MS-DOS Games
 

 
 
Historical
 Software

 
 
Classic PC
 Games

 
 
Software
 Library

 
 
 

 

 
Top

 
 
 
Kodi
 Archive and Support File

 
 
Vintage
 Software

 
 
APK

 
 
MS-DOS

 
 
CD-ROM
 Software

 
 
CD-ROM
 Software Library

 
 
Software Sites
 

 
 
Tucows
 Software Library

 
 
Shareware
 CD-ROMs

 
 
Software
 Capsules Compilation

 
 
CD-ROM Images
 

 
 
ZX Spectrum

 
 
DOOM Level CD
 

 


... (truncated, 6 KB total)
Resource ID: e09fc9ef04adca70 | Stable ID: sid_EDA3WzGX8z