ERROR: The request could not be satisfied
Comments on NIST’s draft document “AI Risk Management Framework: Generative AI Profile.”
Topics: Safety Evals
Entities: Safety Evals
Topic feed
Safety evaluations, red teaming, preparedness, and model risk testing.
Comments on NIST’s draft document “AI Risk Management Framework: Generative AI Profile.”
Topics: Safety Evals
Entities: Safety Evals
We’re developing a blueprint for evaluating the risk that a large language model (LLM) could aid someone in creating a biological threat. In an evaluation involving both biology experts and students, we found that GPT-4 provides at most a mild uplift in...
Topics: Safety Evals
Entities: Safety Evals
To support the safety of highly-capable AI systems, we are developing our approach to catastrophic risk preparedness, including building a Preparedness team and launching a challenge.
Topics: Safety Evals
Entities: Safety Evals