Piloting the world's first double-blind AI evaluations
Building trust in proprietary model benchmarks using cryptographically secure environments
Topics: Benchmarks
Entities: Benchmarks
Building trust in proprietary model benchmarks using cryptographically secure environments
Topics: Benchmarks
Entities: Benchmarks
Google DeepMind and partners are announcing a new technical research funding call of up to $10M for researchers worldwide to strengthen multi-agent safety.
Entities: GoogleGoogle DeepMind
Google DeepMind releases new findings and an evaluation framework to measure AI's potential for harmful manipulation in areas like finance and health, with the goal of enhancing AI safety.
Topics: Safety EvalsTesting Tools
Google DeepMind proposes a cognitive framework to evaluate AGI and launches a Kaggle hackathon to build capability benchmarks
Topics: Benchmarks
Entities: BenchmarksGoogleGoogle DeepMind
Announcing Gemma Scope 2, a comprehensive, open suite of interpretability tools for the entire Gemma 3 family to accelerate AI safety research.
Topics: Safety Evals
Entities: Safety Evals
Google DeepMind and the UK AI Security Institute (AISI) strengthen collaboration through a new research partnership, focusing on critical safety research areas like monitoring AI reasoning and evalua…
Topics: Safety Evals
Entities: Safety EvalsGoogleGoogle DeepMind
The FACTS Benchmark Suite provides a systematic evaluation of Large Language Models (LLMs) factuality across three areas: Parametric, Search, and Multimodal reasoning.
Topics: Benchmarks
Entities: Benchmarks
Today, we’re publishing the third iteration of our Frontier Safety Framework (FSF) — our most comprehensive approach yet to identifying and mitigating severe risks from advanced AI models. This updat…
Entities: GoogleGoogle DeepMind
Kaggle Game Arena is a new platform where AI models compete head-to-head in complex strategic games.
We’re exploring the frontiers of AGI, prioritizing technical safety, proactive risk assessment, and collaboration with the AI community.
Topics: Safety Evals
Entities: Safety Evals