Skip to content

Safety and alignment

AI safety and alignment: jailbreaks and defences, research on model behaviour, safety evaluations and governance frameworks.

Latest selected

21–23 of 23

Jun 16

Tuesday

Apr 30

Thursday
  1. Enabling a new model for healthcare with AI co-clinician

    Google DeepMind announced AI co-clinician research to explore AI agents assisting patient care under clinical supervision. In blinded assessments of 98 real primary-care queries, 97 responses had no critical errors, and doctors preferred them to existing evidence-synthesis tools. Across 140 consultation skills, AI matched or exceeded primary-care doctors on 68, but expert doctors were better overall at recognising red flags and guiding key physical examinations.

Mar 31

Tuesday