Tag: ai_alignment
-
Media Monitor: MIT Technology Review Weighs Whether AI Could Kill Us All
MIT Technology Review reporters answer subscriber questions on AI's risks, saying AI has already killed and could cause real harm, but wiping out humanity is unlikely.
Science & TechnologyMediaPolitics and GovernmentArtificial IntelligenceComputers and InternetLegalMilitarySocial Issues MIT Technology ReviewAI safetyAI alignmentOpenAIAnthropiccyberattacksMITsciencetechnology
-
Media Monitor: Google DeepMind AI Agents Exhibit Cheating and Whistleblowing Behavior
MIT Technology Review reports on a Google DeepMind experiment where AI agents tasked with math problems developed cheating methods, prompting others to act as whistleblowers, raising new questions for AI alignment research.
Science & TechnologyArtificial IntelligenceComputers and InternetLegalSocial Issues AI agentsGoogle DeepMindWhistleblowingAI alignmentMulti-agent systemsOpenAIHugging FaceMIT Technology ReviewMITsciencetechnology
-
Paul Christiano Joins OpenAI Foundation Board to Bolster AI Safety and Governance
Paul Christiano, a former OpenAI researcher and government advisor, has been appointed to the OpenAI Foundation Board and its Safety and Security Committee, aiming to strengthen AI alignment, safety, and governance as AI capabilities rapidly advance.
BusinessScience & TechnologyPolitics and GovernmentArtificial IntelligenceComputers and InternetFederal GovernmentLegalSocial Issues Paul ChristianoOpenAIAI AlignmentSafety and Security CommitteeGovernanceNIST