Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans
The code of conduct lays out general principles that Microsoft AI models should uphold — supporting humans rather than r
Overview
The code of conduct lays out general principles that Microsoft AI models should uphold — supporting humans rather than replacing them, for instance, and accelerating human flourishing — as well as specific safety constraints meant to implement those principles.
Similar Tools
Verified Info
Ratings & Reviews
Rate Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans
Alternatives to Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans
View AllGoogle's AI safety program for government and enterprise security.
Contributes to shared safety standards and evaluation frameworks for advanced AI systems.
Automated red teaming system that tests AI safety through self-play.
OpenAI's cybersecurity AI tools and training for critical infrastructure defenders.
OpenAI's official security report on the Hugging Face breach incident.
Monitors AI model outputs to detect and prevent harmful or non-compliant responses.