Aligning AI With Shared Human Values (ICLR 2021)
-
Updated
Apr 21, 2023 - Python
Aligning AI With Shared Human Values (ICLR 2021)
[AAAI 2018] Implementation of the Ethics Shaping approach proposed in "A low-cost ethics shaping approach for designing reinforcement learning agents"
Code and data for Paper "Enhancing Ethical Explanations of Large Language Models through Iterative Symbolic Refinement"
Reinforcement learning environment for learning ethical behaviours in a SmartGrid use-case.
AI-HPP-Standard: an inspection-ready architecture for accountable AI systems. Vendor-neutral. Audit-ready. High-risk gated. Developed via structured multi-model orchestration with human oversight. Designed to support emerging international AI governance.
Uses Montague semantics and a deterministic graph database to enforce mathematically verified, immutable ethical logic, prevents utilitarian overrides of deontological constraints via an immutable constraint layer.
Ethical AI governance framework for multi-model alignment, integrity, and enterprise oversight.
Add a description, image, and links to the machine-ethics topic page so that developers can more easily learn about it.
To associate your repository with the machine-ethics topic, visit your repo's landing page and select "manage topics."