Core Ethics Principles and Trustworthy AI

The canonical principles — fairness, transparency, explainability, accountability, privacy, safety and robustness, human oversight — where they came from (OECD, UNESCO, EU HLEG), how NIST turns them into trustworthiness characteristics, why fairness definitions mathematically conflict, and what it takes to move from principles to practice.

Lessons

  1. Where the principles came from
  2. The canonical principles, one by one
  3. Fairness: pick your definition, because you cannot have them all
  4. Trustworthy AI: the same principles, three official dialects
  5. From principles to practice — and why principles alone failed