Vertrauenswürdige KI bezeichnet KI-Systeme, die so entworfen, entwickelt und eingesetzt werden, dass sie Zuverlässigkeit, Sicherheit, Verantwortlichkeit, Fairness, Transparenz, Datenschutz und menschliche Aufsicht unterstützen. Der Begriff wird häufig in Governance-Rahmen verwendet, um KI zu beschreiben, auf die in ihrem vorgesehenen Kontext verantwortungsvoll vertraut werden kann.
The NIST AI RMF defines trustworthiness in AI as 'responsive[ness] to a multiplicity of criteria that are of value to interested parties.' It specifies that such values include 'valid and reliable, safe, secure and resilient, accountable and transparent, explainable and interpretable, privacy-enhanced, and fair with harmful bias managed.' The White House Voluntary Commitments specify that 'trust,' together with 'safety' and 'security,' comprise the 'three principles that must be fundamental to the future of AI.'