Надежный ИИ относится к ИИ-системам, проектируемым, разрабатываемым и развертываемым так, чтобы поддерживать надежность, безопасность, подотчетность, справедливость, прозрачность, приватность и человеческий надзор. Этот термин часто используется в governance frameworks для описания ИИ, на который можно ответственно полагаться в его intended context.
The NIST AI RMF defines trustworthiness in AI as 'responsive[ness] to a multiplicity of criteria that are of value to interested parties.' It specifies that such values include 'valid and reliable, safe, secure and resilient, accountable and transparent, explainable and interpretable, privacy-enhanced, and fair with harmful bias managed.' The White House Voluntary Commitments specify that 'trust,' together with 'safety' and 'security,' comprise the 'three principles that must be fundamental to the future of AI.'