La IA confiable se refiere a sistemas de IA diseñados, desarrollados y desplegados de formas que apoyan fiabilidad, seguridad, rendición de cuentas, equidad, transparencia, privacidad y supervisión humana. El término se usa a menudo en marcos de gobernanza para describir IA en la que puede confiarse responsablemente dentro de su contexto previsto.
The NIST AI RMF defines trustworthiness in AI as 'responsive[ness] to a multiplicity of criteria that are of value to interested parties.' It specifies that such values include 'valid and reliable, safe, secure and resilient, accountable and transparent, explainable and interpretable, privacy-enhanced, and fair with harmful bias managed.' The White House Voluntary Commitments specify that 'trust,' together with 'safety' and 'security,' comprise the 'three principles that must be fundamental to the future of AI.'