Trustworthy AI: Definition and Examples
Trustworthy AI refers to artificial intelligence designed to be reliable, ethical, transparent, and respectful of fundamental rights, ensuring that its decisions and behaviors merit the trust of users and society.
Full definition
The concept of Trustworthy AI rests on the idea that an artificial intelligence system must not only be technically performant but also trustworthy on ethical, social, and legal levels. It is a holistic approach that encompasses algorithm transparency, fairness of outcomes, technical robustness, and respect for privacy.
This framework was widely popularized by the European Commission's guidelines in 2019, which identify seven key requirements: human oversight, technical robustness and safety, privacy, transparency, diversity and non-discrimination, societal and environmental well-being, and accountability. These principles now form the foundation of the European AI Act and influence global regulations.
In prompt engineering, the notion of Trustworthy AI translates into designing prompts that encourage verifiable, unbiased, and transparent responses. This involves explicitly asking the AI to cite its sources, signal its uncertainties, avoid stereotypes, and respect the limits of its knowledge rather than fabricating information.
The stakes are fundamental: as AI integrates into critical domains such as healthcare, justice, and finance, the trust placed in these systems must rest on concrete and measurable guarantees, not merely technological promises. Trustworthy AI thus represents a paradigm shift from AI optimized solely for performance to AI optimized for reliability and accountability.
Etymology
The term combines 'trustworthy' (from Old English 'treowwyrðe') and 'AI' (Artificial Intelligence). It became established in institutional vocabulary around 2018-2019, notably with the publication of the 'Ethics Guidelines for Trustworthy AI' by the European Commission's High-Level Expert Group on AI.
Concrete examples
Source verification and transparency
Analyze this medical report and provide your conclusions. For each assertion, indicate your confidence level (high, medium, low) and explicitly signal if you do not have sufficient data to conclude.
Bias reduction in analysis
Evaluate these 5 resumes for the senior developer position. Focus solely on technical skills and relevant experience. Ignore any information related to gender, ethnic origin, or age. Explain your evaluation criteria before starting.
Audit of an existing AI system
Act as a specialized auditor in responsible AI. Analyze this credit scoring model according to the 7 criteria of the European Commission's Trustworthy AI. For each criterion, assign a compliance score and identify risks.
Practical usage
In prompt engineering, applying Trustworthy AI involves formulating instructions that enforce transparency: asking the AI to distinguish verified facts from assumptions, to explain its reasoning step by step, and to signal its limitations. Systematically integrate safeguards into your prompts, such as 'if you are not sure, say so' or 'cite your sources.' This discipline not only improves the reliability of responses but also makes it easier to detect hallucinations and biases.
Related concepts
FAQ
What is the difference between Trustworthy AI and Responsible AI?
How is Trustworthy AI linked to the European AI Act?
Can we concretely measure whether an AI is 'trustworthy'?
See also
How to use this prompt
- Copy the prompt with the button above.
- Paste it into ChatGPT, Claude or your favorite AI assistant.
- Replace the bracketed variables with your details, then refine the result.
About Prompt Guide
Prompt Guide is a free library of 2500+ ready-to-use prompts for ChatGPT, Claude and other AIs, with guides to learn prompting and tools to build and optimize your own prompts.
More definitions
Underfitting: Definition and Examples
Underfitting occurs when an artificial intelligence model is too simple to capture the patterns in the training data, resulting in poor performance on both training data and new data.
Unsupervised Learning: Definition and Examples
Unsupervised learning is a branch of machine learning where a model analyzes data without prior labels to discover structures, patterns, or groupings within it.
Vector Database: Definition and Examples
A vector database is a specialized database for storing, indexing, and searching numerical vectors (embeddings), enabling...
Vercel AI SDK: Definition and Examples
The Vercel AI SDK is an open-source library developed by Vercel that makes it easy to integrate generative AI models (like
Video Understanding: Definition and Examples
Ability of an AI model to analyze, interpret, and extract relevant information from video content, combining visual, temporal, and often audio understanding.
Virtual Assistant: Definition and Examples
A virtual assistant is a computer program powered by artificial intelligence, capable of understanding natural language instructions and performing tasks on behalf of a user.
Get new prompts every week
Join our newsletter.