Concepts

Interpretability

Reference entrySafety, ethics and evaluation
Reference entryInterpretability

Interpretability seeks to understand how a model processes information and reaches outputs, using analyses of behavior, representations, parameters, or internal computations.

Also known as Explainable AI, XAI.

Where it sits