Uncommon · Knowledge
Mechanistic interpretability
Reverse-engineering neural networks
Mechanistic interpretability (sometimes abbreviated as mech interp, mechinterp, or MI) is a subfield of research within explainable artificial intelligence that aims to understand the internal workings of neural networks by analyzing their concrete structures, algorithms and circuits. This approach seeks to analyze neural networks in a manner similar to the reverse engineering of conventional software.
From Wikipedia
Mechanistic interpretability (sometimes abbreviated as mech interp, mechinterp, or MI) is a subfield of research within explainable artificial intelligence that aims to understand the internal workings of neural networks by analyzing their concrete structures, algorithms and circuits. This approach seeks to analyze neural networks in a manner similar to the reverse engineering of conventional software.
Text: Wikipédia, CC BY-SA 4.0. ·
Related cards
-
P★
Principle of maximum entropy
Principle in Bayesian statistics
-
★★
MEMS
Technology of very small devices
-
★★★
Statistical mechanics
Physics of large number of particles' statistical behavior
-
R★★
Rote learning
Memorization technique based on repetition
-
★★★
Mechanism (engineering)
Device designed to transform input forces and movement into a desired set of output forces and movement
-
★★
Explainable artificial intelligence
AI whose processes can be understood by humans