AIXI
Mathematical formalism for artificial general intelligence combining Solomonoff induction with sequential decision theory
AIXI is a theoretical mathematical formalism for artificial general intelligence. It combines Solomonoff induction with sequential decision theory. AIXI was first proposed by Marcus Hutter in 2000 and several results regarding AIXI are proved in Hutter's 2005 book Universal Artificial Intelligence.
Nº Q18204908 ★
Commune · Savoirs
AIXI
Mathematical formalism for artificial general intelligence combining Solomonoff induction with sequential decision theory
AIXI is a theoretical mathematical formalism for artificial general intelligence. It combines Solomonoff induction with sequential decision theory. AIXI was first proposed by Marcus Hutter in 2000 and several results regarding AIXI are proved in Hutter's 2005 book Universal Artificial Intelligence.
Dernier prix
—
Prix plancher
—
Médiane 7 j
—
Ventes 30 j
0
Fourchette 30 j
—
En circulation
0
Cours
médiane
min – max
ventes
Aucune vente sur la période
Voir le tableau
| Date | médiane | Min | Max | ventes |
|---|
Historique des ventes
- Dernière vente
- —
- Moyenne 30 j
- —
- Plus bas 30 j
- —
- Plus haut 30 j
- —
- Ventes 7 j
- 0
- Ventes 30 j
- 0
Aucune vente pour l'instant.
Ventes anonymes : ni acheteur ni vendeur. Les chiffres ne comptent que les ventes entre joueurs.
Sur Wikipédia
Texte en anglais Pas encore d'article dans ta langue : extrait en anglais.
AIXI is a theoretical mathematical formalism for artificial general intelligence. It combines Solomonoff induction with sequential decision theory. AIXI was first proposed by Marcus Hutter in 2000 and several results regarding AIXI are proved in Hutter's 2005 book Universal Artificial Intelligence. AIXI is a reinforcement learning (RL) agent. It maximizes the expected total rewards received from the environment. Intuitively, it simultaneously considers every computable hypothesis (or environment). In each time step, it looks at every possible program and evaluates how many rewards that program generates depending on the next action taken. The promised rewards are then weighted by the subjective belief that this program constitutes the true environment. This belief is computed from the length of the program: longer programs are considered less likely, in line with Occam's razor. AIXI then selects the action that has the highest expected total reward in the weighted sum of all these programs.
Texte : Wikipédia en anglais, CC BY-SA 4.0. ·
Cartes voisines
Intelligence artificielle générale
Type d'intelligence artificielle
Nº Q2264109 ★★★★
Computational learning theory
Theory of machine learning
Nº Q2462783 ★
Apprentissage par renforcement à partir de rétroaction humaine
Technique pour entraîner une IA
Nº Q115570683 ★★★
intelligence artificielle explicable
Domaine de l'intelligence artificielle
Nº Q40890078 ★★
Allen Institute for Artificial Intelligence
Institut de recherche
Nº Q16002567 ★
Apprentissage par renforcement
Sous domaine de l'apprentissage automatique
Nº Q830687 ★★★