AIXI
Mathematical formalism for artificial general intelligence combining Solomonoff induction with sequential decision theory
AIXI is a theoretical mathematical formalism for artificial general intelligence. It combines Solomonoff induction with sequential decision theory. AIXI was first proposed by Marcus Hutter in 2000 and several results regarding AIXI are proved in Hutter's 2005 book Universal Artificial Intelligence.
Nº Q18204908 ★
Common · Knowledge
AIXI
Mathematical formalism for artificial general intelligence combining Solomonoff induction with sequential decision theory
AIXI is a theoretical mathematical formalism for artificial general intelligence. It combines Solomonoff induction with sequential decision theory. AIXI was first proposed by Marcus Hutter in 2000 and several results regarding AIXI are proved in Hutter's 2005 book Universal Artificial Intelligence.
Last price
—
Floor price
—
7-day median
—
30-day sales
0
30-day range
—
In circulation
0
Price history
median
low – high
sales
No sales in this period
Show table
| Date | median | Low | High | sales |
|---|
Sales history
- Last sale
- —
- 30-day average
- —
- 30-day low
- —
- 30-day high
- —
- Sales 7d
- 0
- Sales 30d
- 0
No sales yet.
Anonymous sales: no buyer or seller shown. Figures count player-to-player sales only.
From Wikipedia
AIXI is a theoretical mathematical formalism for artificial general intelligence. It combines Solomonoff induction with sequential decision theory. AIXI was first proposed by Marcus Hutter in 2000 and several results regarding AIXI are proved in Hutter's 2005 book Universal Artificial Intelligence. AIXI is a reinforcement learning (RL) agent. It maximizes the expected total rewards received from the environment. Intuitively, it simultaneously considers every computable hypothesis (or environment). In each time step, it looks at every possible program and evaluates how many rewards that program generates depending on the next action taken. The promised rewards are then weighted by the subjective belief that this program constitutes the true environment. This belief is computed from the length of the program: longer programs are considered less likely, in line with Occam's razor. AIXI then selects the action that has the highest expected total reward in the weighted sum of all these programs.
Text: Wikipédia, CC BY-SA 4.0. ·
Related cards
Artificial general intelligence
Theoretical class of AI able to perform any intelligence-based task humans can
Nº Q2264109 ★★★★
Reinforcement learning from human feedback
Training method using human feedback to rank responses and train a reward model that improves model outputs
Nº Q115570683 ★★★
Explainable artificial intelligence
AI whose processes can be understood by humans
Nº Q40890078 ★★
Reinforcement learning
Type of machine learning where an agent learns how to behave in an environment by performing actions and receiving rewards or penalties in return, aiming to maximize the cumulative reward over time
Nº Q830687 ★★★
MIT Computer Science and Artificial Intelligence Laboratory
Computer Science research institute at MIT
Nº Q1354917 ★
Recursive self-improvement
Ability of an artificial intelligence to modify itself to further improve its capability
Nº Q1768494 ★★★