AIXI
Mathematical formalism for artificial general intelligence combining Solomonoff induction with sequential decision theory
AIXI is a theoretical mathematical formalism for artificial general intelligence. It combines Solomonoff induction with sequential decision theory. AIXI was first proposed by Marcus Hutter in 2000 and several results regarding AIXI are proved in Hutter's 2005 book Universal Artificial Intelligence.
Nº Q18204908 ★
Comum · Saberes
AIXI
Mathematical formalism for artificial general intelligence combining Solomonoff induction with sequential decision theory
AIXI is a theoretical mathematical formalism for artificial general intelligence. It combines Solomonoff induction with sequential decision theory. AIXI was first proposed by Marcus Hutter in 2000 and several results regarding AIXI are proved in Hutter's 2005 book Universal Artificial Intelligence.
Último preço
—
Preço mínimo
—
Mediana 7 d
—
Vendas 30 d
0
Faixa 30 d
—
Em circulação
0
Cotação
mediana
mín – máx
vendas
Sem vendas no período
Ver tabela
| Data | mediana | Mín | Máx | vendas |
|---|
Histórico de vendas
- Última venda
- —
- Média 30 d
- —
- Mínima 30 d
- —
- Máxima 30 d
- —
- Vendas 7 d
- 0
- Vendas 30 d
- 0
Ainda sem vendas.
Vendas anônimas: sem comprador nem vendedor. Os números contam só vendas entre jogadores.
Na Wikipédia
Texto em inglês Ainda não há artigo no seu idioma: trecho em inglês.
AIXI is a theoretical mathematical formalism for artificial general intelligence. It combines Solomonoff induction with sequential decision theory. AIXI was first proposed by Marcus Hutter in 2000 and several results regarding AIXI are proved in Hutter's 2005 book Universal Artificial Intelligence. AIXI is a reinforcement learning (RL) agent. It maximizes the expected total rewards received from the environment. Intuitively, it simultaneously considers every computable hypothesis (or environment). In each time step, it looks at every possible program and evaluates how many rewards that program generates depending on the next action taken. The promised rewards are then weighted by the subjective belief that this program constitutes the true environment. This belief is computed from the length of the program: longer programs are considered less likely, in line with Occam's razor. AIXI then selects the action that has the highest expected total reward in the weighted sum of all these programs.
Texto: Wikipédia em inglês, CC BY-SA 4.0. ·
Cartas próximas
Inteligência artificial geral
Nº Q2264109 ★★★★
Aprendizado por reforço com feedback humano
Nº Q115570683 ★★★
Inteligência artificial explicável
IA em que os resultados da solução podem ser compreendidos por humanos
Nº Q40890078 ★★
Aprendizagem por reforço
Área do aprendizado de máquinas
Nº Q830687 ★★★
MIT Computer Science and Artificial Intelligence Laboratory
Computer Science research institute at MIT
Nº Q1354917 ★
Autoaperfeiçoamento recursivo
Inteligência artificial capaz de se modificar para aprimorar ainda mais suas capacidades
Nº Q1768494 ★★★