Reward is hyperstitional information

·LessWrong··

(This article broadly explains mirror ascent, continuous Bayesian inference and information geometry in full. Title refers to the result in section 3.) The logarithmic scoring rule mjx-container[jax="CHTML"] { line-height: 0; } mjx-container [space="1"] { margin-left: .111em; } mjx-container [space="2"] { margin-left: .167em; } mjx-container [space="3"] { margin-left: .222em; } mjx-container [space="4"] { margin-left: .278em; } mjx-container [space="5"] { margin-left: .333em; } mjx-container [rs...

Read full article →

Related Articles

Nashville uses eminent domain to block data center near zoo
mapping365 · Hacker News · 8h ago
Muse Code and Muse Spark 1.2
paulkrush · Hacker News · 15h ago
Xbox goes down. You can't play games you own on disc
surprisetalk · Hacker News · 1d ago
Civilian plane crash in New Mexico tied to military GPS blocking
dzdt · Hacker News · 23h ago
Ten advances in mathematics and theoretical computer science
milkshakes · Hacker News · 2d ago