metaai·lightalo unofficial · independent
Universe / Games & Strategy / ReBeL
Games & Strategy · 2020

ReBeL

AlphaZero for hidden-information games — the poker code stayed locked

open source

Latest: NeurIPS 2020 paper + open-source Liar's Dice code (December 2020)

Recursive Belief-based Learning: a general RL-plus-search algorithm extending AlphaZero-style self-play to imperfect-information games by operating on public belief states. ReBeL reached superhuman heads-up no-limit hold'em with far less poker-specific knowledge than prior bots. Meta open-sourced a Liar's Dice implementation alongside the NeurIPS 2020 paper.

Why it matters

Extended AlphaZero-style self-play plus search to imperfect-information games by operating on public belief states, reaching superhuman heads-up no-limit hold'em with far less poker-specific knowledge than earlier bots. The Apache-licensed Liar's Dice implementation, with released value-function checkpoints, is the accessible entry point.

Facts

Try it yourself

Lineage

Descends fromPluribus
Led toCICERO

See the whole family tree →

Sources

More in Games & Strategy

ELF OpenGo

Read the Games & Strategy story on the sky →

✦ Open on the map Explore Games & Strategy Quiz me