Multi-armed bandit

Thompson is Python package to evaluate the multi-armed bandit problem. In addition to thompson, Upper Confidence Bound (UCB) algorithm, and randomized results are also implemented.
In probability theory, the multi-armed bandit problem is a problem in which a fixed limited set of resources must be allocated between competing (alternative) choices in a way that maximizes their expected gain, when each choice's properties are only partially known at the time of allocation, and may become better understood as time passes or by allocating resources to the choice. This is a classic reinforcement learning problem that exemplifies the exploration-exploitation tradeoff dilemma wikipedia.
In the problem, each machine provides a random reward from a probability distribution specific to that machine. The objective of the gambler is to maximize the sum of rewards earned through a sequence of lever pulls. The crucial tradeoff the gambler faces at each trial is between "exploitation" of the machine that has the highest expected payoff and "exploration" to get more information about the expected payoffs of the other machines. The trade-off between exploration and exploitation is also faced in machine learning. In practice, multi-armed bandits have been used to model problems such as managing research projects in a large organization like a science foundation or a pharmaceutical company wikipedia.

⭐️ Star this repo if you like it ⭐️

Install thompson from PyPI

pip install thompson

Import thompson package

import thompson as th

Documentation pages

On the documentation pages you can find detailed information about the working of the thompson with examples.

Examples

Example: Compute multi-armed bandit using Thompson

Example: Compute multi-armed bandit using UCB-Upper confidence Bound

Example: Compute multi-armed bandit using randomized data

References

https://en.wikipedia.org/wiki/Multi-armed_bandit

Maintainers

Erdogan Taskesen, github: erdogant

Contribute

All kinds of contributions are welcome!
If you wish to buy me a Coffee for this work, it is very appreciated :)

Name		Name	Last commit message	Last commit date
Latest commit History 45 Commits
.github		.github
docs		docs
thompson		thompson
.gitignore		.gitignore
CITATION.cff		CITATION.cff
LICENSE		LICENSE
MANIFEST.in		MANIFEST.in
README.md		README.md
make_clean.sh		make_clean.sh
make_sphinx_and_commit.sh		make_sphinx_and_commit.sh
requirements.txt		requirements.txt
setup.cfg		setup.cfg
setup.py		setup.py

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

Multi-armed bandit

Install thompson from PyPI

Import thompson package

Documentation pages

Examples

References

Maintainers

Contribute

About

Releases 3

Sponsor this project

Packages

Languages

License

erdogant/thompson

Folders and files

Latest commit

History

Repository files navigation

Multi-armed bandit

Install thompson from PyPI

Import thompson package

Documentation pages

Examples

References

Maintainers

Contribute

About

Topics

Resources

License

Stars

Watchers

Forks

Releases 3

Sponsor this project

Packages 0

Languages

Packages