Papers › Self-timed Reinforcement Learning using Tsetlin Machine

Self-timed Reinforcement Learning using Tsetlin Machine

2 Sep 2021arXiv:2109.00846archive 2025-07-28

Adrian Wheeldon, Alex Yakovlev, Rishad Shafik

We present a hardware design for the learning datapath of the Tsetlin machine algorithm, along with a latency analysis of the inference datapath. In order to generate a low energy hardware which is suitable for pervasive artificial intelligence applications, we use a mixture of asynchronous design techniques - including Petri nets, signal transition graphs, dual-rail and bundled-data. The work builds on previous design of the inference hardware, and includes an in-depth breakdown of the automaton feedback, probability generation and Tsetlin automata. Results illustrate the advantages of asynchronous design in applications such as personalized healthcare and battery-powered internet of things devices, where energy is limited and latency is an important figure of merit. Challenges of static timing analysis in asynchronous circuits are also addressed.

PaperPDFCode

Code

cair/TsetlinMachine officialmentioned in papermentioned on GitHub report
cair/PyTsetlinMachineCUDA mentioned on GitHub report
cair/pyTsetlinMachine mentioned on GitHub report
cair/pyTsetlinMachineMT mentioned on GitHub report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Reinforcement LearningReinforcement Learning (RL)reinforcement-learning

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections