Bandit based Monte-Carlo planning [pdf]
web.engr.oregonstate.edu
web.engr.oregonstate.edu
It is also very easy to work with, you can easily tweak the algorithm and add heuristics for your specific domains.
Also relevant:
UCT applied to partially observable game (Poc-man) http://papers.nips.cc/paper/4031-monte-carlo-planning-in-lar...
Another approach for Monte-Carlo planning http://papers.nips.cc/paper/5189-despot-online-pomdp-plannin...
However, the tutorial doesn't work towards a working implementation. I think you can verify your results against benchmark problems. There are a number of good implementations around:
UCT implementation with many MDP benchmark problems: https://github.com/bonetblai/mdp-engine
My favorite implementation. The code is quite easy to read: http://www0.cs.ucl.ac.uk/staff/D.Silver/web/Applications_fil...