6 years, 3 months ago
Highlights
6 years, 3 months ago
alvaro said:the way to use a policy heuristic is to use it to initialize counts to some fake values that will make the first few playouts through this node approximately follow…
6 years, 3 months ago
Hello,I want to add heuristics to an MCTS implementation but I still want MCTS to “take over” and make the final decision. Say I have a function policy() that returns a value fr…
6 years, 3 months ago
If I use MCTS but with "reward" as -1, 0, and 1 for lose, draw, and win respectively, can I use the UCT formula as is? uct = node.rewards/(node.visits+1.0) + explorationRate * s…
8 years ago
Edited for brevity. I made a post 8 years ago about this Stratego-like game: https://www.gamedev.net/forums/topic/585572-game-of-the-generals-ai/. Are there advances in AI in th…
8 years, 1 month ago
Latest Activity
See all in Discussionsalvaro said:the way to use a policy heuristic is to use it to initialize counts to some fake values that …
6 years, 3 months ago
Thanks for your reply & apologies for the confusion. I've asked questions about MCTS/UCT and discussed extensively in this forum …
6 years, 3 months ago
Hello,I want to add heuristics to an MCTS implementation but I still want MCTS to “take over” and make the …
6 years, 3 months ago