I may be all wrong about this (and welcome any dissenting opinions) but it seems to me that:
(1) Bayesian data analysis isn't actually necessary for most AI/ML work. Bayes theorem is fundamental of course, but Bayesian data analysis (BDA) seems to rarely come up in practice. ML algorithms like Naive Bayes don't really require knowledge of BDA, plus PGMs aren't really that common (and not within BDA's scope anyway).
The computational methods behind BDA (mostly MCMC-based) are also fairly heavy and I don't know too many ML shops that actually use MCMC-based Bayesian ML.
To me, BDA's primary value is to "upgrade" the type of NHST-based data analysis done in science and social science, and ground it on what I think is more solid epistemology. I don't know if BDA is really that practical for machine learning, where the goal is not analysis but prediction.
(2) Kruschke might not be the best book for learning BDA. I own a copy of Krushke (puppies on cover and all), and found the first few chapters interesting, but it then quickly got tedious. It seemed to me that Kruschke, in an effort to make things accessible to social scientists, belabors the subject a little in later chapters without adding pedagogical value (I realize this is controversial statement to Kruschke fans, and I am prepared to change my mind).
Gelman's BDA on the other hand (I also own a copy) is less accessible to beginners, but ultimately rewards the reader in a more consistent manner.