I am only familiar with the basics of Reinforcement Learning and came across two terms that I don't understand: path-wise cost function and the terminal cost for the given system. Are these terms analogous to those in the Bellman Equation? Attaching a snippet of the paper for context.