The NUC model assumes that rational agents undergo a hierarchical planning process: First select an intention (a sequence/plan of goals), then take actions to achieve a goal.
Would it be possible to model agents using a "flat" planning process, where agents directly optimize for subjective utility (rewards minus costs), without going through the process of selecting goals?
If so, how would that generative model differ in its predictions?