tau2's current reward_info carries 'reward' (composite) + db_check / action_checks; the scorer read the old environment_reward / communication_reward split that no longer exists -- simulations scored 1.0 came out 0.0. Fallback to the old split kept for older engines. Co-Authored-By: Claude <noreply@anthropic.com>