An extra parsing layer damages the same reply.
Across all eight same-window configurations, replaying the same raw-conditioned reply through one added parser lowers success by 55.4–73.2 points. Disclosing that boundary before generation produces substantial compensation in six configurations.