That would make sense. But, would still need to be quantified.
I can create an anomaly in the "encoded signal" but still wouldn't know the extent of which would trigger any sort of "recovery mechanism". E.g., if I zero out a few samples, does that just become an imperceptible click/pop? Likewise, if I inject several 0xFFFF, would I be able to aurally identify them in the audio stream?
At what point do I have to decide to take alternative LISTENING measures?
If the player squelches the output to hide an anomaly. Then, unsquelches and squelches the NEXT anomaly a fraction of a second later (etc), I suspect the listener would quickly become annoyed. Would the player's designers assume the listener should be the ultimate "handler" for such a condition? Or, would they take measures (to protect the perceived value of their technology) to avoid such an annoying presentation?
[Contrast with how dreadfully DTV fails when it does]