Some video encoders already incorporate perceptual optimizations (eg, x264's psy-rd and psy-trellis) that "look" better but lead to objectively worse results with traditional image quality metrics.
Audio codecs, however, have been using psychoacoustic models for decades. Frequencies outside the human hearing range are clipped, masked noises are discarded, voice codecs emphasize the range of human speech, etc.
Audio codecs, however, have been using psychoacoustic models for decades. Frequencies outside the human hearing range are clipped, masked noises are discarded, voice codecs emphasize the range of human speech, etc.