In recent years, there has been rapid development in learned image compression techniques that prioritize ratedistortion-perceptual compression, preserving fine details even at lower bit-rates. However, current learning-based image compression methods often sacrifice human-friendly compression and require long decoding times. In this paper, we propose enhancements to the backbone network and loss function of existing image compression model, focusing on improving human perception and efficiency. Our proposed approach achieves competitive subjective results compared to state-of-the-art end-to-end learned image compression methods and classic methods, while requiring less decoding time and offering human-friendly compression. Through empirical evaluation, we demonstrate the effectiveness of our proposed method in achieving outstanding performance, with more than 25% bit-rate saving at the same subjective quality.
翻译:近年来,优先考虑率失真-感知压缩的学习型图像压缩技术发展迅速,即使在较低比特率下也能保留精细细节。然而,当前基于学习的图像压缩方法往往牺牲了人性化压缩,且解码时间较长。本文对现有图像压缩模型的骨干网络和损失函数进行了改进,重点提升人类感知效果和效率。与最先进的端到端学习型图像压缩方法及经典方法相比,我们提出的方法在减少解码时间的同时实现了具有竞争力的主观结果,并能提供人性化压缩。通过实证评估,我们证明了所提方法在实现卓越性能方面的有效性——在相同主观质量下可节省超过25%的比特率。