In this paper, techniques for improving multichannel lossless coding are examined. A method is proposed for the simultaneous coding of two or more different renderings (mixes) of the same content. The signal model uses both past samples of the upmix, and the current time samples of downmix samples to predict the upmix. Model parameters are optimized via a general linear solver, and the prediction residual is Rice coded. Additionally, the use of an SVD projection prior to residual coding is proposed. A comparison is made against various baselines, including FLAC. The proposed methods show improved compression ratios for the storage and transmission of immersive audio.
翻译:本文研究了改进多通道无损编码技术。针对同一内容的两种或多种不同渲染(混音)的同步编码,提出了一种方法。该信号模型利用上混的过去样本与下混的当前时间样本共同预测上混结果。模型参数通过通用线性求解器优化,预测残差采用Rice编码。此外,在残差编码前引入了SVD投影预处理。通过与包括FLAC在内的多种基准方法比较,所提方法在沉浸式音频的存储与传输中展现出更优的压缩比。