AMBIQUAL: Towards a quality metric for headphone rendered compressed ambisonic spatial audio

Miroslaw Narbutt, Jan Skoglund, Andrew Allen, Michael Chinen, Dan Barry, Andrew Hines

Research output: Contribution to journalArticlepeer-review

Abstract

Spatial audio is essential tor creating a sense of immersion in virtual environments. Efficient encoding methods are required to deliver spatial audio over networks without compromising Quality of Service (QoS). Streaming service providers such as YouTube typically transcode content into various bit rates and need a perceptually relevant audio quality metric to monitor users' perceived quality and spatial localization accuracy. The aim of the paper is two-fold. First, it is to investigate the effect of Opus codec compression on the quality of spatial audio as perceived by listeners using subjective listening tests. Secondly it is to introduce AMBIQUAL, a full reference objective metric for spatial audio quality, which derives both listening quality and localization accuracy metrics directly from the B-format Ambisonic audio. We compare AMBIQUAL quality predictions with subjective quality assessments across a variety of audio samples which have been compressed using the Opus 1.2 codec at various bit rates. Listening quality and localization accuracy of first and third-order Ambisonics were evaluated. Several fixed and dynamic audio sources (single and multiple) were used to evaluate localization accuracy. Results show good correlation regarding listening quality and localization accuracy between objective quality scores using AMBIQUAL and subjective scores obtained during listening tests.

Original languageEnglish
Article number3188
JournalApplied Sciences (Switzerland)
Volume10
Issue number9
DOIs
Publication statusPublished - 1 May 2020

Keywords

  • Ambisonics
  • Audio coding
  • Audio compression
  • Audio qualitv
  • MUSHRA
  • Opus codec
  • QoE
  • Spatial audio
  • Virtual reality

Fingerprint

Dive into the research topics of 'AMBIQUAL: Towards a quality metric for headphone rendered compressed ambisonic spatial audio'. Together they form a unique fingerprint.

Cite this