This page will soon be deactivated—explore our new, faster, mobile-friendly site, now centralized in MyWorkspace!

Connecting the world and beyond

  •  
PP 2026

ITU-T Recommendations

Search by number:
Others:
Skip Navigation Links
Content search
Advanced search
Provisional name
Equivalent number
Formal description
Study Groups tree viewExpand Study Groups tree view

ITU-T P.565.1 (11/2021)

عربي | 中文 | English | Español | Français | Русский
Machine learning model for the assessment of transmission network impact on speech quality for mobile packet-switched voice services
Recommendation ITU-T P.565.1 is based on the ITU-T P.565 framework. It provides a machine learning based model that predicts the impact on the speech quality from the Internet Protocol (IP) transport and underlying transport, as well as a standardized or pre-defined jitter buffer in the end client; thus, providing a network centric view on the speech quality service delivered on mobile packet switched networks. This is expressed in terms of a mean opinion score-listening quality objective (MOS-LQO) under the assumption of an otherwise clean transmission, without background noise, non-standard-conformant encoding on sending device, automatic gain control, voice enhancement devices, transcoding, bridging, frequency response, non-standard-conformant jitter-buffer (for IMS mobile calls) or decoding, clock drift or any other impairment not caused by the IP transport and underlying transport.
The model supports the uses cases and applications defined in revised ITU-T P.565 for IMS mobile calls (VoLTE/VoNR with EVS, AMRWB codecs) and OTT/WhatsApp. In addition, it meets the minimum performance requirements for the provided test vectors (see ITU-T P.565, Annex D) and it also passed an independent validation on an additional unknown live recorded data set (see ITU-T P.565, Annex D).
The model enables the assessment of transmission network impact on speech quality for mobile packet-switched voice services. In addition, if this predictor is used together with perceptual speech analysis or perceptual speech quality metrics like [ITU-T P.863], it is possible to identify if the source of problems resides inside or outside the transport network observed by the predictor.

Keywords
Speech Quality Prediction, HD voice, EVS, AMR/AMR-WB, FB, OTT voice, IMS mobile voice, VoLTE, VoNR, OTT, Machine Learning (ML), intrusive parametric, generic OTT client
Citation: https://handle.itu.int/11.1002/1000/14823
Series title: P series: Telephone transmission quality, telephone installations, local line networks
  P.500-P.599: Objective measuring apparatus
Approval date: 2021-11-29
Provisional name:P.VSQMTF-1
Approval process:AAP
Status: In force
Maintenance responsibility: ITU-T Study Group 12
Further details: Patent statement(s)
Development history