Commune · Histoire
Speech coding
Type de compression de données
Speech coding is an application of data compression to digital audio signals containing speech. Speech coding uses speech-specific parameter estimation using audio signal processing techniques to model the speech signal, combined with generic data compression algorithms to represent the resulting modeled parameters in a compact bitstream.
Sur Wikipédia
Texte en anglais Pas encore d'article dans ta langue : extrait en anglais.
Speech coding is an application of data compression to digital audio signals containing speech. Speech coding uses speech-specific parameter estimation using audio signal processing techniques to model the speech signal, combined with generic data compression algorithms to represent the resulting modeled parameters in a compact bitstream. Common applications of speech coding are mobile telephony and voice over IP (VoIP). The most widely used speech coding technique in mobile telephony is linear predictive coding (LPC), while the most widely used in VoIP applications are the LPC and modified discrete cosine transform (MDCT) techniques. The techniques employed in speech coding are similar to those used in audio data compression and audio coding, where appreciation of psychoacoustics is used to transmit only data that is relevant to the human auditory system. For example, in voiceband speech coding, only information in the frequency band 400 to 3500 Hz is transmitted, but the reconstructed signal retains adequate intelligibility. Speech coding differs from other forms of audio coding in that speech is a simpler signal than other audio signals, and statistical information is available about the properties of speech. As a result, some auditory information that is relevant in general audio coding can be unnecessary in the speech coding context. Speech coding stresses the preservation of intelligibility and pleasantness of speech while using a constrained amount of transmitted data. In addition, most speech applications require low coding delay, as latency interferes with speech interaction.
Texte : Wikipédia en anglais, CC BY-SA 4.0. · Image : Ramon Boixadera Planas (Public domain) ·
Cartes voisines
-
E★★★
Encodage-pourcent
Mécanisme de codage de l’information dans un Uniform Resource Identifier
-
M★
Microsoft Speech API
Application programming interface for Microsoft Windows
-
★★
audiométrie vocale
Measurement of the ability to hear speech under various conditions of intensity and noise interference using sound-field as well as earphones and bone oscillators
-
★★
Reconnaissance automatique de la parole
Sous-domaine interdisciplinaire de la linguistique informatique qui développe des méthodologies et des technologies permettant la reconnaissance et la traduction du langage parlé en texte par ordinateur
-
D★
Datagram Congestion Control Protocol
Protocole réseau
-
N★
Note d'opinion moyenne
-
★★★★
Unicode
Standard industriel permettant de coder, représenter et traiter de façon cohérente des textes exprimés dans la plupart des écritures du monde
-
★★
Système auditif
Système sensoriel permettant l'audition
-
★
Voyant Tools
Application d'analyse de données textuelles
-
★★
MPEG-2
-
★★★
Orthophonie
Discipline paramédicale qui a pour but de diagnostiquer, évaluer et aider les personnes ayant des troubles de la parole, de la communication, du langage, de la voix et de la déglutition
-
★
Channel code
Technical process
-
M★
MPEG-H 3D Audio
-
C★★
Communication verbale
Communication par des mots
-
★★
Vorbis
Algorithme libre et gratuit de compression et de décompression audio
-
★
Voix de poitrine
-
★★★
Free Lossless Audio Codec
Codec audio libre de compression audio sans perte
-
C★★★
Compression de données
Opération informatique