# Speech coding <!-- MICROSIMGEN:BEGIN v1.7 — generated by g08_place_microsims.py; three.js first (§15); do not hand-edit inside --> ## Microsims — p5.js ### Speech coding (p5.js) · `vocoder` <div class="microsim-player"> <iframe src="https://editor.p5js.org/sciencenibber/full/Gxz9oSmSV" width="100%" height="480" frameborder="0" loading="lazy" sandbox="allow-scripts allow-same-origin" title="Speech coding — p5.js microsim"></iframe> </div> *Compress speech by modeling how the voice is produced, squeezing a call into a few kilobits.* **Open in the editor:** [&#9654; fork this sketch](https://editor.p5js.org/sciencenibber/sketches/Gxz9oSmSV) · movement *V · Coding, compression & communication* · library `p5js` ### Related microsims Live sims on neighbouring articles — 6 of them inside this article's own Wikipedia link tree: - [[Companding]] *(in tree)* - [[Convolution]] *(in tree)* - [[Data_compression]] *(in tree)* - [[Delta_modulation]] *(in tree)* - [[Differential_pulse-code_modulation]] *(in tree)* - [[Digital_signal_processing]] *(in tree)* *Sim hosted off-article; the article owns the reference, not the runtime (WIKI_RULES §10.4). Placed by `g08_place_microsims.py`.* <!-- MICROSIMGEN:END --> ## Links (Wikipedia order) <!-- injected from _registry/childlinks/Speech_coding.json (2026-07-30T02:09:12Z) --> `842_(compression_algorithm)` · `A-law_algorithm` · `AAC-LD` · `ATRAC` · `Adaptive_Huffman_coding` · `Adaptive_Multi-Rate_audio_codec` · `Adaptive_coding` · `Adaptive_differential_pulse-code_modulation` · `Algebraic_code-excited_linear_prediction` · `Apple_Inc.` · `Arithmetic_coding` · `Asymmetric_numeral_systems` · `Audio_Engineering_Society` · `Audio_bit_depth` · `Audio_codec` · `Audio_signal_processing` · `Autoencoder` · `Average_bitrate` · `Bit_rate` · `Brotli` · `Burrows–Wheeler_transform` · `Byte-pair_encoding` · `Bzip2` · `CDMA2000` · `CELT` · `Canonical_Huffman_code` · `Chain_code` · `Chroma_subsampling` · `Code-excited_linear_prediction` · `Codec_2` · `Coding_tree_unit` · `Color_space` · [[Companding]] · `Compressed_data_structure` · `Compressed_suffix_array` · `Compression_artifact` · `Constant_bitrate` · `Context_mixing` · `Context_tree_weighting` · [[Convolution]] · [[Data_compression]] · `Data_compression_symmetry` · `Daubechies_wavelet` · `David_A._Huffman` · `Deblocking_filter` · `Deep_learning_speech_synthesis` · `Deflate` · `Delta_encoding` · [[Delta_modulation]] · `Dictionary_coder` · [[Differential_pulse-code_modulation]] · `Digital_audio` · `Digital_radio` · [[Digital_signal_processing]] · [[Discrete_cosine_transform]] · `Discrete_sine_transform` · [[Discrete_wavelet_transform]] · `Display_resolution` · `Dynamic_Markov_compression` · [[Dynamic_range]] · `Elias_gamma_coding` · `Embedded_zerotrees_of_wavelet_transforms` · `Enhanced_full_rate` · [[Entropy_(information_theory)]] · `Entropy_coding` · `Exponential-Golomb_coding` · `FM-index` · `FaceTime` · [[Fast_Fourier_transform]] · `Fibonacci_coding` · `Film_frame` · `Fourier_transform` · `Fractal_compression` · `Frame_rate` · `Free_software` · `Full_Rate` · `Fundamental_frequency` · `G.711` · `G.722` · `G.722.1` · `G.723.1` · `G.726` · `G.728` · `G.729` · `G.729.1` · `GSM` · `GitHub` · `Golomb_coding` · `Grammar-based_code` · `Half_Rate` · `Huffman_coding` · `Hutter_Prize` · [[Image_compression]] · `Image_resolution` · `Incremental_encoding` · [[Information_theory]] · `Intelligibility_(communication)` · `Interlaced_video` · `Kolmogorov_complexity` · `LHA_(file_format)` · `LZ4_(compression_algorithm)` · `LZ77_and_LZ78` · `LZFSE` · `LZMA` · `LZRW` · `LZWL` · `LZX` · `Lapped_transform` · `Latency_(audio)` · `Lempel–Ziv–Oberhumer` · `Lempel–Ziv–Stac` · `Lempel–Ziv–Storer–Szymanski` · `Lempel–Ziv–Welch` · `Levenshtein_coding` · `Line_spectral_pairs` · `Linear_prediction` · `Linear_predictive_coding` · `Log_area_ratio` · `Lossless_compression` · `Lossy_compression` · `Lyra_(codec)` · [[Machine_learning]] · `Macroblock` · `Mark_Adler` · `Mobile_telephony` · `Modified_Huffman_coding` · `Modified_discrete_cosine_transform` · `Motion_compensation` · `Motion_estimation` · `Move-to-front_transform` · `Mu-law_algorithm` · [[Nyquist–Shannon_sampling_theorem]] · `Opus_(audio_format)` · `PAQ` · `Peak_signal-to-noise_ratio` · `Phil_Katz` · `Pixel` · `PlayStation_4` · `PlayStation_Network` · `Prediction_by_partial_matching` · `Prefix_code` · `Psychoacoustics` · `Pyramid_(image_processing)` · `Quantization_(image_processing)` · [[Quantization_(signal_processing)]] · `Range_coding` · `Rate–distortion_theory` · `Re-Pair` · `Redundancy_(information_theory)` · `Run-length_encoding` · `SILK` · [[Sampling_(signal_processing)]] · `Satellite_phone` · `Secure_voice` · `Selectable_Mode_Vocoder` · `Sequitur_algorithm` · `Set_partitioning_in_hierarchical_trees` · `Shannon_coding` · `Shannon–Fano_coding` · `Shannon–Fano–Elias_coding` · `Silence_compression` · `Smallest_grammar_problem` · `Snappy_(compression)` · `Sound_quality` · `Speech` · `Speech_processing` · `Speech_synthesis` · `Speex` · `Standard_test_image` · `Sub-band_coding` · `Texture_compression` · `The_Register` · `Timeline_of_information_theory` · `Transform_coding` · `Tunstall_coding` · `Unary_coding` · `Unified_Speech_and_Audio_Coding` · `Universal_code_(data_compression)` · `Variable_bitrate` · `Vector_quantization` · `Video` · `Video_codec` · `Video_compression_picture_types` · `Video_quality` · `Vocoder` · `Voice_over_IP` · `Warped_linear_predictive_coding` · `Wavelet_transform` · [[Wayback_Machine]] · `WhatsApp` · `Zstd` > Signal Processing concept · part of the Signal Processing Portal · movement V · !74 記号 kigō.svg <!-- RENDER-THUMB:START --> !480 *Rendered from the live microsim (▶ motion).* <!-- RENDER-THUMB:END --> ## See it next [![Speech signal processing|200](Speech_signal_processing_thumb.png)](Speech_signal_processing) *→ Speech signal processing* <!-- VISUAL-LINK:END --> --- Back to Signal Processing Portal · the room · Semiotic gateway <!-- REAL-GENERATIVE-MEDIA:START --> ## What it is Speech coding is the compression of digitized speech signals using models tuned specifically to the properties of the human voice. ## How it works / why it matters Speech coders such as CELP and LPC-based vocoders represent the voice as a model of the vocal tract excited by a source signal, transmitting compact model parameters rather than raw waveform samples. This exploits speech's predictable structure to achieve very low bit rates, making it essential to mobile telephony, VoIP, and digital radio. ## Signs & universals Instantiates encoding, signal and tradeoff_balance. ## Related Related: Speech signal processing, Speech communication, Audio data compression, [[Data_compression]], Codec. <!-- VISUAL-LINK:START --> ## From the Real GENERATIVE library > Speech coding is an application of data compression to digital audio signals containing speech. Speech coding uses speech-specific parameter estimation using audio signal processing techniques to model the speech signal, combined with generic data compression algorithms to represent the resulting modeled parameters in a compact bitstream.[1] ([Wikipedia](https://en.wikipedia.org/wiki/Speech_coding)) <!-- REAL-GENERATIVE-MEDIA:END --> <!-- CRAFT-LINK:START g12 --> *Built to the [[WT!P5_js_Microsim_Master_Class|p5.js Master Class]].* <!-- CRAFT-LINK:END --> ## Wikipedia : Wikitube **Strict pair:** [Wikipedia](https://en.wikipedia.org/wiki/Speech_coding) : [Wikitube](https://en.wikitube.io/wiki/Speech_coding) ## Previous hub tags Tree parent: [[Information_theory]]. Legacy hubs: none. --- *Sources: 1 legacy note. Minted wave 1, 2026-07-30 (v1.6 order).*