Slide Lập trình C nâng cao - Fit Lec 11 (HUST) GV.AnhTT
Génération de l'aperçu...
Slide bài giảng về nén dữ liệu (Data Compression) với trọng tâm là mã Huffman, bao gồm các khái niệm về mã độ dài biến (variable-length encoding) và thuật toán xây dựng cây Huffman để nén văn bản một cách hiệu quả.
Description
Data compression anhtt-fit@mail.hut.edu.vn dungct@it-hut.edu.vn Data Compression Data in memory have used fixed length for representation For data transfer (in particular), this method is inefficient. For speed and storage efficiencies, data symbols should use the minimum number of bits possible for representation. Methods Used For Compression: Encode high probability symbols with fewer bits Shannon-Fano, Huffman, UNIX compact Encode sequences of symbols with location of sequence in a dictionary PKZIP, ARC, GIF, UNIX compress, V.42bis Lossy compression JPEG and MPEG Variable Length Bit Codings Suppose ‘A’ appears 50 times in text, but ‘B’ appears only 10 times ASCII coding assigns 8 bits per character, so total bits for ‘A’ and ‘B’ is 60 * 8 = 480 If ‘A’ gets a 4-bit code and ‘B’ gets a 12-bit code, total is 50 * 4 + 10 * 12 = 320 Compression rules: Use minimum number of bits No code is the prefix of another code Enables left-to-right, unambiguous decoding Variable Length Bit Codings No code is a prefix of another For example, can’t have ‘A’ map to 10 and ‘B’ map to 100, because 10 is a prefix (the start of) 100. Enables left-to-right, unambiguous decoding That is, if you see 10, you know it’s ‘A’, not the start of another character. Variable-length encoding Use different number of bits to encode different characters. Ex. Morse code. Issue: ambiguity. SOS ? IAMIE ? EEWNI ? V7O ? Huffman code Constructed by using a code tree, but starting at the leave
Résumé IA
- Nom du document
- Slide Lập trình C nâng cao - Fit Lec 11 (HUST) GV.AnhTT
- École / Cours
- Đại học Bách khoa Hà Nội · Lập trình C
- Auteur (dans le document)
- anhtt-fit@mail.hut.edu.vn, dungct@it-hut.edu.vn
- Contenu
- Tài liệu tập trung vào nén dữ liệu bằng cách sử dụng mã hóa độ dài biến đổi, đặc biệt là thuật toán Huffman. Nó giải thích nguyên tắc hoạt động và cách xây dựng cây mã Huffman để tối ưu hóa việc biểu diễn ký hiệu.
- Table des matières
- Ce document n'a pas de table des matières claire.
- Pages
- 22 pages
- Téléversé par
- lienhejb
Foire aux questions
Ce document est-il gratuit ?
Oui. « Slide Lập trình C nâng cao - Fit Lec 11 (HUST) GV.AnhTT » est gratuit — il suffit de vous connecter et de cliquer sur Télécharger pour obtenir le fichier original.
Combien de pages compte ce document ?
Le document contient 22 pages, pour le cours Lập trình C. Vous pouvez le prévisualiser en ligne avant de le télécharger.
Puis-je prévisualiser avant de télécharger ?
Oui. Vous pouvez prévisualiser ce document directement sur cette page avec le lecteur en ligne, puis décider de le télécharger ou non.
Slide Lập trình C nâng cao - Fit Lec 11 (HUST) GV.AnhTT
Génération de l'aperçu...
Data compression anhtt-fit@mail.hut.edu.vn dungct@it-hut.edu.vn Data Compression Data in memory have used fixed length for representation For data transfer (in particular), this method is inefficient. For speed and storage efficiencies, data symbols should use the minimum number of bits possible for representation. Methods Used For Compression: Encode high probability symbols with fewer bits Shannon-Fano, Huffman, UNIX compact Encode sequences of symbols with location of sequence in a dictionary PKZIP, ARC, GIF, UNIX compress, V.42bis Lossy compression JPEG and MPEG Variable Length Bit Codings Suppose ‘A’ appears 50 times in text, but ‘B’ appears only 10 times ASCII coding assigns 8 bits per character, so total bits for ‘A’ and ‘B’ is 60 * 8 = 480 If ‘A’ gets a 4-bit code and ‘B’ gets a 12-bit code, total is 50 * 4 + 10 * 12 = 320 Compression rules: Use minimum number of bits No code is the prefix of another code Enables left-to-right, unambiguous decoding Variable Length Bit Codings No code is a prefix of another For example, can’t have ‘A’ map to 10 and ‘B’ map to 100, because 10 is a prefix (the start of) 100. Enables left-to-right, unambiguous decoding That is, if you see 10, you know it’s ‘A’, not the start of another character. Variable-length encoding Use different number of bits to encode different characters. Ex. Morse code. Issue: ambiguity. SOS ? IAMIE ? EEWNI ? V7O ? Huffman code Constructed by using a code tree, but starting at the leave
Lire le document entier
- Nom du document
- Slide Lập trình C nâng cao - Fit Lec 11 (HUST) GV.AnhTT
- École / Cours
- Đại học Bách khoa Hà Nội · Lập trình C
- Auteur (dans le document)
- anhtt-fit@mail.hut.edu.vn, dungct@it-hut.edu.vn
- Contenu
- Tài liệu tập trung vào nén dữ liệu bằng cách sử dụng mã hóa độ dài biến đổi, đặc biệt là thuật toán Huffman. Nó giải thích nguyên tắc hoạt động và cách xây dựng cây mã Huffman để tối ưu hóa việc biểu diễn ký hiệu.
- Table des matières
- Ce document n'a pas de table des matières claire.
- Pages
- 22 pages
- Téléversé par
- lienhejb
Commentaires (0)
Aucun commentaire pour le moment. Soyez le premier !
Lập trình C nâng cao - Fit Lec 7 (HUST) GV.AnhTT
Lập trình C nâng cao - Fit Lec 8 (HUST) GV.AnhTT
Lập trình C nâng cao - Fit Lec 4 (HUST) GV.AnhTT
Lập trình C nâng cao - Fit Lec 5 (HUST) GV.AnhTT
Lập trình C nâng cao - Fit Lec 3 (HUST) GV.AnhTT
K5 Bộ đề luyện thi Trạng Nguyên Tiếng Việt (NXB DHQG)
K2 Bộ đề luyện thi Trạng Nguyên Tiếng Việt (NXB DHQG)
K3 Bộ đề luyện thi Trạng Nguyên Tiếng Việt (NXB DHQG)
K4 Bộ đề luyện thi Trạng Nguyên Tiếng Việt (NXB DHQG)
K1 Bộ đề luyện thi Trạng Nguyên Tiếng Việt (NXB DHQG)
Commentaires (0)
Aucun commentaire pour le moment. Soyez le premier !