Slide Lập trình C nâng cao - Fit Lec 11 (HUST) GV.AnhTT
正在生成预览...
Slide bài giảng về nén dữ liệu (Data Compression) với trọng tâm là mã Huffman, bao gồm các khái niệm về mã độ dài biến (variable-length encoding) và thuật toán xây dựng cây Huffman để nén văn bản một cách hiệu quả.
描述
Data compression anhtt-fit@mail.hut.edu.vn dungct@it-hut.edu.vn Data Compression Data in memory have used fixed length for representation For data transfer (in particular), this method is inefficient. For speed and storage efficiencies, data symbols should use the minimum number of bits possible for representation. Methods Used For Compression: Encode high probability symbols with fewer bits Shannon-Fano, Huffman, UNIX compact Encode sequences of symbols with location of sequence in a dictionary PKZIP, ARC, GIF, UNIX compress, V.42bis Lossy compression JPEG and MPEG Variable Length Bit Codings Suppose ‘A’ appears 50 times in text, but ‘B’ appears only 10 times ASCII coding assigns 8 bits per character, so total bits for ‘A’ and ‘B’ is 60 * 8 = 480 If ‘A’ gets a 4-bit code and ‘B’ gets a 12-bit code, total is 50 * 4 + 10 * 12 = 320 Compression rules: Use minimum number of bits No code is the prefix of another code Enables left-to-right, unambiguous decoding Variable Length Bit Codings No code is a prefix of another For example, can’t have ‘A’ map to 10 and ‘B’ map to 100, because 10 is a prefix (the start of) 100. Enables left-to-right, unambiguous decoding That is, if you see 10, you know it’s ‘A’, not the start of another character. Variable-length encoding Use different number of bits to encode different characters. Ex. Morse code. Issue: ambiguity. SOS ? IAMIE ? EEWNI ? V7O ? Huffman code Constructed by using a code tree, but starting at the leave
AI 摘要
- 文档名称
- Slide Lập trình C nâng cao - Fit Lec 11 (HUST) GV.AnhTT
- 学校 / 课程
- Đại học Bách khoa Hà Nội · Lập trình C
- 作者(文档中)
- anhtt-fit@mail.hut.edu.vn, dungct@it-hut.edu.vn
- 内容
- Tài liệu tập trung vào nén dữ liệu bằng cách sử dụng mã hóa độ dài biến đổi, đặc biệt là thuật toán Huffman. Nó giải thích nguyên tắc hoạt động và cách xây dựng cây mã Huffman để tối ưu hóa việc biểu diễn ký hiệu.
- 目录
- 此文档没有清晰的目录。
- 页数
- 22 页
- 上传者
- lienhejb
常见问题
此文档免费吗?
是的。“Slide Lập trình C nâng cao - Fit Lec 11 (HUST) GV.AnhTT”是免费的 — 只需登录并点击“下载”即可获取原始文件。
这份文档有多少页?
该文档共有 22 页,适用于课程 Lập trình C。您可以在下载前进行在线预览。
我可以在下载前预览吗?
是的。您可以通过在线阅读器直接在本页面预览此文档,然后再决定是否下载。
Slide Lập trình C nâng cao - Fit Lec 11 (HUST) GV.AnhTT
正在生成预览...
Data compression anhtt-fit@mail.hut.edu.vn dungct@it-hut.edu.vn Data Compression Data in memory have used fixed length for representation For data transfer (in particular), this method is inefficient. For speed and storage efficiencies, data symbols should use the minimum number of bits possible for representation. Methods Used For Compression: Encode high probability symbols with fewer bits Shannon-Fano, Huffman, UNIX compact Encode sequences of symbols with location of sequence in a dictionary PKZIP, ARC, GIF, UNIX compress, V.42bis Lossy compression JPEG and MPEG Variable Length Bit Codings Suppose ‘A’ appears 50 times in text, but ‘B’ appears only 10 times ASCII coding assigns 8 bits per character, so total bits for ‘A’ and ‘B’ is 60 * 8 = 480 If ‘A’ gets a 4-bit code and ‘B’ gets a 12-bit code, total is 50 * 4 + 10 * 12 = 320 Compression rules: Use minimum number of bits No code is the prefix of another code Enables left-to-right, unambiguous decoding Variable Length Bit Codings No code is a prefix of another For example, can’t have ‘A’ map to 10 and ‘B’ map to 100, because 10 is a prefix (the start of) 100. Enables left-to-right, unambiguous decoding That is, if you see 10, you know it’s ‘A’, not the start of another character. Variable-length encoding Use different number of bits to encode different characters. Ex. Morse code. Issue: ambiguity. SOS ? IAMIE ? EEWNI ? V7O ? Huffman code Constructed by using a code tree, but starting at the leave
阅读全文
- 文档名称
- Slide Lập trình C nâng cao - Fit Lec 11 (HUST) GV.AnhTT
- 学校 / 课程
- Đại học Bách khoa Hà Nội · Lập trình C
- 作者(文档中)
- anhtt-fit@mail.hut.edu.vn, dungct@it-hut.edu.vn
- 内容
- Tài liệu tập trung vào nén dữ liệu bằng cách sử dụng mã hóa độ dài biến đổi, đặc biệt là thuật toán Huffman. Nó giải thích nguyên tắc hoạt động và cách xây dựng cây mã Huffman để tối ưu hóa việc biểu diễn ký hiệu.
- 目录
- 此文档没有清晰的目录。
- 页数
- 22 页
- 上传者
- lienhejb
评论 (0)
暂无评论。快来抢沙发吧!
K5 Bộ đề luyện thi Trạng Nguyên Tiếng Việt (NXB DHQG)
K2 Bộ đề luyện thi Trạng Nguyên Tiếng Việt (NXB DHQG)
K3 Bộ đề luyện thi Trạng Nguyên Tiếng Việt (NXB DHQG)
K4 Bộ đề luyện thi Trạng Nguyên Tiếng Việt (NXB DHQG)
K1 Bộ đề luyện thi Trạng Nguyên Tiếng Việt (NXB DHQG)
评论 (0)
暂无评论。快来抢沙发吧!