Data-Intensive Text Processing with MapReduce (Xử lý văn bản chuyên sâu dữ liệu với MapReduce) - Jimmy Lin and Chris Dyer
- ページ数
- 175
- 形式
- サイズ
- 1.7 MB
- 言語
- EN · English
- 年
- 2010
- 閲覧数
- 0
- コメント
- 0
- Lượt tải
- 0
プレビューを生成中...
Cuốn sách này giới thiệu về xử lý văn bản chuyên sâu dữ liệu sử dụng MapReduce, bao gồm các khái niệm cơ bản, thiết kế thuật toán và ứng dụng trong truy xuất văn bản.
- ドキュメント名
- Data-Intensive Text Processing with MapReduce (Xử lý văn bản chuyên sâu dữ liệu với MapReduce) - Jimmy Lin and Chris Dyer
- 著者(ドキュメント内)
- Jimmy Lin and Chris Dyer
- 内容
- Tài liệu này cung cấp kiến thức nền tảng và chuyên sâu về xử lý dữ liệu văn bản quy mô lớn bằng MapReduce, bao gồm các nguyên lý hoạt động, thiết kế thuật toán và ứng dụng thực tế như xây dựng chỉ mục đảo ngược.
- 目次
- Contents
- Introduction
- Computing in the Clouds
- Big Ideas
- Why Is This Different?
- What This Book Is Not
- MapReduce Basics
- Functional Programming Roots
- Mappers and Reducers
- The Execution Framework
- Partitioners and Combiners
- The Distributed File System
- Hadoop Cluster Architecture
- Summary
- MapReduce Algorithm Design
- Local Aggregation
- Combiners and In-Mapper Combining
- Algorithmic Correctness with Local Aggregation
- Pairs and Stripes
- Computing Relative Frequencies
- Secondary Sorting
- Relational Joins
- Reduce-Side Join
- Map-Side Join
- Memory-Backed Join
- Summary
- Inverted Indexing for Text Retrieval
- Web Crawling
- Inverted Indexes
- Inverted Indexing: Baseline Implementation
- Inverted Indexing: Revised Implementation
- Index Compression
- ページ数
- 175 ページ
- アップロード者
- Uni24h
説明
Trích nội dung tài liệu
i Data-Intensive Text Processing with MapReduce Jimmy Lin and Chris Dyer University of Maryland, College Park Manuscript prepared April 11, 2010 This is the pre-production manuscript of a book in the Morgan & Claypool Synthesis Lectures on Human Language Technologies. Anticipated publication date is mid-2010. ii Contents Contents . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . ii 1 2 3 Introduction . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 1 1.1 Computing in the Clouds. . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . .6 1.2 Big Ideas. . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . .9 1.3 Why Is This Different? . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 15 1.4 What This Book Is Not . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 17 MapReduce Basics . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 18 2.1 Functional Programming Roots . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 20 2.2 Mappers and Reducers . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 22 2.3 The Execution Framework . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 26 2.4 Partitioners and Combiners . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 28 2.5 The Distributed File System . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 31 2.6 Hadoop Cluster Architecture . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . .
よくある質問
このドキュメントは無料ですか?
はい。「Data-Intensive Text Processing with MapReduce (Xử lý văn bản chuyên sâu dữ liệu với MapReduce) - Jimmy Lin and Chris Dyer」は無料です。ログインして「ダウンロード」をクリックするだけで、元のファイルを取得できます。
このドキュメントは何ページありますか?
このドキュメントは 175 ページあります。ダウンロードする前にオンラインでプレビューできます。
ダウンロードする前にプレビューできますか?
はい。このページにあるオンラインリーダーでドキュメントをプレビューし、その後ダウンロードするかどうかを決めることができます。
Data-Intensive Text Processing with MapReduce (Xử lý văn bản chuyên sâu dữ liệu với MapReduce) - Jimmy Lin and Chris Dyer
プレビューを生成中...
Trích nội dung tài liệu
i Data-Intensive Text Processing with MapReduce Jimmy Lin and Chris Dyer University of Maryland, College Park Manuscript prepared April 11, 2010 This is the pre-production manuscript of a book in the Morgan & Claypool Synthesis Lectures on Human Language Technologies. Anticipated publication date is mid-2010. ii Contents Contents . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . ii 1 2 3 Introduction . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 1 1.1 Computing in the Clouds. . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . .6 1.2 Big Ideas. . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . .9 1.3 Why Is This Different? . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 15 1.4 What This Book Is Not . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 17 MapReduce Basics . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 18 2.1 Functional Programming Roots . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 20 2.2 Mappers and Reducers . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 22 2.3 The Execution Framework . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 26 2.4 Partitioners and Combiners . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 28 2.5 The Distributed File System . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 31 2.6 Hadoop Cluster Architecture . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . .
- ドキュメント名
- Data-Intensive Text Processing with MapReduce (Xử lý văn bản chuyên sâu dữ liệu với MapReduce) - Jimmy Lin and Chris Dyer
- 著者(ドキュメント内)
- Jimmy Lin and Chris Dyer
- 内容
- Tài liệu này cung cấp kiến thức nền tảng và chuyên sâu về xử lý dữ liệu văn bản quy mô lớn bằng MapReduce, bao gồm các nguyên lý hoạt động, thiết kế thuật toán và ứng dụng thực tế như xây dựng chỉ mục đảo ngược.
- 目次
- Contents
- Introduction
- Computing in the Clouds
- Big Ideas
- Why Is This Different?
- What This Book Is Not
- MapReduce Basics
- Functional Programming Roots
- Mappers and Reducers
- The Execution Framework
- Partitioners and Combiners
- The Distributed File System
- Hadoop Cluster Architecture
- Summary
- MapReduce Algorithm Design
- Local Aggregation
- Combiners and In-Mapper Combining
- Algorithmic Correctness with Local Aggregation
- Pairs and Stripes
- Computing Relative Frequencies
- Secondary Sorting
- Relational Joins
- Reduce-Side Join
- Map-Side Join
- Memory-Backed Join
- Summary
- Inverted Indexing for Text Retrieval
- Web Crawling
- Inverted Indexes
- Inverted Indexing: Baseline Implementation
- Inverted Indexing: Revised Implementation
- Index Compression
- ページ数
- 175 ページ
- アップロード者
- Uni24h
コメント (0)
まだコメントはありません。最初のコメントを書きましょう!
Ngân hàng đề thi môn: Hệ thống thông tin quản lý
Đề thi môn Cơ sở dữ liệu (kèm Đáp án) - Đại học Sư phạm kỹ thuật
Đề thi và đáp án môn Hệ thống thông tin kế toán
Đề thi và đáp án môn Cấu trúc dữ liệu giải thuật
Đáp án đề thi môn Mạng máy tính - ĐH Công nghệ thông tin (CNTT)
Tổng hợp Đề Toán 5 - Luyện thi vào Lớp 6 - CLB EMath
Bài giảng vật lý đại cương (Chương 3) - Đỗ Ngọc Uấn
Chương 8.Nguyên tử - Vật lý đại cương 3 - TS.Nguyễn Thị Trang
Chương 7.Cơ học lượng tử - Vật lý đại cương 3 - TS.Nguyễn Thị Trang
Chương 6.Quang học lượng tử - Vật lý đại cương 3 - TS.Nguyễn Thị Trang

コメント (0)
まだコメントはありません。最初のコメントを書きましょう!