The Enterprise Big Data Lake

Alex Gorelik

  • 出版商: O'Reilly
  • 出版日期: 2019-04-16
  • 定價: $2,600
  • 售價: 9.0$2,340
  • 語言: 英文
  • 頁數: 200
  • 裝訂: Paperback
  • ISBN: 1491931558
  • ISBN-13: 9781491931554
  • 相關分類: Hadoop大數據 Big-dataData Science
  • 立即出貨 (庫存=1)

買這商品的人也買了...

商品描述

Enterprises are experimenting with using Hadoop to build Big Data Lakes, but many projects are stalling or failing because the approaches that worked at Internet companies have to be adopted for the enterprise. This practical handbook guides managers and IT professionals from the initial research and decision-making process through planning, choosing products, and implementing, maintaining, and governing the modern data lake.

You'll explore various approaches to starting and growing a Data Lake, including Data Warehouse off-loading, analytical sandboxes, and "Data Puddles." Author Alex Gorelik shows you methods for setting up different tiers of data, from raw untreated landing areas to carefully managed and summarized data. You'll learn how to enable self-service to help users find, understand, and provision data; how to provide different interfaces to users with different skill levels; and how to do all of that in compliance with enterprise data governance policies.

商品描述(中文翻譯)

企業正在嘗試使用Hadoop建立大數據湖,但許多項目因為在企業中必須採用適合的方法而陷入停滯或失敗。這本實用手冊將引導管理人員和IT專業人員從最初的研究和決策過程,到規劃、選擇產品,以及實施、維護和管理現代數據湖。

您將探索各種啟動和發展數據湖的方法,包括數據倉庫卸載、分析沙盒和“數據水坑”。作者Alex Gorelik向您展示了建立不同數據層的方法,從原始未處理的登陸區域到精心管理和總結的數據。您將學習如何實現自助服務,幫助用戶找到、理解和提供數據;如何為不同技能水平的用戶提供不同的界面;以及如何在符合企業數據治理政策的情況下完成所有這些工作。