The Enterprise Big Data Lake

时间:2022-05-07 03:19:56
【文件属性】:

文件名称:The Enterprise Big Data Lake

文件大小:10.52MB

文件格式:PDF

更新时间:2022-05-07 03:19:56

大数据 数据湖 数据科学

Author(s): Alex Gorelik Publisher: O’Reilly Media, Year: 2019 ISBN: 1491931558,9781491931554 Description: Enterprises are experimenting with using Hadoop to build Big Data Lakes, but many projects are stalling or failing because the approaches that worked at Internet companies have to be adopted for the enterprise. This practical handbook guides managers and IT professionals from the initial research and decision-making process through planning, choosing products, and implementing, maintaining, and governing the modern data lake. You'll explore various approaches to starting and growing a Data Lake, including Data Warehouse off-loading, analytical sandboxes, and "Data Puddles." Author Alex Gorelik shows you methods for setting up different tiers of data, from raw untreated landing areas to carefully managed and summarized data. You'll learn how to enable self-service to help users find, understand, and provision data; how to provide different interfaces to users with different skill levels; and how to do all of that in compliance with enterprise data governance policies.


网友评论