Organizations these days have gravitated toward services such as AWS Glue that undertake undifferentiated heavy lifting and provide serverless Spark, enabling you to create and manage data lakes in a serverless fashion. This guide shows you how AWS Glue can be used to solve real-world problems along with helping you learn about data processing, data integration, and building data lakes.
Beginning with AWS Glue basics, this book teaches you how to perform various aspects of data analysis such as ad hoc queries, data visualization, and real-time analysis using this service. It also provides a walk-through of CI/CD for AWS Glue and how to shift left on quality using automated regression tests. You’ll find out how data security aspects such as access control, encryption, auditing, and networking are implemented, as well as getting to grips with useful techniques such as picking the right file format, compression, partitioning, and bucketing. As you advance, you’ll discover AWS Glue features such as crawlers, Lake Formation, governed tables, lineage, DataBrew, Glue Studio, and custom connectors. The concluding chapters help you to understand various performance tuning, troubleshooting, and monitoring options.
By the end of this AWS book, you’ll be able to create, manage, troubleshoot, and deploy ETL pipelines using AWS Glue.
Paga fácilmente con tarjeta, Klarna, Apple Pay o Google Pay. ¿No estás contento? Siempre tienes 14 días de garantía de devolución. Lee más en nuestros términos. Si tienes preguntas, escríbenos a hello@memmo.org.
Memmo hace que estudiar sea más fácil, estés donde estés. Aquí tienes tus libros de texto y herramientas de estudio inteligentes en un solo lugar: resúmenes, quizzes, podcasts y flashcards. Y también a Ted, tu compañero de estudio que responde a todo lo que te preguntes. Más de 50 000 estudiantes ya estudian aquí. Está hecho para que aprendas más rápido y te estreses menos.