- Home Page /
- Books /
- Computers & Technology /
- Databases & Big Data /
- Data Modeling & Design /
- Data Engineering with Scala and Spark: Build ...
Data Engineering with Scala and Spark: Build streaming and batch pipelines that process massive amounts of data using Scala
ISK 7886
Price Details
Excluding Shipping & Custom charges ( Shipping and custom charges will be calculated on checkout )
*All items will import from US
QTY:
Ubuy works hard to protect your security and privacy. Our advanced payment security system ensures confidentiality by encrypting your information during transmission using AES (Advanced Encryption Standards) and SSL (Secure Socket Layer) protocols. Your payment details are 100% secure as we do not share your payment details with third party sellers.
Take your data engineering skills to the next level by learning how to utilize Scala and functional programming to create continuous and scheduled pipelines.
Fast
Shipping
Free
Return*
Secure Packaging
100% Original Products
PCI DSS Compliance
ISO 27001 Certified
Product Details
- Take your data engineering skills to the next level by learning how to utilize Scala and functional programming to create continuous and scheduled pipelines that ingest, transform, and aggregate dataKey FeaturesTransform data into a clean and trusted source of information for your organization using ScalaBuild streaming and batch-processing pipelines with step-by-step explanationsImplement and orchestrate your pipelines by following CI/CD best practices and test-driven development (TDD)Purchase of the print or Kindle book includes a free PDF eBookBook DescriptionMost data engineers know that performance issues in a distributed computing environment can easily lead to issues impacting the overall efficiency and effectiveness of data engineering tasks. While Python remains a popular choice for data engineering due to its ease of use, Scala shines in scenarios where the performance of distributed data processing is paramount. This book will teach you how to leverage the Scala programming language on the Spark framework and use the latest cloud technologies to build continuous and triggered data pipelines. You’ll do this by setting up a data engineering environment for local development and scalable distributed cloud deployments using data engineering best practices, test-driven development, and CI/CD. You’ll also get to grips with DataFrame API, Dataset API, and Spark SQL API and its use. Data profiling and quality in Scala will also be covered, alongside techniques for orchestrating and performance tuning your end-to-end pipelines to deliver data to your end users. By the end of this book, you will be able to build streaming and batch data pipelines using Scala while following software engineering best practices.What you will learnSet up your development environment to build pipelines in ScalaGet to grips with polymorphic functions, type parameterization, and Scala implicitsUse Spark DataFrames, Datasets, and Spark SQL with ScalaRead and write data to object storesProfile and clean your data using DeequPerformance tune your data pipelines using ScalaWho this book is forThis book is for data engineers who have experience in working with data and want to understand how to transform raw data into a clean, trusted, and valuable source of information for their organization using Scala and the latest cloud technologies. Table of ContentsScala Essentials for Data EngineersEnvironment SetupAn Introduction to Apache Spark and Its APIs – DataFrame, Dataset, and Spark SQLWorking with DatabasesObject Stores and Data LakesUnderstanding Data TransformationData Profiling and Data QualityTest-Driven Development, Code Health, and MaintainabilityCI/CD with GitHubData Pipeline OrchestrationPerformance TuningBuilding Batch Pipelines Using Spark and ScalaBuilding Streaming Pipelines Using Spark and Scala
| Publisher | Packt Publishing |
| Publication date | January 31, 2024 |
| Language | English |
| Print length | 300 pages |
| ISBN-10 | 1804612588 |
| ISBN-13 | 978-1804612583 |
| Item Weight | 1.14 pounds (520 grams) |
| Dimensions | 7.5 x 0.68 x 9.25 inches (19.1 x 1.7 x 23.5 cm) |
Product Description
Data Engineering with Scala and Spark: Build streaming and batch pipelines that process massive amounts of data using Scala
Customer Questions & Answers
-
Question:
Who is this book intended for?
Answer: This book is for data engineers with experience in handling data, aiming to learn Scala for data transformation. -
Question:
What technologies does this book cover?
Answer: It covers Scala, Apache Spark, cloud technologies, and data engineering best practices. -
Question:
What will I achieve after reading this book?
Answer: You will be able to build efficient streaming and batch data pipelines while following software engineering best practices.
Data Modeling & Design Editorial Review
Customer Reviews & Ratings
-
5 Star
36%
-
4 Star
27%
-
3 Star
37%
-
2 Star
0%
-
1 Star
0%
Review this product
Share your thoughts with other customers
Platform Trust & Buyer Confidence
“Excellent quality and original too,when ubuy send original things I will appreciate that ,I very satisfied thank you”
“The order and delivery progress was communicated very well. Package arrived within the estimated time. Products arrived as expected in good condition.”
“Good supply of products. Safe payment methods, and shipment worldwide! Genuine products.”
“Was my first time buying a product from Ubuy, but I found it so helpful. This is reliable. Gonna place a new order!”
“Ubuy is a great online platform to buy stuff. Reliable, fast delivery time and great value for money. I've been using Ubuy online shopping platform for three years now and will continue to do so.”
Product Price History
Important information
- Limitations : For products shipped internationally, please note that any manufacturer warranty may not be valid; manufacturer service options may not be available; product manuals, instructions, and safety warnings may not be in destination country languages; the products (and accompanying materials) may not be designed in accordance with destination country standards, specifications, and labeling requirements; and the products may not conform to destination country voltage and other electrical standards (requiring use of an adapter or converter if appropriate). The recipient is responsible for assuring that the product can be lawfully imported to the destination country. When ordering from Ubuy or its affiliates, the recipient is the importer of record and must comply with all laws and regulations of the destination country.
- Not all the products listed on Ubuy are for sale, as Ubuy is a global search engine. Products are subject to export/trade regulations.
ISK 7886
Order now and get it around Tuesday, September 08
This item is not restrict in my country.(Please click on above link if this item is not restrict in your country, So our team will review and allow.)
QTY:
PCI DSS compliant and ISO 27001:2022 certified, with encrypted payments and full buyer protection on every order.
Features & Benefits
- Enhance your data engineering skills with Scala.
- Learn to build efficient streaming and batch-processing pipelines.
- Implement CI/CD best practices and test-driven development.
- Develop a data engineering environment for both local and cloud deployments.
- Master Spark APIs: DataFrames, Datasets, and Spark SQL.
- Transform raw data into clean, trusted information.
Ubuy Assurance
Experience worry-free shopping with 100% original products, PCI DSS-compliant payment security, ISO 27001-certified data protection, the fastest cross-border delivery, free returns *, and secure packaging on every order.
