Big Data Analytics with Hadoop and Spark : A hands-on guide to big data engineering and scalable analytics

Description: Technologies like Hadoop and Spark, powered by the Cloudera platform, have become essential for storing, processing, and analyzing big data across various industries, including finance, healthcare, e-commerce, and research in today's data-driven world. This book systematically navi...

Ful tanımlama

Kaydedildi:
Detaylı Bibliyografya
Yazar: Mehta, Shikha
Materyal Türü: Livre numérique
Dil:Anglais
Baskı/Yayın Bilgisi: New Delhi : BPB Publications 2026.
Paris : Cyberlibris
Online Erişim:Accès Université d'Orléans et IFPM
Not: Couverture. https://static2.cyberlibris.com/books_upload/300pix/9789365894745.jpg
Cyberlibris (ScholarVox) corpus Informatique
Autres localisations: Voir dans le Sudoc
Edition sous un autre format:• Big Data Analytics with Hadoop and Spark, A hands-on guide to big data engineering and scalable analytics, Shikha Mehta, New Delhi, BPB Publications, 2026, 1 vol. (576 p.), 978-93-6589-474-5
LEADER 03978nam a22002177a 4500
001 1478014
008 260708s2026 xxg ||| |||| 00| 0 eng d
009 PPN298023482
041 0 |a eng 
100 1 |a Mehta, Shikha. 
245 1 0 |a Big Data Analytics with Hadoop and Spark :  |b A hands-on guide to big data engineering and scalable analytics   |c Shikha Mehta. 
260 |a New Delhi :  |b BPB Publications. 
260 |a Paris :  |b Cyberlibris,  |c 2026. 
500 |a Couverture. https://static2.cyberlibris.com/books_upload/300pix/9789365894745.jpg 
500 |a Cyberlibris (ScholarVox) corpus Informatique 
505 0 |a 1. Exploring Big Data -- 2. Introduction to Hadoop -- 3. Hadoop Distributed File System and MapReduce -- 4. Big Data Analysis with Cloudera -- 5. Stock Data Analysis with Cloudera -- 6. Understanding Pig for Big Data Processing -- 7. Operators in Pig Latin -- 8. Functions in Apache Pig -- 9. Hive-data Warehousing and SQL-like Queries -- 10. Data Analysis Using Hive -- 11. Data Storage and Processing Using HBase -- 12. MongoDB -- 13. Introduction to Spark for Big Data Processing -- 14. Getting Started with Scala Programming -- 15. Data Analysis with Spark SQL -- 16. Machine Learning Application Using PySpark 
506 |a L'accès en ligne est réservé aux établissements ou bibliothèques ayant souscrit l'abonnement. Cyberlibris 
520 |a Description: Technologies like Hadoop and Spark, powered by the Cloudera platform, have become essential for storing, processing, and analyzing big data across various industries, including finance, healthcare, e-commerce, and research in today's data-driven world. This book systematically navigates the entire ecosystem, starting with big data fundamentals, security, and HDFS architecture before mastering MapReduce through weather and stock data case studies. Readers will gain hands-on experience with the Cloudera framework, learning high-level scripting with Pig Latin and structured data warehousing using HiveQL's Metastore and partitions. Additionally, it explores NoSQL versatility with HBase and MongoDB's CAP theorem, followed by Scala programming and Spark's high-speed in-memory engine. You will learn to optimize queries with the Catalyst optimizer and process complex Parquet or JSON files using Spark SQL DataFrames. The book also covers machine learning pipelines with spark.ml for professional-grade classification and clustering applications. By the end of this book, readers will be able to develop strong conceptual clarity and practical expertise in big data analytics. This will enable them to confidently design, implement, and manage scalable data processing solutions, preparing them to solve real-world data challenges and take on professional roles in big data engineering and analytics. What you will learn: Understand big data concepts, architecture, ethics, and applications; Build scalable storage using HDFS and MapReduce; Perform data analysis using Pig and Hive; Develop NoSQL solutions using HBase and MongoDB; Process large datasets using Apache Spark; Analyze data using Spark SQL and DataFrames; Implement machine learning using PySpark. Who this book is for: This book is ideal for students, researchers, and academicians. It empowers aspiring big data engineers, data scientists, and software engineers. Readers should possess basic programming knowledge and database fundamentals to master Hadoop and Spark for professional-grade data science and faculty-level instruction 
776 0 |t Big Data Analytics with Hadoop and Spark  |o A hands-on guide to big data engineering and scalable analytics  |f Shikha Mehta  |c New Delhi  |n BPB Publications  |d 2026  |p 1 vol. (576 p.)  |z 978-93-6589-474-5 
856 4 |5 452349901:894031309  |u https://ezproxy.univ-orleans.fr/login?url=https://univ.scholarvox.com/book/88980637  |z Accès Université d'Orléans et IFPM 
997 |0 1478014  |1 Livre numérique  |a Ressource numérique  |b INSA  |b ENSA  |c 0/Bibliothèque numérique/  |c 1/Bibliothèque numérique/ScholarVox (ebooks)/