Question 1
What are the challenges of the conventional data management systems when handling Big Data?
Explain the anatomy of a MapReduce job run, including failures and task execution.
What are the challenges of the conventional data management systems when handling Big Data?
Explain the anatomy of a MapReduce job run, including failures and task execution.
Describe the steps involved in setting up a Hadoop cluster with basic commands.
Write a Spark program to implement the PageRank algorithm using PySpark or Scala.
Write a MapReduce program to find the highest and lowest temperatures from a dataset containing daily temperature recordings.
Write a Python program using the MongoDB driver (PyMongo) to perform CRUD (Create, Read, Update, Delete) operations on a MongoDB collection storing user profiles.
What is the difference between master-slave and peer-to-peer NoSQL distribution models?
What is the importance of sampling data in a stream? Provide an example of its application.
Explain the importance of Stream Computing and give examples of its real- time applications, such as stock market predictions.
What is Spark's role in Big Data analytics? Discuss its advantages over traditional data processing systems.
customers and orders, and find all the orders placed by each customer.
Provide examples of real-world applications.
Circulate this solved paper with KaTeX formulas and 1-click AI step solvers to your batchmates on WhatsApp or Telegram.
Official Gujarat Technological University (GTU) examination paper and step-by-step solutions for Big Data Analytics (Summer 2025, B.E. · IT Engineering, Sem 6). Features complete 70-mark regular & remedial examination pattern, official marking distribution across all 5 questions, and direct 1-click official PDF download.
Transcribed for student exam preparation from Gujarat Technological University official examination archives. Questions, syllabus guidelines, and curriculum marking schemes remain the intellectual property of Gujarat Technological University.