Gujarat Technological UniversitySummer 2025 Examination

GTU 3170722 Big Data Analytics (BDA) Summer 2025 Paper Solution & PDF

B.E. · Computer Engineering · Semester 7 · Subject Code: 3170722
Download Official GTU PDF
Share:
Total Marks70 MarksExternal theory exam
Passing Marks23 Marks33% minimum cutoff
Exam Duration2.5 Hours2:30 PM – 5:00 PM
Paper Structure5 QuestionsWith internal OR choices
Jump toQ1Q2Q3Q4Q5

Question 1

14 MarksMedium
(a)
What is Big Data? Explain the five V’s of Big Data.
3 Marks
(b)

Discuss the key differences between structured, unstructured, and semi-structured data with examples.

4 Marks
(c)

Explain Hadoop's architecture in detail with a diagram. How does the NameNode and DataNode work together in Hadoop?

7 Marks

Question 2

14 MarksMedium
(a)

What is a MapReduce function? Explain the purpose of the Map and Reduce phases.

3 Marks
(b)

Discuss the concept of Data Locality in Hadoop. Why is it significant?

4 Marks
(c)

Write a detailed note on HDFS (Hadoop Distributed File System) and explain its significance in Big Data processing.

7 Marks
OR OPTION
(c)

Discuss how MapReduce handles failures during the execution of jobs.

7 Marks

Question 3

14 MarksMedium
(a)

Define NoSQL databases and explain how they differ from relational databases.

3 Marks
(b)
Explain the Document store model in NoSQL databases
4 Marks
(c)

Compare and contrast different types of NoSQL databases: Key- Value, Document, Column-family, and Graph databases.

7 Marks
OR OPTION
(a)

Describe how NoSQL databases handle scalability and high availability.

3 Marks
(b)

What is Sharding in NoSQL databases? How does it help in handling Big Data?

4 Marks
(c)

Describe the architecture of MongoDB. Discuss its data model and how it handles queries and indexing.

7 Marks

Question 4

14 MarksMedium
(a)

Explain the role of Apache Spark in Big Data Analytics. How is it different from Hadoop MapReduce?

3 Marks
(b)

Discuss Spark’s RDD (Resilient Distributed Dataset) and its features.

4 Marks
(c)

Write and explain a simple Spark application for word count using PySpark

7 Marks
OR OPTION
(a)

What is in-memory processing? Why is it beneficial in Big Data analytics?

3 Marks
(b)

Explain Spark’s transformation and action operations with examples.

4 Marks
(c)

How is Spark used to perform Machine Learning tasks? Discuss Spark MLlib with an example of a classification algorithm.

7 Marks

Question 5

14 MarksMedium
(a)

What is Mahout? Explain its use in scalable machine learning with Big Data.

3 Marks
(b)

Explain the concept of data streaming. How does Spark Streaming process real-time data?

4 Marks
(c)

Explain how data mining techniques are applied in Big Data Analytics. Mention some of the challenges faced when mining large datasets.

7 Marks
OR OPTION
(a)

Describe the use of Big Data in retail industries. How can companies benefit from Big Data Analytics in decision-making?

3 Marks
(b)
What are the security challenges associated with Big Data?
4 Marks
(c)

Explain data warehousing in the context of Big Data. How does Hive enable querying large datasets?

7 Marks
College Exam Groups

Studying for Big Data Analytics?

Circulate this solved paper with KaTeX formulas and 1-click AI step solvers to your batchmates on WhatsApp or Telegram.

About this Examination Paper & Attribution

Official Gujarat Technological University (GTU) examination paper and step-by-step solutions for Big Data Analytics (BDA) (Summer 2025, B.E. · Computer Engineering, Sem 7). Features complete 70-mark regular & remedial examination pattern, official marking distribution across all 5 questions, and direct 1-click official PDF download.

Transcribed for student exam preparation from Gujarat Technological University official examination archives. Questions, syllabus guidelines, and curriculum marking schemes remain the intellectual property of Gujarat Technological University.

Download PDF