B.TechSemester 82022-23Big DataKOE-097

Big Data (KOE-097) - AKTU Question Paper 2022-23

B.Tech · Semester 8 · Free PDF Download

This is the official AKTU Big Data Previous Year Question Paper for B.Tech Semester 8, academic session 2022-23. Published by Dr. A.P.J. Abdul Kalam Technical University (AKTU/UPTU), Lucknow. Free PDF download — no login required.

Course:B.Tech
Semester:Semester 8
Session:2022-23
University:AKTU / UPTU

Rate this paper

Questions Asked in 2022-23

Big Data (KOE-097) — complete question paper

Section AAttempt all questions in brief. 2 x 10 = 20
  • a
    Explain benefits of HDFS over NFS
  • b
    Differentiate between structured, semi-structured and unstructured data
  • c
    Explain sources of data in big data
  • d
    Define Metadata in HDFS
  • e
    Differentiate between Map & Reduce
  • g
    Differentiate between shuffle and sort operation
  • h
    Explain TF-IDF. (i) Explain name node, data node, job tracker and task tracker . (j) Define file name and block size for Windows, Linux and Hadoop
Section BAttempt any three of the following: 10x3=30
  • a
    Explain the 5 Vs of Big Data. Also discuss their importance in the context of Big Data?
  • b
    Illustrate the history of Hadoop and its evolution over time into the Apache Hadoop platform that is widely used today?
  • c
    Explain the concept of data replication in HDFS and its benefits and challenges
  • d
    Compare and contrast the fair and capacity schedule rs used in the Hadoop YARN framework
  • e
    Explain Pig and its execution modes. Compare pi g with databases
Section CAttempt any one part of the following: 10 x1=10
  • a
    Discuss the role of security, compliance, auditing and protection in Big Data. Also discuss are the key features of Big Data in terms of security and privacy
  • b
    Explain the challenges of conventional data systems . Discuss the process of providing a solution to these challenges by Big Data?
  • a
    Discuss the Hadoop Distributed File System, and dis cuss its role to allow for the storage and processing of large data sets acros s distributed computing clusters
  • b
    What is the anatomy of a Map Reduce job run?
  • a
    Illustrate the data flows and data ingest meth ods in Hadoop, including Flume and Scoop?
  • b
    Discuss the support provided by Hadoop for compression, serialization, Avro, and file-based data structures in Hadoop I/O
  • b
    Discuss SCALA and its basic features, including cla sses and objects, basic types and operators, control structures, functions, and closures?
  • a
    Illustrate HBase and how its differ ence with RDBMS? Discuss the advantages of HBase's advanced indexing and schema design?
  • b
    Discuss the ro le of Zoo Keeper in monitoring a cluster. Also disc uss the process of building applications with Zoo Keeper

Question text is extracted from the official AKTU question paper PDF above. Hindi translations are omitted — every question is printed in English in the original paper. Last verified: 2026-08-23.

Repeated Questions — KOE-097

Questions that appeared in more than one session, found by comparing 4 years of Big Data papers (2021-22, 2022-23, 2023-24, 2024-25)

4x

Differentiate between structured, semi-structured and unstructured data

Appeared in: 2021-22 · 2022-23 · 2023-24 · 2024-25

3x

What is the anatomy of a Map Reduce job run?

Appeared in: 2021-22 · 2022-23 · 2023-24

2x

Explain sources of data in big data

Appeared in: 2022-23 · 2024-25

2x

Define Metadata in HDFS

Appeared in: 2022-23 · 2023-24

2x

Differentiate between Map & Reduce

Appeared in: 2022-23 · 2024-25

2x

Differentiate between shuffle and sort operation

Appeared in: 2022-23 · 2024-25

Big Data — Other Year Papers

AKTU Big Data PYQs from other sessions

Syllabus & More PYQs

Paper solve karne se pehle unit-wise syllabus dekh lo