Big Data (KOE-097) - AKTU Question Paper 2022-23
B.Tech · Semester 8 · Free PDF Download
This is the official AKTU Big Data Previous Year Question Paper for B.Tech Semester 8, academic session 2022-23. Published by Dr. A.P.J. Abdul Kalam Technical University (AKTU/UPTU), Lucknow. Free PDF download — no login required.
Rate this paper
Questions Asked in 2022-23
Big Data (KOE-097) — complete question paper
- aExplain benefits of HDFS over NFS
- bDifferentiate between structured, semi-structured and unstructured data
- cExplain sources of data in big data
- dDefine Metadata in HDFS
- eDifferentiate between Map & Reduce
- gDifferentiate between shuffle and sort operation
- hExplain TF-IDF. (i) Explain name node, data node, job tracker and task tracker . (j) Define file name and block size for Windows, Linux and Hadoop
- aExplain the 5 Vs of Big Data. Also discuss their importance in the context of Big Data?
- bIllustrate the history of Hadoop and its evolution over time into the Apache Hadoop platform that is widely used today?
- cExplain the concept of data replication in HDFS and its benefits and challenges
- dCompare and contrast the fair and capacity schedule rs used in the Hadoop YARN framework
- eExplain Pig and its execution modes. Compare pi g with databases
- aDiscuss the role of security, compliance, auditing and protection in Big Data. Also discuss are the key features of Big Data in terms of security and privacy
- bExplain the challenges of conventional data systems . Discuss the process of providing a solution to these challenges by Big Data?
- aDiscuss the Hadoop Distributed File System, and dis cuss its role to allow for the storage and processing of large data sets acros s distributed computing clusters
- bWhat is the anatomy of a Map Reduce job run?
- aIllustrate the data flows and data ingest meth ods in Hadoop, including Flume and Scoop?
- bDiscuss the support provided by Hadoop for compression, serialization, Avro, and file-based data structures in Hadoop I/O
- bDiscuss SCALA and its basic features, including cla sses and objects, basic types and operators, control structures, functions, and closures?
- aIllustrate HBase and how its differ ence with RDBMS? Discuss the advantages of HBase's advanced indexing and schema design?
- bDiscuss the ro le of Zoo Keeper in monitoring a cluster. Also disc uss the process of building applications with Zoo Keeper
Question text is extracted from the official AKTU question paper PDF above. Hindi translations are omitted — every question is printed in English in the original paper. Last verified: 2026-08-23.
Repeated Questions — KOE-097
Questions that appeared in more than one session, found by comparing 4 years of Big Data papers (2021-22, 2022-23, 2023-24, 2024-25)
Differentiate between structured, semi-structured and unstructured data
Appeared in: 2021-22 · 2022-23 · 2023-24 · 2024-25
What is the anatomy of a Map Reduce job run?
Appeared in: 2021-22 · 2022-23 · 2023-24
Explain sources of data in big data
Appeared in: 2022-23 · 2024-25
Define Metadata in HDFS
Appeared in: 2022-23 · 2023-24
Differentiate between Map & Reduce
Appeared in: 2022-23 · 2024-25
Differentiate between shuffle and sort operation
Appeared in: 2022-23 · 2024-25
Big Data — Other Year Papers
AKTU Big Data PYQs from other sessions
More B.Tech Semester 8 (2022-23) Papers
Other subjects from same semester and session
Syllabus & More PYQs
Paper solve karne se pehle unit-wise syllabus dekh lo