Page 1
FOR CBSE CLASS 10 EXAM PREPARATION
CBSE Class 10
2026
Question Paper ·
Science
EXAM YEAR TYPE SUBJECT
CBSE Class 10 2026 Question Paper Science
DETAILS
Data
Notes · Sample Papers · Previous Year Papers · Mock Tests
Page 2
Series : 2LKNM SET ~ 4
्ቚश्न-प्ቔ कोड
*DATA mSCIENCE* 106
रोल नं.
m .co Q.P. Code
s e m
Roll No.
e परी्षा्वी ्ऺश्न-प्ऴ कोड को उ्तर-पस्ु तिका के la
las मख
ु -पृ्ी पर अवश्य स्िखें । ag
ag Candidates must write the Q.P. Code
on the title page of the answer-book.
नोट / NOTE :
(I) कृ पया जााँच कर िें स्क इस ्ऺश्न-प्ऴ में मस्ु िि पृ्ी 19 हैं ।
Please check that this question paper contains 19 printed pages.
(II) ्ऺश्न-प्ऴ में दास्हने हा्व की ओर स्दए गए ्ऺश्न-प्ऴ कोड को परी्षा्वी उ्तर-पस्ु तिका के मख
ु -पृ्ी पर
स्िखें ।
m
.co
Q.P. Code given on the right hand side of the question paper should be
em
written on the title page of the answer-book by the candidate.
s
la
(III) कृ पया जााँच कर िें स्क इस ्ऺश्न-प्ऴ में 21 ्ऺश्न हैं ।
g
aु करने से पहले, उ्ቈर-पलु तिका में यथा तथान पर ्ቚश्न का
Please check that this question paper contains 21 questions.
(IV) कृपया ्ቚश्न का उ्ቈर ललखना शरू
्ቅमांक अवश्य ललखें ।
Please write down the Serial Number of the question in the
answer-book at the given place before attempting it.
(V) इस ्ऺश्न-प्ऴ को पढ़ने के स्िए 15 स्मनट का समय स्दया गया है । ्ऺश्न-प्ऴ का स्विरण पवू ाा्ቡ में
10.15 बजे स्कया जाएगा । 10.15 बजे से 10.30 बजे िक परी्षा्वी के वि ्ऺश्न-प्ऴ को पढ़ेंगे और
इस अवस्ि के दौरान वे उ्तर-पस्ु तिका पर कोई उ्तर नहीं स्िखेंगे ।
m
m 15 minute time has been allotted to read this question paper. The
.co
m .co question paper will be distributed at 10.15 a.m. From 10.15 a.m. to
s e m
se
10.30 a.m., the candidates will read the question paper only and will not
l a
ag
write any answer on the answer-book during this period. []
डेटा लव्ሺान
DATA SCIENCE
निर्धारित समय : 2 घण्टे अनर्कतम अक
ं : 50
Time allowed : 2 hours Maximum Marks : 50
106 [] Page 1 of 20 P.T.O.
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 1 of 24
Page 3
सामान्य नि्ቖेश :
(i) कृ पयध नि्शेशों को ध्यधि से पढेा ।
(ii) इस ्ቚश्ि-प्ቔ मेा ्ቖो खण्डों मेा 21 ्ቚश्ि हैं : खण्ड क औि खण्ड ख ।
(iii) खण्ड क मेा वस्तनु ि् ्ቚकधि के ्ቚश्ि हैं जबनक खण्ड ख मेा नवषयपिक ्ቚकधि के ्ቚश्ि हैं ।
(iv) न्शए गए (5 + 16) = 21 ्ቚश्िों मेा से, उम्मी्शवधि को 2 घटं े के आबनं टत (अनर्कतम) समय मेा
(5 + 10) = 15 ्ቚश्िों के उ्ቈि ्शेिे हैं ।
(v) नकसी नवशेष खण्ड के सभी ्ቚश्िों को सही ्ቅम मेा कििे कध ्ቚयधस नकयध जधिध चधनहए ।
(vi) खण्ड क : वस्तनु ि् ्ቚकधि के ्ቚश्ि ( 24 अंक) :
(a) इस खण्ड मेा 5 ्ቚश्ि हैं ।
(b) कोई िकधिधत्मक अक ं ि िहीं है ।
(c) न्शए गए नि्शेशों के अिस ु धि कीनजए ।
(d) ्ቚत्येक ्ቚश्ि/भधग के सधमिे आबनं टत अक ं ों कध उल्लेख नकयध गयध है ।
(vii) खण्ड ख : नवषयपिक ्ቚकधि के ्ቚश्ि (26 अक ं ):
(a) इस खण्ड मेा 16 ्ቚश्ि हैं ।
(b) उम्मी्शवधि को 10 ्ቚश्ि कििे हैं ।
(c) न्शए गए नि्शेशों के अिस ु धि कीनजए ।
(d) ्ቚत्येक ्ቚश्ि/भधग के सधमिे आबनं टत अक ं ों कध उल्लेख नकयध गयध है ।
खण्ड क $
(वतिुलन् ्ቚकार के ्ቚश्न) (24 अंक)
1. रोज़गार कौशि पर स्दए गए 6 ्ऺश्नों में से स्कन्हीं 4 के उ्तर दीस्जए । 41=4
(i) सच
ं ार में सदं श
े के ्ऺस्ि ्ऺा्किाा की तवीकृ स्ि और ्ऺस्िस्िया को ________ कहा जािा है ।
(A) फीडबैक (B) चैनि
(C) सच ू ना (D) त्वानांिरण
(ii) कायात्वि में तविं्ऴ रूप से काया करने की ्षमिा क्यों महቈኚवपणू ा है ?
(A) सहकस्मायों पर स्नर्ारिा बढ़ाने के स्िए
(B) नौकरी की स्जम्मेदाररयों को कम करने के स्िए
(C) टीम वका और सहयोग से बचने के स्िए
(D) व्यस्िगि स्वकास और स्नणाय िेने को बढ़ावा देने के स्िए
106 [] Page 2 of 20
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 2 of 24
Page 4
General Instructions :
(i) Please read the instructions carefully.
(ii) This question paper consists of 21 questions in two sections : Section A
and Section B.
m
.co
(iii) Section A has Objective Type Questions whereas Section B contains
e m
s
Subjective Type Questions.
e m la
las
(iv) Out of the given (5 + 16) = 21 questions, a candidate has to answer
(5 + 10) = 15 questions in the allotted (maximum) time of 2 hours.
ag
(v)
ag
All questions of a particular section must be attempted in the correct order.
(vi) Section A : Objective Type Questions (24 marks) :
(a) This section has 5 questions.
(b) There is no negative marking.
(c) Do as per the instructions given.
(d) Marks allotted are mentioned against each question/part.
(vii) Section B : Subjective Type Questions (26 marks) :
m
.co
(a) This section has 16 questions.
(b) A candidate has to do 10 questions.
e m
(c) Do as per the instructions given.
l as
(d)
g
Marks allotted are mentioned against each question/part.
a
SECTION A
(Objective Type Questions) (24 Marks)
1. Answer any 4 out of the given 6 questions on Employability skills. 41=4
(i) The receiver’s acknowledgement and response to the message is
called ___________ in communication.
m
m (A) feedback (B) channel
.co
m .co (C) information (D) transfer
s e m
se (ii)
g l a
Why is the ability to work independently important in the
workplace ? a
(A) To increase dependence on colleagues
(B) To reduce job responsibilities
(C) To avoid team work and collaboration
(D) To promote personal growth and decision-making
106 [] Page 3 of 20 P.T.O.
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 3 of 24
Page 5
(iii) स्नम्नस्िस्खि में से कौन-सी िनाव ्ऺबिं न िकनीक िहीं है ?
(A) शारीररक व्यायाम
(B) ध्यान
(C) िनावों के बारे में स्चंिा करना
(D) योग
(iv) कंप्यटू र ्ऺदशान को अनक
ु ू स्िि करने के स्िए स्नयस्मि सफाई के दौरान स्कस ्ऺकार की फाइिें
आमिौर पर हटा दी जािी हैं ?
(A) स्सतटम फाइिें
(B) सॉफ्टवेयर एस्प्िके शन
(C) अत्वायी फाइिें
(D) महቈኚवपणू ा दतिावेज़
(v) एक उ्यमी के कौन-से काया में एक नई स्वस्ि, स्वचार या उत्पाद बनाना शास्मि है ?
(A) नवाचार
(B) आय का स्वर्ाजन
(C) जोस्खम ्ऺबंिन
(D) स्नणाय िेना
(vi) स्नम्नस्िस्खि में से कौन-सा पाररस्त्वस्िक िं्ऴ और जैव-स्वस्वििा को नक
ु सान पहचाँ ाने वािी
अस्त्वर ्ऺ्वाओ ं का पररणाम है ?
(A) आस््वाक स्वकास
(B) पयाावरणीय स्गरावट
(C) अल्पकास्िक मनु ाफा
(D) जनसंख्या वृस्ि
106 [] Page 4 of 20
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 4 of 24
Page 6
(iii) Which of the following is not a stress management technique ?
(A) Physical exercise
m
(B) Meditation
(C) .co
Worrying about stressors
s e m
s emYoga g la
la
(D)
g a
a Which types of files are typically removed during a routine
(iv)
clean-up to optimize computer performance ?
(A) System files
(B) Software applications
(C) Temporary files
m
(D) Important documents
m .co
s e
(v)
l? a
Which function of an entrepreneur involves creating of a new
g
a
method, idea or product
(A) Innovation
(B) Division of income
(C) Managing risk
(D) Making decisions
m
m .co
.co
(vi) Which of the following is a consequence of unsustainable practices
e m
m that harm ecosystem and biodiversity ?
as
se (A) Economic growth
ag
l
(B) Environmental degradation
(C) Short-term profits
(D) Population growth
106 [] Page 5 of 20 P.T.O.
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 5 of 24
Page 7
2. स्दए गए 6 ्ऺश्नों में से स्कन्हीं 5 के उ्तर दीस्जए । 51=5
(i) डेटा-आिाररि उपसमायोजन (subsetting) का उपयोग िब स्कया जािा है :
(A) जब सर्ी पंस्ियों (Rows) को चनु स्िया जाए
(B) उपसमायोजन (Subsetting) स्वस्श्ि डेटा दशाओ ं पर आिाररि है
(C) जब सप ं णू ा डेटा सेट का सत्यापन हो जाए
(D) जब कॉिमों (Columns) और पंस्ियों (Rows) को यादृस्छिक (randomly) रूप
से चनु स्िया जाए
(ii) सांस्ख्यकी समतया-समािान (Statistical Problem-solving) ्ऺस्िया का मख्ु य ्ऺयोजन
क्या है ?
(A) स्वतिृि ररपोटा िैयार करना
(B) ऑनिाइन डेटा एक्ऴ करना
(C) डेटा का ्ऺयोग करिे हए खोजी ्ऺश्नों के उ्तर देना
(D) चाटटास देखना
(iii) डेटा स्व्ञान में पव ू ाा्ቇह (bias) क्या है ?
(A) एक ्ऺकार के डेटा का अविोकन (visualization)
(B) डेटा ्ऺदशान (representation) की ्ऺस्िया
(C) डेटा स्वश्िेषण (data analysis) की एक पिस्ि
(D) डेटा में ्ऺत्यास्शि पररणाम से पररविान (स्वचिन)
(iv) अस्र्क्वन (A) : स्वस्र्न्न परी्षणों के ्ऺा्ांकों (scores) की िि ु ना करने में शिमक
(percentiles) मदद करिे हैं ।
कारण (R) : शिमक (percentiles) एक स्निााररि डेटा प्वॉइटं पर या उससे नीचे
वैल्यजू (values) के अनपु ाि को ्ऺदस्शाि करिे हैं ।
(A) अस्र्क्वन (A) और कारण (R) दोनों सही हैं और कारण (R), अस्र्क्वन (A) का
सही तपष्टीकरण है ।
(B) अस्र्क्वन (A) और कारण (R) दोनों सही हैं, परंिु कारण (R), अस्र्क्वन (A) का
सही तपष्टीकरण िहीं है ।
(C) अस्र्क्वन (A) सही है, परंिु कारण (R) ग़िि है ।
(D) अस्र्क्वन (A) ग़िि है, परंिु कारण (R) सही है ।
(v) बिाइए स्क सही है या ग़िि :
स्वश्िेषण के स्िए स्ቇं हीि डेटा को कर्ी र्ी मानव इछिा में हति्षेप (interfere) नहीं करना
चास्हए ।
106 [] Page 6 of 20
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 6 of 24
Page 8
2. Answer any 5 out of the given 6 questions. 51=5
(i) Data-based subsetting is used when :
(A) All the rows are selected
(B)
o m
Subsetting is based on specific data conditions
m
. c s e
emColumns and rows are randomly selected a
(C) The entire dataset is verified
s g l
la a
(D)
g
(ii)a What is the main purpose of the statistical problem-solving
process ?
(A) To prepare detailed report
(B) To collect online data
(C) To answer investigative questions using data
(D) To visualize charts
m
(iii) What is bias in data science ?
.co
(A)
s em
A type of data visualization
g la
(B) The process of representing data
(C) A method of dataa analysis
(D) A deviation from the expected outcome in data
(iv) Assertion (A) : Percentiles help compare scores across different
tests.
Reason (R) : Percentiles represent the proportion of values at or
below a certain data point.
m
m (A) Both Assertion (A) and Reason (R) are true and Reason (R) is
.co
m .co the correct explanation of Assertion (A).
e m
s Reason (R) is
se l
(B) Both Assertion (A) and Reason (R) are true, but a
ag
not the correct explanation of Assertion (A).
(C) Assertion (A) is true, but Reason (R) is false.
(D) Assertion (A) is false, but Reason (R) is true.
(v) State True or False :
Data collected for analysis should never interfere with human will.
106 [] Page 7 of 20 P.T.O.
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 7 of 24
Page 9
(vi) सं्ቇहण उपकरण (storage device) में डेटा का सं्ቇहण करिे समय डेटा का/की
_________ एक अछिी पिस्ि है, िास्क हैकसा आपके डेटा को न पढ़ सकें ।
(A) ्ऺस्िस्िस्प (Copy) करना (B) फोटो िेना
(C) कोडीकरण (Encrypt) करना (D) फॉमेट (Format) करना
3. स्दए गए 6 ्ऺश्नों में से स्कन्हीं 5 के उ्तर दीस्जए । 51=5
(i) अस्र्क्वन (A) : जहााँ डेटा सेट में स्र्न्निा (outlier) हों वहााँ माध्यक (median) के न्िीय
्ऺवृस््त (central tendency) का अस्िक ्ऺर्ावशािी उपाय है ।
कारण (R) : स्कसी डेटा सेट में माध्यक सर्ी मानों (values) का औसि (average) है ।
(A) अस्र्क्वन (A) और कारण (R) दोनों सही हैं और कारण (R), अस्र्क्वन (A) का
सही तपष्टीकरण है ।
(B) अस्र्क्वन (A) और कारण (R) दोनों सही हैं, परंिु कारण (R), अस्र्क्वन (A) का
सही तपष्टीकरण िहीं है ।
(C) अस्र्क्वन (A) सही है, परंिु कारण (R) ग़िि है ।
(D) अस्र्क्वन (A) ग़िि है, परंिु कारण (R) सही है ।
(ii) स्नम्नस्िस्खि ्ቇाफ ्षारा ्ऺदस्शाि स्विरण (distribution) के स्कतम का नाम बिाइए :
(iii) एक अनसु ंिानकिाा के वि उन पररणामों को देखिा है जो वह अपने अध्ययन में देखना चाहिा
है । यह _________ का एक उदाहरण है ।
(A) पस्ु ्िकरण पव ू ाा्ቇह (Confirmation bias) (B) रै स्खक पवू ाा्ቇह (Linearity bias)
(C) तमरण पव ू ाा्ቇह (Recall bias)$ (D) चयन पवू ाा्ቇह (Selection bias)$
(iv) ज़ेड-तकोर (Z-Score) ्ऺदस्शाि करिा है :
(A) डेटा का सवाास्िक सामान्य (common) वैल्यू
(B) एक तकोर से नीचे के वैल्यज ू (values) का ्ऺस्िशि
(C) मानक स्वचिनों (standard deviations) की वह सख् ं या जो औसि (mean) से
डेटा प्वॉइटं में हो
(D) डेटा प्वॉइटं ट स का कुि योग
106 [] Page 8 of 20
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 8 of 24
Page 10
(vi) While storing data in a storage device, it is a good practice to
___________ the data, so that hackers cannot read your data.
(A) copy (B) take photographs of
om
(C) encrypt (D) format
. c e m
3. Answer any 5 out of the given 6 questions.
m
e (A) : Median is a more effective measure of centralgla s 51=5
las
tendency where there are outliers in the data set. a
(i) Assertion
g
a Reason (R) : Median is the average of all the values in a dataset.
(A) Both Assertion (A) and Reason (R) are true and Reason (R) is
the correct explanation of Assertion (A).
(B) Both Assertion (A) and Reason (R) are true, but Reason (R) is
not the correct explanation of Assertion (A).
(C) Assertion (A) is true, but Reason (R) is false.
m
(D) Assertion (A) is false, but Reason (R) is true.
(ii) .co
Name the type of distribution represented by the following graph :
m
s e
g l a
a
(iii) A researcher only sees the results they want to see in their study.
m
.co
This is an example of :
m
.co m
(A) Confirmation Bias (B) Linearity Bias
m s e
se
(C) Recall Bias (D) Selection Bias
l a
(iv) The Z-score represents : ag
(A) The most common value of data
(B) The percentage of values below a score
(C) The number of standard deviations a data point is from the
mean
(D) The total sum of the data points
106 [] Page 9 of 20 P.T.O.
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 9 of 24
Page 11
(v) एक डेटा सेट में Q1 = 20, Q3 = 50. IQR क्या है ?
(A) 30 (B) 20
(C) 50 (D) 70
(vi) स्नम्नस्िस्खि में से कौन-सा डेटा स्वश्िेषण के स्िए नैस्िक स्दशास्नदेश (ethical
guideline) िहीं है ?
(A) डेटा के सही और स्वश्वसनीय ्ቨोि सस्ु नस्श्चि करना
(B) स्नजिा और गोपनीयिा बनाए रखना
(C) स्वश्िेषण के स्िए के वि सगं ि डेटा (relevant data) का उपयोग करना
(D) डेटा स्वश्िेषक अपनी व्यस्िगि पसंदों के आिार पर डेटा चनु सकिे हैं
4. स्दए गए 6 ्ऺश्नों में से स्कन्हीं 5 के उ्तर दीस्जए । 51=5
(i) स्दन के समय चाय या कॉफी पीने वािे अनेक िोगों पर एक सवे्षण स्कया गया और नीचे स्दए
गए फॉमेट में एक सा्व आाँकडे (data) रखे गए :
पेय (beverage) ्ऺाि: सायंकाि कुि
कॉफी 43 56 99
चाय 60 72 132
योग 103 128 231
इस ्ऺकार की सारणी (table) का क्या नाम है ?
(A) सब्सेट टेबि (Subset table)
(B) टू-वे ्ቛीक्वेंसी टेबि (Two-way frequency table)
(C) टू-वे ररिेस्टव ्ቛीक्वेंसी टेबि (Two-way relative frequency table)
(D) तटैंडडा डेस्वएशन टेबि (Standard deviation table)
(ii) स्नम्नस्िस्खि में से कौन-सा सांस्ख्यकी खोजी ्ऺश्न (Statistical investigative
question) िहीं है ?
(A) एक दस वषा का बािक स्किना िेज दौड सकिा है ?
(B) क्या जो बछचे उस्चि नाश्िा करिे हैं अस्िक िेज दौड सकिे हैं ?
(C) बछचे के कायाकिापों को नींद स्कस िरह ्ऺर्ास्वि करिी है ?
(D) बछचे ने 200 मीटर दौडने में स्किना समय िगाया ?
106 [] Page 10 of 20
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 10 of 24
Page 12
(v) In a dataset, Q1 = 20, Q3 = 50. What is the IQR ?
(A) 30 (B) 20
(C) 50 (D) 70
m
(vi)
.co
Which of the following is not an ethical guideline for data
m s e m
as e
analysis ?
g la
l (A) Ensure accurate and reliable sources of data
a
ag (B) Maintain privacy and confidentiality
(C) Use only relevant data for analysis
(D) Data analysts may select data based on their personal
preferences
4. Answer any 5 out of the given 6 questions. 51=5
m
(i)
.co
A survey was done on the number of people consuming tea or coffee
em
during the day and the data was put together in the format given
s
below :
g l a
Beverage a
Morning Evening Total
Coffee 43 56 99
Tea 60 72 132
Total 103 128 231
What is the name of this type of table ?
(A) Subset table
m
.co
(B) Two-way frequency table
m
.co m
(C) Two-way relative frequency table
m s e
se
(D) Standard deviation table
l a
(ii) ag question ?
Which of the following is not a statistical investigative
(A) How fast can a ten-year-old child run ?
(B) Do children who have proper breakfast run faster ?
(C) How does sleep affect the performance of a child ?
(D) What time did the child take to run 200 m ?
106 [] Page 11 of 20 P.T.O.
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 11 of 24
Page 13
(iii) एक शहर में िा्ऴों की औसि िंबाई का पिा िगाने के स्िए एक सवे्षण स्कया जािा है । सर्ी
िा्ऴों की िम्बाई नापने के बजाय अनसु ंिानकिाा स्वस्र्न्न तकूिों से साइज 40 के नमनू े िेिा है
और ्ऺत्येक नमनू े के व्यस्िगि औसि की गणना करिा है । इसके बाद इन व्यस्िगि नमनू ा
औसिों का औसि स्नकािा जािा है । यह पाया गया स्क िा्ऴों की नमनू ा-औसि-िम्बाई का
आयिस्च्ऴ (histogram) सामान्य स्विरण (normal distribution) से समानिा रखिा
है । यहााँ कौन-सी सांस्ख्यकी अविारणा (statistical concept) ्ऺदस्शाि की गई है ?
(A) संर्ाव्यिा (Probability)
(B) रै स्खक पव ू ाा्ቇह (Linear bias)
(C) के न्िीय सीमा ्ऺमेय (Central Limit Theorem)
(D) सच ू ना स्ቇं हण मल्ू याक
ं न (Data evaluation)
(iv) स्नम्नस्िस्खि में से कौन-सा चि्वु ाक (Quartiles) से सही रूप में मेि खािा है ?
(A) Q1 – 25%, Q2 – 50%, Q3 – 75%, Q4 – 100%
(B) Q1 – 10%, Q2 – 20%, Q3 – 30%, Q4 – 50%
(C) Q1 – 33%, Q2 – 66%, Q3 – 100%, Q4 – 99%
(D) Q1 – 50%, Q2 – 75%, Q3 – 100%, Q4 – 25%
(v) औसि से इसकी दरू ी के संदर्ा में स्कसी प्वॉइटं की स्त्वस्ि को बिाने वािे इस सांस्ख्यकी शब्द
की वैल्यू मानक स्वचिन यस्ू नटों (standard deviation units) में नापी जािी है । इस
शब्द (term) को क्या कहिे हैं ?
(vi) डेटा सचं ािन सरं चना (data governance framework) का मख्ु य उ्देश्य क्या है ?
(A) िार् मास्जान को बढ़ाना
(B) डेटा का मानकीकरण, एकीकरण, संर्षण और सं्ቇहण
(C) के वि डुप्िीके ट ररकॉडटास को हटाना
(D) डेटा ्ऺत्य्षीकरण (data visualizations) को िैयार करना
5. स्दए गए 6 ्ऺश्नों में से स्कन्हीं 5 के उ्तर दीस्जए । 51=5
(i) स्नम्नस्िस्खि में से कौन-सा मानक स्वचिन (standard deviation) का वातिस्वक-जीवन
में अन्ऺु योग (real-life application) है ?
(A) Excel में डेटा ्ቇाफ बनाना
(B) के वि औसि वषाा की गणना करना
(C) परी्षा में िा्ऴों के परफॉमेन्स त्ऺेड (performance spread) का मापन करना
(D) त्ऺेडशीट में कॉिमों की संख्या का पिा करना
106 [] Page 12 of 20
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 12 of 24
Page 14
(iii) A survey is conducted to find the average height of students in a
city. Instead of measuring all students, a researcher takes samples
of size 40 from different schools and calculates the individual mean
m
of each sample. Thereafter the mean of these individual sample
.co
means is calculated. It is noticed that the histogram of sample
s e m
em
mean heights of students resembles the normal distribution. Which
s g la
g la(A) Probability
statistical concept is demonstrated here ?
a
a (B) Linear bias
(C) Central Limit Theorem (D) Data evaluation
(iv) Which of the following correctly matches quartiles ?
(A) Q1 – 25%, Q2 – 50%, Q3 – 75%, Q4 – 100%
(B) Q1 – 10%, Q2 – 20%, Q3 – 30%, Q4 – 50%
(C) Q1 – 33%, Q2 – 66%, Q3 – 100%, Q4 – 99%
(D)
o m
Q1 – 50%, Q2 – 75%, Q3 – 100%, Q4 – 25%
c
. that describes the position of a
(v)
m
efrom the mean, when it is measured in
The value of this statistical term
a s
l Name this term.
point in terms of its distance
a g
standard deviation units.
(vi) What is the main aim of a data governance framework ?
(A) To increase profit margins
(B) To standardize, integrate, protect, and store data
(C) To delete duplicate records only
(D) To design data visualizations
m
m
c. o Answer any 5 out of given 6 questions. m .co
m
5.
e 51=5
s of standard
se l a
ag
(i) Which of the following is a real-life application
deviation ?
(A) Plotting data graphs in Excel
(B) Calculating average rainfall only
(C) Measuring students’ performance spread in a test
(D) Finding the number of columns in a spreadsheet
106 [] Page 13 of 20 P.T.O.
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 13 of 24
Page 15
(ii) डेटा सं्ቇहण (collection) स्डज़ाइन को डेटा की पररविानशीििा (variability) को
तवीकार करना चास्हए । डेटा में पररविानशीििा को कम करने और उसका पिा िगाने के स्िए
कुि पिस्ियों का ्ऺयोग स्कया जािा है । स्नम्नस्िस्खि में से स्कस पिस्ि का उपयोग स्कया जा
सकिा है ?
(A) सांस्ख्यकीय ्ऺस्िया स्नयं्ऴण (Statistical Process Control)
(B) संर्ाव्यिा (Probability)
(C) स्विरण (Distribution)
(D) इवेंट (Event)
(iii) अस्र्क्वन (A) : पवू ाा्ቇहपणू ा डेटा (Biased data) से ग़िि पवू ाानमु ास्नि मॉडि हो
सकिा है ।
कारण (R) : पवू ाानमु ास्नि मॉडि (Predictive models) के वि उन डेटा पर स्वचार
करिे हैं जो इसके ्ऺस्श्षण के स्िए ्ऺणािी (system) में फीड स्कया
गया हो ।
(A) अस्र्क्वन (A) और कारण (R) दोनों सही हैं और कारण (R), अस्र्क्वन (A) का
सही तपष्टीकरण है ।
(B) अस्र्क्वन (A) और कारण (R) दोनों सही हैं, परंिु कारण (R), अस्र्क्वन (A) का
सही तपष्टीकरण िहीं है ।
(C) अस्र्क्वन (A) सही है, परंिु कारण (R) ग़िि है ।
(D) अस्र्क्वन (A) ग़िि है, परंिु कारण (R) सही है ।
(iv) आाँकडों के स्विय (data merging) की ्ऺस्िया के बारे में स्नम्नस्िस्खि में से कौन-सा सही
है ?
(A) स्विीन (merge) स्कए जा रहे सर्ी आाँकडा ्ቨोि (data sources) हमेशा समान
िरीके से वगीकृ ि (grouped) होिे हैं ।
(B) बहि आाँकडा ्ቨोिों (multiple data sources) के बीच काफी अंिर होिा है ।
(C) सर्ी आाँकडा ्ቨोि एक ही उ्देश्य से एक ही समय में िैयार स्कए जािे हैं ।
(D) स्वस्र्न्न आाँकडा ्ቨोिों से स्विय स्कए गए आाँकडों को सुिारने की कोई आवश्यकिा
नहीं है ।
106 [] Page 14 of 20
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 14 of 24
Page 16
(ii) Data collection designs must acknowledge variability in data. Few
methods are used to reduce and detect variability in data. Which of
the following method can be used ?
(A)
o m
Statistical Process Control
m
. c s e
m
(B) Probability
e Distribution la
s
la(D) Event
(C)
ag
ag
(iii) Assertion (A) : Biased data can lead to inaccurate predictive
models.
Reason (R) : The predictive models consider only the data that is
fed into the system for training it.
m
(A)
.co
Both Assertion (A) and Reason (R) are true and Reason (R) is
s em
the correct explanation of Assertion (A).
g a Reason (R) are true, but Reason (R) is
land
a
(B) Both Assertion (A)
not the correct explanation of Assertion (A).
(C) Assertion (A) is true, but Reason (R) is false.
(D) Assertion (A) is false, but Reason (R) is true.
(iv) Which of the following is true about the process of data merging ?
m
m (A) All data sources being merged are always grouped in similar
.co
m .co manner.
s e m
se l a
ag
(B) There is a lot of difference between multiple data sources.
(C) All data sources are created with the same objective and at
the same time.
(D) No correction is required on the data merged from different
data sources.
106 [] Page 15 of 20 P.T.O.
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 15 of 24
Page 17
(v) स्नम्नस्िस्खि क्वन (statements) एक-से-एक जडु ाव (one-to-one join) के संबंि में
हैं :
(i) एक टेबि में ्ऺत्येक पस्ं ि (row) को दसू री टेबि के एकि पस्ं ि (row) से जोडा
जािा है ।
(ii) ‘एक छा्ቔ अनेक पाठ्य्ቅमों मे पंजीकरण कर सकिा है’ इस जड ु ाव (join) का
वैि (मान्य) उदाहरण है ।
(iii) दो टेबल्स की स्संगि पंस्ि (single row) के बीच जोडने का काया की कॉिम (Key
column) के उपयोग से स्कया जािा है ।
(iv) टेबल्स को जोडने के स्िए ्ऺयि ु की फील्ड (Key field) को अस््षिीय मूल्यों
(Unique values) को अंिस्वा्ि (Contain) करने के स्िए िैयार स्कया जािा है ।
(v) ‘एक छा्ቔ के पास के वल एक पहचान प्ቔ (Id) हो सकिा है’ इस जड ु ाव (join)
का एक वैि (मान्य) उदाहरण है ।
स्दए गए क्वनों में से कौन-सा सही हैं ?
(A) (i), (ii) और (iv) (B) (i), (iii), (iv) और (v)
(C) (ii), (iv) और (v) (D) (ii), (iii), (iv) और (v)
(vi) सहमस्ि के सा्व स्कसी व्यस्ि से एकस््ऴि स्नजी आाँकडे ऐसे होने चास्हए स्क :
(A) स्जसको आवश्यकिा हो ऐसे स्कसी र्ी व्यस्ि के सा्व इनकी स्बना रोक-टोक के
साझेदारी हो सके
(B) आाँकडा स्वश्िेषण के स्िए इन्हें उपिब्ि स्कया जा सके
(C) हमेशा गोपनीयिा के सा्व इनका संचािन स्कया जा सके
(D) स्कसी र्ी पररस्त्वस्ि में कर्ी र्ी इनकी िेखापरी्षा (audit) न हो
खण्ड ख $
(लवषयपरक ्ቚकार के ्ቚश्न) (26 अंक)
िोज़गधि कौशल पि न्शए गए 5 ्ቚश्िों मेा से नकन्हीं 3 के उ्ቈि ्शीनजए । ्ቚत्येक ्ቚश्ि कध उ्ቈि 20 – 30 शब््शों मेा
्शीनजए । 32=6
6. सचं ार में र्ाषाई बािाएाँ क्या हैं ?
7. र्ावनात्मक बस्ु िम्ता को ्ऺबस्ं िि करने के चरण बिाइए ।
8. (क) तपैम (SPAM) मेि क्या हैं ?
(ख) क्या हमें उनका जवाब देना चास्हए ?
9. उ्यस्मिा के बारे में कोई दो स्म्वक स्िस्खए ।
10. हम संपोषणीय (sustainable) शहर कै से बना सकिे हैं ?
106 [] Page 16 of 20
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 16 of 24
Page 18
(v) The following statements are with respect to one-to-one join :
(i) Each row in one table is linked to a single row in another
table.
(ii) ‘A student can register in multiple courses’ is a valid
o m
example of this join.
m
. c s e
emkey column.
(iii) The linking between single row of two tables is done using a
s g l a
la(iv) The key field used to link the tables is designed to containa
ag unique values.
(v) ‘A student can have only one Student Id’ is a valid
example of this join.
Which of the given statements are correct ?
(A) (i), (ii) and (iv) (B) (i), (iii), (iv) and (v)
(C) (ii), (iv) and (v) (D) (ii), (iii), (iv) and (v)
m
(vi)
.co
Private data collected from a person with consent should :
m
(A)
s e
Be freely shared with anyone who needs it
g la data analysis
(B) Be made available for
a with confidentiality
(C) Always be handled
(D) Never be audited under any circumstance
SECTION B
(Subjective Type Questions) (26 Marks)
Answer any 3 out of the given 5 questions on Employability skills. Answer each
m
m
question in 20 – 30 words.
c. o What are linguistic barriers to communication ?
32=6
m .co
m
6.
s e
se 7. l a
State the steps to manage emotional intelligence.
ag
8. (a) What are SPAM mails ?
(b) Should we respond to them ?
9. Write any two myths about entrepreneurship.
10. How can we create sustainable cities ?
106 [] Page 17 of 20 P.T.O.
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 17 of 24
Page 19
न्शए गए 6 ्ቚश्िों मेा से नकन्हीं 4 के उ्ቈि 20 – 30 शब््शों (्ቚत्येक) मेा ्शीनजए । 42=8
11. स्नम्नस्िस्खि डेटासेट पर स्वचार कीस्जए :
[5, 10, 2, 1, 20, 6, 15]
इसका माध्य (mean) और माध्यक (median) ्ञाि कीस्जए ।
12. डेटा पृ्वक (discrete) या सिि (continuous) हो सकिा है । पृ्वक और सिि डेटा के बीच एक
अंिर बिाइए । ्ऺत्येक का एक उदाहरण र्ी दीस्जए ।
13. चयन पवू ाा्ቇह (selection bias) को पररर्ास्षि कीस्जए । यह कब होिा है ?
14. यस्द Z-score का वैल्यू (value) िनात्मक (positive) है या ॠणात्मक (negative), िो आप
इससे क्या समझिे हैं ? अपने उ्तर के सम्वान में उदाहरण दीस्जए ।
15. अनेक के सा्व अनेक जडु ाव (Many-to-many join) को एक उदाहरण के सा्व तप्ि कीस्जए ।
16 एक कंपनी, ्ቇाहकों के खरीददारी आाँकडे का स्वश्िेषण कर रही है । डेटासेट को िोटा बनाने के स्िए
स्व्ቨेषक स्बना औस्चत्य के 18 वषा से नीचे के ्ቇाहकों के सारे ररकॉडा िोडने का स्नणाय िेिा है ।
(a) इस मामिे में नैस्िक म्दु े का पिा कीस्जए ।
(b) इसे स्नकाििे (discard) समय क्या कंपनी को डेटा सॉफ्ट स्डिीट कर देना चास्हए । क्यों या
क्यों नहीं ?
न्शए गए 5 ्ቚश्िों मेा से नकन्हीं 3 के उ्ቈि 50 – 80 शब््शों (्ቚत्येक) मेा ्शीनजए । 34=12
17. औसि स्नरपे्ष स्वचिन (Mean Absolute Deviation) (MAD) को पररर्ास्षि कीस्जए ।
उदाहरण के रूप में स्नम्नस्िस्खि डेटासेट को िेकर MAD संगणना के कदमों (steps) को तप्ि
कीस्जए ।
[10, 12, 14, 16, 18]
18. डेटा स्व्ञान में स्विरण (distribution) से आप क्या समझिे हैं ? यह संर्ाव्यिा (probability) से
स्कस ्ऺकार स्र्न्न है ? स्सक्का उिािना (Tossing the coin) इवेंट की सहायिा से तप्ि कीस्जए ।
19. (a) सास्ं ख्यकी में के न्िीय सीमा ्ऺमेय (Central Limit Theorem) (CLT) को महቈኚवपणू ा
क्यों माना जािा है ?
(b) CLT के सदं र्ा में, यस्द डेटा का नमनू ा आकार बढ़िा है, िो ्ऴस्ु ट बढ़ेगी या कम होगी ?
(c) CLT के स्कन्हीं दो वातिस्वक दस्ु नया में अन्ऺ ु योगों (real world applications) का
उल्िेख कीस्जए ।
20. (a) दशमक (deciles) क्या हैं ?
(b) 15 संख्याओ ं के स्नम्नस्िस्खि डेटासेट पर स्वचार कीस्जए :
[77, 60, 63, 36, 54, 57, 36, 72, 55, 51, 32, 56, 33, 42, 55]
D1 ि्वा D5 के डेटा स्त्वस्ि और decile value की गणना कीस्जए ।
21. डेटा का ्ऺयोग हो जाने के बाद उसे अछिी िरह स्नकाि देना (discard) क्यों महቈኚवपणू ा है ? गोपनीय
डेटा की हाडा कॉपी (physical copies) हटाने के िीन िरीकों को तप्ि कीस्जए ।
106 [] Page 18 of 20
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 18 of 24
Page 20
Answer any 4 out of the given 6 questions in 20 – 30 words each. 42=8
11. Consider the following dataset :
[5, 10, 2, 1, 20, 6, 15]
m
Find the mean and median.
12.
.co
The data can be discrete or continuous. Give one difference between
m s e m
e bias. When does it occur ? la
discrete and continuous data. Also give one example of each.
13. s
Defineaselection
ldo you infer if the value of Z-score is positive or negative ? Giveag
14. ag
What
example to support your answer.
15. Explain Many-to-Many join with an example.
16. A company is analyzing purchase data of customers. The analyst decides
to discard all records from customers under 18 without justification, to
make the dataset smaller.
(a) Identify the ethical issue in this case.
m
.co
(b) Should the company soft delete the data while discarding it ?
Why/Why not ?
s em
Answer any 3 out of the given 5 questions in 50 – 80 words each.
a
34=12
g l (MAD). Explain the steps to calculate
a as an example :
17. Define Mean Absolute Deviation
MAD taking the following dataset
[10, 12, 14, 16, 18]
18. What do you mean by distribution in data science ? How is it different
from probability ? Explain with the help of the event – Tossing the coin.
19. (a) Why is Central Limit Theorem (CLT) considered important in
statistics ?
m
.co
(b) With reference to CLT, if the sample size of data increases, will the
m
c. o (c)
error increase or decrease ?
e m
s
Mention any two real world applications of CLT.
m a
se 20. (a)
(b)
What are deciles ?
Consider the following dataset of 15 numbers :a
g l
[77, 60, 63, 36, 54, 57, 36, 72, 55, 51, 32, 56, 33, 42, 55]
Calculate the Data Position and decile value for D1 and D5.
21. Why is it important to discard data properly once its use is over ? Explain
the three methods to discard physical copies of confidential data.
106 [] Page 19 of 20 P.T.O.
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 19 of 24
Page 21
106 [] Page 20 of 20 P.T.O.
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 20 of 24
Page 22
m
m .co s e m
s e g la
g la a
a
m
m .co
s e
g l a
a
m
m .co
m .co s e m
se l a
ag
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 21 of 24
Page 23
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 22 of 24
Page 24
m
m .co s e m
s e g la
g la a
a
m
m .co
s e
g l a
a
m
m .co
m .co s e m
se l a
ag
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 23 of 24
Page 25
For more Question Papers, Sample Papers, Notes & Syllabus visit Page 24 of 24