Skip to content
CS-503 (A) · Data Analytics/Important Questions

Data Analytics (CS-503 (A)) - Important Questions

  1. Unit 17 Marks High Priority

    Suppose the weights of 800 male students are normally distributed with mean 68.8 kg and standard deviation 2.06 kg. Find the number of students whose weights are (i) between 68 kg and 72 kg and (ii) above 72 kg.

    Predicted for DEC-2026

  2. Unit 17 Marks High Priority

    A discrete random variable X has the following probability function with unknown constant k: P(X=0)=k, P(X=1)=2k, P(X=2)=3k, P(X=3)=4k. Determine (i) the value of k (ii) mean and (iii) variance of X.

    Predicted for DEC-2026

  3. Unit 17 Marks High Priority

    Explain normal, binomial and Poisson probability distributions with examples. State their properties and applications in data analytics.

    Predicted for DEC-2026

  4. Unit 27 Marks High Priority

    Discuss the trends in big data generation and acquisition.

    Predicted for DEC-2026

  5. Unit 214 Marks High Priority

    Explain the following in detail with examples: (i) Drivers for Big Data (ii) Predictive analytics and its applications.

    Predicted for DEC-2026

  6. Unit 27 Marks High Priority

    With an example, explain the term social media analytics.

    Predicted for DEC-2026

  7. Unit 27 Marks High Priority

    What are the various stages in big data analytics life cycle? Illustrate with a figure, explaining each of them.

    Predicted for DEC-2026

  8. Unit 47 Marks High Priority

    Explain the MapReduce programming model and its main components.

    Predicted for DEC-2026

  9. Unit 27 Marks High Priority

    What is Hadoop? Describe the role of Hadoop in big data analysis. Also explain its architecture and core components.

    Predicted for DEC-2026

  10. Unit 47 Marks High Priority

    Describe the structure of HDFS in a Hadoop ecosystem using a diagram. Explain the architecture and data storage process with an example.

    Predicted for DEC-2026

  11. Unit 27 Marks High Priority

    Why to choose Hadoop for processing Big Data in detail and explain the concept of distributed and parallel computing challenges?

    Predicted for DEC-2026

  12. Unit 27 Marks High Priority

    Explain in detail the interacting process with Hadoop ecosystem. List out various big data processing technologies.

    Predicted for DEC-2026

  13. Unit 27 Marks High Priority

    What are the key drivers for Big Data? Explain in detail with suitable examples.

    Predicted for DEC-2026

  14. Unit 57 Marks High Priority

    Explain Pig Data Model in detail and discuss how it will help for effective data flow.

    Predicted for DEC-2026

  15. Unit 57 Marks High Priority

    Draw and explain architecture of Apache Hive. Explain various data insertion techniques in Hive with example.

    Predicted for DEC-2026

Go to where you left off?

Quick Add to Notes

Save questions, your own notes and screenshots into notes filed by unit. It takes a free account.

Create free account

Have an account? Log in