Skip to content
IT-603 (B) · Data Mining/Important Questions

Data Mining (IT-603 (B)) - Important Questions

  1. 7 Marks High Priority Asked: 2025

    Compare ROLAP, MOLAP, OLAP. Explain different OLAP operations.

    Appeared 1x (2025)

  2. 7 Marks High Priority Asked: 2025

    Draw a snowflake schema diagram for "University" data warehouse which consist of four dimensions student, course, semester and instructor and two measures count and avg_grade. At the lowest conceptual level, the avg_grade measure stores the actual course grade of student. At higher level avg_grades store average grade for given schema.

    Appeared 1x (2025)

  3. 7 Marks High Priority Asked: 2025

    Describe the steps involved in the knowledge discovery process (KDD) in data mining.

    Appeared 2x (2025)

  4. 7 Marks High Priority Asked: 2025

    Explain data transformation strategies in detail.

    Appeared 1x (2025)

  5. 7 Marks High Priority Asked: 2025

    In real world data, tuples with missing values for some attributes are a common occurrence. Describe various methods for handling this problem.

    Appeared 1x (2025)

  6. 7 Marks High Priority Asked: 2025

    Give an example showing that items in a strong association rule can be negatively correlated.

    Appeared 1x (2025)

  7. 7 Marks High Priority Asked: 2025

    Apriori algorithm makes use of prior knowledge of subset support properties. Prove that all non empty subsets of a frequent itemset must also be frequent.

    Appeared 1x (2025)

  8. 7 Marks High Priority Asked: 2025

    Explain why tree pruning is useful in decision tree induction and a drawback of using a separate validation set to evaluate pruning.

    Appeared 1x (2025)

  9. 7 Marks High Priority Asked: 2025

    Give an example showing why k-means may not find the global optimum of within-cluster variation.

    Appeared 1x (2025)

  10. 7 Marks High Priority Asked: 2025

    Explain a clustering-based outlier detection method that can detect outliers at different granularity levels in a cluster hierarchy.

    Appeared 1x (2025)

  11. 7 Marks High Priority Asked: 2025

    Explain the advantage of using unlabeled objects in the training dataset for semi-supervised outlier detection.

    Appeared 1x (2025)

  12. 7 Marks High Priority Asked: 2025

    Write the k-means algorithm with an example.

    Appeared 1x (2025)

  13. 7 Marks High Priority Asked: 2025

    What is Web mining? Explain its types.

    Appeared 1x (2025)

  14. 7 Marks High Priority Asked: 2025

    Describe the security issues in data mining.

    Appeared 1x (2025)

  15. 7 Marks High Priority Asked: 2025

    Describe spatial and Temporal mining with suitable example.

    Appeared 1x (2025)

Go to where you left off?

Quick Add to Notes

Save questions, your own notes and screenshots into notes filed by unit. It takes a free account.

Create free account

Have an account? Log in