Skip to content
IT-603 (B) · Data Mining/Important Questions

Data Mining (IT-603 (B)) - Important Questions

  1. Unit 27 Marks High Priority

    Describe the steps involved in the knowledge discovery process (KDD) in data mining.

    Predicted for DEC-2026

  2. Unit 17 Marks High Priority

    Compare ROLAP, MOLAP and OLAP. Explain different OLAP operations.

    Predicted for DEC-2026

  3. Unit 27 Marks High Priority

    Explain data transformation strategies in detail.

    Predicted for DEC-2026

  4. Unit 27 Marks High Priority

    In real world data, tuples with missing values for some attributes are a common occurrence. Describe various methods for handling this problem.

    Predicted for DEC-2026

  5. Unit 17 Marks High Priority

    Draw a snowflake schema diagram for "University" data warehouse which consists of four dimensions student, course, semester and instructor and two measures count and avg_grade. At the lowest conceptual level, show the normalized dimension tables.

    Predicted for DEC-2026

  6. Unit 27 Marks High Priority

    Explain data mining functionalities and classification of data mining systems.

    Predicted for DEC-2026

  7. Unit 37 Marks High Priority

    Give an example showing that items in a strong association rule can be negatively correlated.

    Predicted for DEC-2026

  8. Unit 37 Marks High Priority

    Apriori algorithm makes use of prior knowledge of subset support properties. Prove that all non-empty subsets of a frequent itemset must also be frequent.

    Predicted for DEC-2026

  9. Unit 47 Marks High Priority

    Explain why tree pruning is useful in decision tree induction and a drawback of using a separate validation set to evaluate pruning.

    Predicted for DEC-2026

  10. Unit 47 Marks High Priority

    Give an example showing why k-means may not find the global optimum of within-cluster variation.

    Predicted for DEC-2026

  11. Unit 47 Marks High Priority

    Explain a clustering-based outlier detection method that can detect outliers at different granularity levels in a cluster hierarchy.

    Predicted for DEC-2026

  12. Unit 47 Marks High Priority

    Explain the advantage of using unlabeled objects in the training dataset for semi-supervised outlier detection.

    Predicted for DEC-2026

  13. Unit 57 Marks High Priority

    What is Web mining? Explain its types.

    Predicted for DEC-2026

  14. Unit 57 Marks High Priority

    Describe the security issues in data mining.

    Predicted for DEC-2026

  15. Unit 57 Marks High Priority

    Describe spatial and temporal mining with suitable example.

    Predicted for DEC-2026

  16. Unit 47 Marks High Priority

    Write the k-means algorithm with an example.

    Predicted for DEC-2026

Go to where you left off?

Quick Add to Notes

Save questions, your own notes and screenshots into notes filed by unit. It takes a free account.

Create free account

Have an account? Log in