What is this course about?
This course focuses on big data analytics with an emphasis on text data, which constitutes a large portion of the data generated daily in modern systems. The course aims to answer a central question: how to represent large-scale data, how to learn from it, and how to apply it to real-world problems such as search and recommendation.
Students will study core methods for text representation, including sparse and dense models, and learn how these representations are used in text classification, graph-based learning, and modern search and retrieval pipelines.
