Course Detail (Course Description By Faculty)

Data Intelligence (41215)

This course sits at the intersection of Data Science and Artificial Intelligence.

It is designed for future business leaders who want to engage with data directly rather than defer to the analysts in the room.

Its primary objectives are to:

  • become fluent in the language of “data science and AI”,
  • learn how to effectively communicate uncertainty,
  • cultivate the ability to craft compelling narratives grounded in data,
  • use data analytics for informed decision-making in uncertain environments,

  • develop skills to explore and make sense of large and messy data,

  • build and interpret (predictive) models with confidence,

  • learn to direct AI coding with Claude Code to produce a defensible data analysis.

Ideal for students preparing for careers in data-rich environments, the course emphasizes data storytelling. The goal is to learn how to turn data into a clear, honest narrative a decision-maker can act on.

Each lecture features two to three authentic datasets, whose analysis is demonstrated live in class. Examples range across consumer database mining, internet and social media tracking, asset pricing, network analysis, healthcare, sports analytics, and text mining, so the methods always arrive attached to a real business decision. For example, we will analyze a social network of the Medici family in Florence, decode who wrote the Federalist, build song recommendation systems, understand the demographics of Titanic survivors, or determine biomarkers for leukemia.

The curriculum spans topics from classical statistics (e.g. hypothesis-driven decisions), data science (dimensionality reduction) to modern machine learning techniques (e.g deep learning). It also explores cutting-edge advancements in generative AI. The course puts a particular emphasis on the analysis of text data in the context of both small and Large Language Models (LLM) that form a basis of popular text-generating systems. Techniques covered include large-scale testing and false discovery rates, modern regression and model choice, machine-learning based classification, network analysis, language and topic models, principal components, clustering, Bayesian analysis, deep learning, transformers and attention.

By the end of the course, students will be equipped to perform machine-supported intelligent data analysis and communicate findings effectively.

OPTIONAL PREREQUISITES
Students might benefit most from this course if they have had prior exposure to basic concepts in probability, such as random variables and normal distributions. That said, these foundational topics
are reviewed/introduced during the course, so students without a formal background in statistics can still succeed–though they may experience a steeper learning curve early on.

This course might appeal to students who have already taken Business Statistics (BUS 41000) and/or Advanced Business Statistics (BUS 41001). Another possibly useful prerequisite is Data Analysis with R and Python (BUS 32100) and Artificial Intelligence (BUS 32200).

Cannot enroll in BUSN 41215 if 41201 taken previously.

This course is designed for those with a strong interest in hands-on data analysis using real-world datasets, rather than purely abstract conversations.

POSSIBLE POSTREQUISITES

This course includes the key concepts and tools that data scientists find valuable in business environments, and it is also designed to act as a primer for continued study.
The course may be a useful prerequisite into deep-dive courses and/or field-specific variants such as Machine Learning in Finance (BUS 35137), Machine Learning (BUS 41204), Generative
Thinking (BUS 32210), Data Science for Marketing Decision Making (BUS 37105), Data-driven Marketing (BUS 37103), Causal Inference in Business Applications (BUS 41207).

Grades will be determined by homework (20%), a take-home midterm exam (45%), and a final project (35%). Late assignments, exams, or projects, will not be accepted.

 

Description and/or course criteria last updated: July 28 2026
SCHEDULE
  • Autumn 2025
    Section: 41215-01
    T 1:30 PM-4:30 PM
    Harper Center
    C01
    In-Person Only
  • Autumn 2025
    Section: 41215-81
    T 6:00 PM-9:00 PM
    Gleacher Center
    406
    In-Person Only

Data Intelligence (41215) - Rockova, Veronika>>

This course sits at the intersection of Data Science and Artificial Intelligence.

It is designed for future business leaders who want to engage with data directly rather than defer to the analysts in the room.

Its primary objectives are to:

  • become fluent in the language of “data science and AI”,
  • learn how to effectively communicate uncertainty,
  • cultivate the ability to craft compelling narratives grounded in data,
  • use data analytics for informed decision-making in uncertain environments,

  • develop skills to explore and make sense of large and messy data,

  • build and interpret (predictive) models with confidence,

  • learn to direct AI coding with Claude Code to produce a defensible data analysis.

Ideal for students preparing for careers in data-rich environments, the course emphasizes data storytelling. The goal is to learn how to turn data into a clear, honest narrative a decision-maker can act on.

Each lecture features two to three authentic datasets, whose analysis is demonstrated live in class. Examples range across consumer database mining, internet and social media tracking, asset pricing, network analysis, healthcare, sports analytics, and text mining, so the methods always arrive attached to a real business decision. For example, we will analyze a social network of the Medici family in Florence, decode who wrote the Federalist, build song recommendation systems, understand the demographics of Titanic survivors, or determine biomarkers for leukemia.

The curriculum spans topics from classical statistics (e.g. hypothesis-driven decisions), data science (dimensionality reduction) to modern machine learning techniques (e.g deep learning). It also explores cutting-edge advancements in generative AI. The course puts a particular emphasis on the analysis of text data in the context of both small and Large Language Models (LLM) that form a basis of popular text-generating systems. Techniques covered include large-scale testing and false discovery rates, modern regression and model choice, machine-learning based classification, network analysis, language and topic models, principal components, clustering, Bayesian analysis, deep learning, transformers and attention.

By the end of the course, students will be equipped to perform machine-supported intelligent data analysis and communicate findings effectively.

OPTIONAL PREREQUISITES
Students might benefit most from this course if they have had prior exposure to basic concepts in probability, such as random variables and normal distributions. That said, these foundational topics
are reviewed/introduced during the course, so students without a formal background in statistics can still succeed–though they may experience a steeper learning curve early on.

This course might appeal to students who have already taken Business Statistics (BUS 41000) and/or Advanced Business Statistics (BUS 41001). Another possibly useful prerequisite is Data Analysis with R and Python (BUS 32100) and Artificial Intelligence (BUS 32200).

Cannot enroll in BUSN 41215 if 41201 taken previously.

This course is designed for those with a strong interest in hands-on data analysis using real-world datasets, rather than purely abstract conversations.

POSSIBLE POSTREQUISITES

This course includes the key concepts and tools that data scientists find valuable in business environments, and it is also designed to act as a primer for continued study.
The course may be a useful prerequisite into deep-dive courses and/or field-specific variants such as Machine Learning in Finance (BUS 35137), Machine Learning (BUS 41204), Generative
Thinking (BUS 32210), Data Science for Marketing Decision Making (BUS 37105), Data-driven Marketing (BUS 37103), Causal Inference in Business Applications (BUS 41207).

Grades will be determined by homework (20%), a take-home midterm exam (45%), and a final project (35%). Late assignments, exams, or projects, will not be accepted.

 

Description and/or course criteria last updated: July 28 2026
SCHEDULE
  • Autumn 2025
    Section: 41215-01
    T 1:30 PM-4:30 PM
    Harper Center
    C01
    In-Person Only
  • Autumn 2025
    Section: 41215-81
    T 6:00 PM-9:00 PM
    Gleacher Center
    406
    In-Person Only