The paper studies a mathematical model for how infectious diseases spread through a population where people are not all the same. In a standard epidemic model, everyone is assumed to have identical susceptibility to infection and identical ability to infect others. Here, each person carries individual characteristics, called covariates, that shape how likely they are to catch or transmit a disease. This matters practically because ignoring such differences can lead to badly wrong estimates of key public health quantities like the basic reproduction number (roughly, how many people one infected person typically infects) or the herd immunity threshold (the fraction of a population that needs to be immune to stop an outbreak).
To handle this mathematically, the authors represent the population as a large collection of interacting particles, where each particle is a person whose disease status evolves randomly over time according to rules that depend on their individual traits and on what is happening in the rest of the population. They then prove two major mathematical results. The first, called a Functional Law of Large Numbers, says that as the population grows large, the messy random behavior of this system can be well approximated by a smooth, deterministic description, making the model tractable. The second result, called propagation of chaos, shows that as the population size grows, individuals become approximately statistically independent of one another, even though they are interacting through the spread of infection. This is because each person's influence on any single other person becomes negligible when the population is huge.
A practical payoff of propagation of chaos is that it justifies writing down a simple product-form expression for the likelihood of observing a given dataset, meaning the probability of the data can be written as a product of individual contributions. This underpins a statistical method called Dynamic Survival Analysis, which allows researchers to estimate epidemic parameters from sparse, real-world data where full outbreak records are rarely available. Together, the theoretical results give rigorous mathematical foundations for using heterogeneous epidemic models in both analysis and data-driven inference.