Learn pathway

How do I judge a market claim?

Samples, uncertainty, bias, testing and replication without statistical theatre.

38guides in sequence
  1. 01
    Reading the evidence · beginner

    Why sample size matters

    A perfect-looking rate can still be weak evidence when it rests on very few observations.

    Read guide
  2. 02
    Reading the evidence · beginner

    A result needs a baseline

    A result becomes meaningful only when it is compared with what would ordinarily happen under a relevant definition.

    Read guide
  3. 03
    Reading the evidence · beginner

    Uncertainty is part of the result

    A result is more useful when you can see the range of values still compatible with the available evidence.

    Read guide
  4. 04
    Reading the evidence · intermediate

    Backtesting without time travel

    A backtest replays an idea on old data. It is useful only when the replay respects what could have been known at the time.

    Read guide
  5. 05
    Reading the evidence · beginner

    Probability

    Probability represents uncertainty on a scale from impossible to certain under a defined model or evidence base. It describes possible outcomes, not a hidden guarantee about one event.

    Read guide
  6. 06
    Reading the evidence · beginner

    Chance and randomness

    Randomness is variation not predictably explained by the information or model being used. It can produce streaks, clusters and convincing patterns even when no durable signal exists.

    Read guide
  7. 07
    Reading the evidence · intermediate

    What is a p-value?

    A p-value asks how unusual the observed result would be if a particular chance-only model were true.

    Read guide
  8. 08
    Reading the evidence · intermediate

    What does statistical significance mean?

    Statistical significance means a result crossed a pre-chosen threshold under a particular test. It does not mean important or profitable.

    Read guide
  9. 09
    Reading the evidence · intermediate

    What is a confidence interval?

    A confidence interval shows a range of values compatible with the estimate under a stated statistical method.

    Read guide
  10. 10
    Reading the evidence · beginner

    False positives

    A false positive occurs when a test signals an effect or condition that is not present under the relevant truth definition. The rate depends on thresholds, prevalence and the testing process.

    Read guide
  11. 11
    Reading the evidence · beginner

    False negatives

    A false negative occurs when a test fails to detect an effect or condition that is present. Small samples, noisy measurement or strict thresholds can make detection unlikely.

    Read guide
  12. 12
    Reading the evidence · intermediate

    Multiple testing

    Multiple testing occurs when many hypotheses, groups, thresholds or outcomes are examined. Even valid individual tests can produce an impressive-looking winner by chance across the wider search.

    Read guide
  13. 13
    Reading the evidence · intermediate

    What is overfitting?

    Overfitting happens when a rule learns the accidents in old data so closely that it struggles with new data.

    Read guide
  14. 14
    Reading the evidence · intermediate

    Look-ahead bias

    Look-ahead bias lets information unavailable at a historical decision time influence that decision or its measured outcome. It gives the backtest a form of time travel.

    Read guide
  15. 15
    Reading the evidence · intermediate

    Data snooping

    Data snooping is repeated exploration of the same dataset until a favourable pattern is found, often without carrying the full search into the reported uncertainty.

    Read guide
  16. 16
    Reading the evidence · beginner

    In-sample testing

    In-sample testing evaluates an idea on data available for developing, selecting or tuning it. It shows fit to the workshop material, not independent validation.

    Read guide
  17. 17
    Reading the evidence · intermediate

    What does out of sample mean?

    Out-of-sample testing applies a frozen rule to data that did not help create or tune it.

    Read guide
  18. 18
    Reading the evidence · intermediate

    Prospective testing

    Prospective testing records a rule, inputs and expectation before future outcomes exist, then attaches those outcomes without rewriting the original record.

    Read guide
  19. 19
    Reading the evidence · intermediate

    Correlation and causation

    Correlation shows that variables move together under a sample; causation means changing one would change the other under specified conditions. Shared causes, selection and chance can create correlation without causation.

    Read guide
  20. 20
    Reading the evidence · intermediate

    Survivorship bias

    Survivorship bias occurs when analysis includes entities that remain visible while omitting those that failed, closed or left the dataset. The survivors can make history look safer or stronger.

    Read guide
  21. 21
    Reading the evidence · intermediate

    Selection bias

    Selection bias arises when inclusion in a sample depends on factors related to the outcome, making the observed group unrepresentative of the intended population.

    Read guide
  22. 22
    Reading the evidence · intermediate

    Publication bias

    Publication bias occurs when results are more likely to become visible because they are positive, striking or statistically significant. The available record then overstates effects.

    Read guide
  23. 23
    Reading the evidence · beginner

    Confirmation bias

    Confirmation bias is the tendency to seek, interpret and remember information in ways that support an existing belief while discounting disconfirming evidence.

    Read guide
  24. 24
    Reading the evidence · intermediate

    Regression to the mean

    Regression to the mean is the tendency for an extreme noisy observation to be followed by one closer to the typical level, even without a causal intervention.

    Read guide
  25. 25
    Reading the evidence · intermediate

    Independent observations

    Observations are independent when knowing one does not change the relevant probability distribution of another under the model. Market events often share firms, dates or shocks and are not fully independent.

    Read guide
  26. 26
    Reading the evidence · intermediate

    Effect size

    Effect size describes the magnitude of a difference or relationship in meaningful or standardised units. It complements uncertainty and helps distinguish detectable effects from useful ones.

    Read guide
  27. 27
    Reading the evidence · advanced

    Statistical power

    Statistical power is the probability that a specified test rejects its null when a specified alternative effect is true. It depends on effect size, sample, variability and threshold.

    Read guide
  28. 28
    Reading the evidence · intermediate

    Robustness

    Robustness is the degree to which a conclusion survives reasonable changes in assumptions, definitions, samples and methods. It is accumulated evidence, not one favourable alternative specification.

    Read guide
  29. 29
    Reading the evidence · intermediate

    Replication

    Replication repeats a study or test using the same method, independent implementation, new data or a new setting. Each form checks a different source of error.

    Read guide
  30. 30
    Reading the evidence · beginner

    What is a hypothesis?

    A hypothesis is a precise, testable statement about a relationship or outcome. It turns a broad observation into something that data could support, weaken or leave unresolved.

    Read guide
  31. 31
    Reading the evidence · intermediate

    Falsifiability

    Falsifiability means a claim permits some possible evidence to count against it. A claim that explains every outcome after the fact cannot be meaningfully tested.

    Read guide
  32. 32
    Reading the evidence · beginner

    Signal and noise

    Signal is repeatable information relevant to the question; noise is variation that obscures or imitates it under the chosen model. The separation is uncertain and context-dependent.

    Read guide
  33. 33
    Reading the evidence · intermediate

    Win rate and profitability

    Win rate is the proportion of positive outcomes; profitability depends on the size of wins and losses, costs, exposure and sequence. A strategy can win often and still lose money.

    Read guide
  34. 34
    Reading the evidence · beginner

    Mean, median and outliers

    The mean adds values and divides by their count; the median is the middle ordered value. Outliers and skew can separate them, so each describes a different centre.

    Read guide
  35. 35
    Reading the evidence · beginner

    Reading a distribution

    A distribution shows which values occurred or are possible and how frequently or plausibly they appear. Shape reveals spread, skew, clusters and tails hidden by one average.

    Read guide
  36. 36
    Reading the evidence · intermediate

    Sampling error

    Sampling error is the variation between a sample estimate and the population quantity caused by observing only part of the population. It differs from bias and measurement error.

    Read guide
  37. 37
    Reading the evidence · intermediate

    Pre-registering a test

    Pre-registration records the hypothesis, data, exclusions and analysis plan before outcomes are examined. It separates planned confirmation from later exploration and makes deviations visible.

    Read guide
  38. 38
    Reading the evidence · intermediate

    Data quality and missing information

    Data quality is fitness for a specific use across completeness, accuracy, consistency, timing, provenance and meaning. Clean formatting does not establish that observations represent the intended facts.

    Read guide
Return to the full Learn library