Here are the steps of a Cohort Study explained in the simplest possible way:
Steps of a Cohort Study (Think of it as a Story)
Step 1 - Select Your Study Subjects (The Cohort)
Who do you pick?
Pick people who are currently healthy - they should NOT have the disease you are studying yet.
You divide them into two groups:
- 🟥 Exposed group - people who have the risk factor (e.g., smokers)
- 🟩 Non-exposed group - people without the risk factor (e.g., non-smokers)
Example: You pick 1,000 factory workers. 500 work near asbestos (exposed), 500 work in the office (not exposed). All 1,000 are currently healthy.
Step 2 - Obtain Data on Exposure
How do you confirm who is exposed?
Before the study starts, you collect information about their exposure:
- Interviews
- Medical records
- Special tests or examinations
- Environmental surveys
You also classify the level of exposure:
- Simply: Exposed vs. Not exposed
- Or in detail: Low exposure / Medium / High exposure (for dose-response analysis)
Example: You measure how many hours per day each worker is near asbestos dust.
Step 3 - Select a Comparison Group
Who do you compare the exposed group against?
Three options:
| Option | Meaning | Example |
|---|
| Internal comparison | Compare subgroups within your own cohort | Heavy smokers vs. light smokers |
| External comparison | Compare your exposed group with an outside non-exposed group | Radiologists vs. general population doctors |
| General population rate | Compare with national/regional disease rates | Compare your factory workers' cancer rate with national cancer rate |
Step 4 - Follow Up the People Over Time
This is where you WAIT and WATCH
You follow your two groups for months or years and track:
- Who develops the disease?
- Who stays healthy?
- Who dropped out? (loss to follow-up)
How do you follow them?
- Regular medical check-ups
- Hospital/physician records
- Death certificates
- Phone calls, questionnaires, home visits
Example: You track all 1,000 workers every year for 10 years and record who develops lung cancer.
Step 5 - Analysis (The Maths Part)
Now you count and calculate.
First, fill in the 2×2 table:
| Disease ✅ | No Disease ❌ | Total |
|---|
| Exposed | a | b | a+b |
| Not Exposed | c | d | c+d |
| Total | a+c | b+d | a+b+c+d |
Then calculate:
A) Incidence Rate (IR)
How many people got the disease per 1000?
- IR in exposed = a ÷ (a+b) × 1000
- IR in non-exposed = c ÷ (c+d) × 1000
B) Relative Risk (RR)
How many times MORE likely is the exposed group to get the disease?
RR = IR in exposed ÷ IR in non-exposed
| RR value | Meaning |
|---|
| RR = 1 | Exposure makes NO difference |
| RR > 1 | Exposure INCREASES the risk |
| RR < 1 | Exposure is PROTECTIVE |
Example: RR = 8 means smokers are 8 times more likely to get lung cancer than non-smokers.
C) Attributable Risk (AR)
How much of the disease is directly CAUSED by the exposure?
AR = IR in exposed - IR in non-exposed
Example: If 300 per lakh smokers get lung cancer and 10 per lakh non-smokers get it, AR = 290 per lakh. That 290 is directly attributable to smoking.
D) Population Attributable Risk (PAR)
What proportion of ALL disease cases in the population is due to this exposure?
PAR = (Incidence in population - Incidence in non-exposed) ÷ Incidence in population × 100
Useful for public health decisions - tells you how much disease would disappear if you removed the exposure.
Summary - The 5 Steps at a Glance
Step 1 → SELECT healthy people, divide into Exposed vs. Not Exposed
Step 2 → COLLECT data on their exposure level
Step 3 → CHOOSE a comparison group
Step 4 → FOLLOW UP over time, track who gets sick
Step 5 → ANALYSE using 2×2 table → calculate IR, RR, AR, PAR
Memory Tip
"S-O-S-F-A"
Select subjects → Obtain exposure data → Select comparison group → Follow up → Analyse
This follows the same order as your textbook notes!