Methods of regression analysis

Last updated: November 01, 2022

Methods of regression analysis

Revision for finals delete as I go

Revision for finals delete as I go

DNA synthesis inhibitors: Fluoroquinolones
DNA synthesis inhibitors: Metronidazole
Antimetabolites: Sulfonamides and trimethoprim
Cell wall synthesis inhibitors: Penicillins
Cell wall synthesis inhibitors: Cephalosporins
Type III hypersensitivity
Type IV hypersensitivity
Type I hypersensitivity
Type II hypersensitivity
Thymus histology
Viral structure and functions
Wound healing
Pharmacodynamics: Drug-receptor interactions
Pharmacodynamics: Agonist, partial agonist and antagonist
Somatosensory pathways
Somatosensory receptors
Ascending and descending spinal tracts
Pyramidal and extrapyramidal tracts
Nervous system anatomy and physiology
Muscle spindles and golgi tendon organs
Sciatica
Anatomy of the cranial meninges and dural venous sinuses
Anatomy of the blood supply to the brain
Anatomy of the abdominal viscera: Blood supply of the foregut, midgut and hindgut
Blood histology
Anatomy of the brainstem
Optic pathways and visual fields
Anatomy and physiology of the eye
Cerebral circulation
Stroke: Clinical
Cranial nerves
Cranial nerves rap
Cranial nerve pathways
Pharmacokinetics: Drug absorption and distribution
Anatomy of the basal ganglia
Basal ganglia: Direct and indirect pathway of movement
Cerebellum
Anatomy of the cerebellum
Auditory transduction and pathways
Vestibular transduction
Complement system
Dementia and delirium: Clinical
Substance misuse and addiction: Clinical
Tricyclic antidepressants
Typical antipsychotics
Metabolic alkalosis
Seizures and epilepsy
Free radicals and cellular injury
Traumatic brain injury: Clinical
Concussion and traumatic brain injury
Schizophrenia
Major depressive disorder
Bone remodeling and repair
Lambert-Eaton myasthenic syndrome
Pediatric orthopedic conditions: Clinical
Bone histology
Skin histology
Colon histology
Stomach histology
Cartilage histology
Ovary histology
Paget disease of bone
Non-steroidal anti-inflammatory drugs
Osteomalacia and rickets
Osgood-Schlatter disease (traction apophysitis)
Slipped capital femoral epiphysis
Developmental dysplasia of the hip
Rotator cuff tear
Thoracic outlet syndrome
Klumpke paralysis
Erb-Duchenne palsy
Carpal tunnel syndrome
Compartment syndrome
Osteomyelitis
Osteoporosis
Lordosis, kyphosis, and scoliosis
Osteoarthritis
Rheumatoid arthritis
Gout
Ankylosing spondylitis
Muscular dystrophy
Inclusion body myopathy
Dermatomyositis
Fibromyalgia
Myasthenia gravis
Opioid agonists, mixed agonist-antagonists and partial agonists
Antigout medications
Osteoporosis medications
Brachial plexus
Anatomy of the brachial plexus
Rheumatoid arthritis: Clinical
Introduction to biostatistics
Types of data
Probability
Mean, median, and mode
Range, variance, and standard deviation
Standard error of the mean (Central limit theorem)
Normal distribution and z-scores
Paired t-test
Two-sample t-test
Hypothesis testing: One-tailed and two-tailed tests
One-way ANOVA
Two-way ANOVA
Repeated measures ANOVA
Correlation
Methods of regression analysis
Linear regression
Logistic regression
Spearman's rank correlation coefficient
Mann-Whitney U test
Kappa coefficient
Chi-squared test
Fisher's exact test
Kaplan-Meier survival analysis
Type I and type II errors
Sensitivity and specificity
Positive and negative predictive value
Test precision and accuracy
Incidence and prevalence
Relative and absolute risk
Odds ratio
Attributable risk (AR)
Mortality rates and case-fatality
DALY and QALY
Direct standardization
Indirect standardization
Study designs
Ecologic study
Cross sectional study
Case-control study
Cohort study
Randomized control trial
Clinical trials
Sample size
Placebo effect and masking
Disease causality
Selection bias
Information bias
Confounding
Interaction
Bias in interpreting results of clinical studies
Bias in performing clinical studies
Prevention
Anatomy of the pelvic girdle
Anatomy of the pelvic cavity
Arteries and veins of the pelvis
Anatomy of the male reproductive organs of the pelvis
Nerves and lymphatics of the pelvis
Anatomy clinical correlates: Male pelvis and perineum
Anatomy of the breast
Anatomy of the female urogenital triangle
Anatomy clinical correlates: Breast
Development of the reproductive system
Prostate gland histology
Penis histology
Testis, ductus deferens, and seminal vesicle histology
Mammary gland histology
Fallopian tube and uterus histology
Cervix and vagina histology
Anatomy and physiology of the male reproductive system
Testosterone
Anatomy and physiology of the female reproductive system
Estrogen and progesterone
Menstrual cycle
Menopause
Pregnancy
Stages of labor
Oxytocin and prolactin
Breastfeeding
Benign prostatic hyperplasia
Prostate cancer
Erectile dysfunction
Amenorrhea
Polycystic ovary syndrome
Premature ovarian failure
Endometritis
Endometriosis
Pelvic inflammatory disease
Preeclampsia & eclampsia
Placenta previa
Placental abruption
Potter sequence
Postpartum hemorrhage
Congenital cytomegalovirus (NORD)
Miscarriage
Ectopic pregnancy
Fetal alcohol syndrome
PDE5 inhibitors
Adrenergic antagonists: Alpha blockers
Estrogens and antiestrogens
Progestins and antiprogestins
Aromatase inhibitors
Uterine stimulants and relaxants
Newborn management: Clinical
Neonatal jaundice: Clinical
Human development days 1-4
Human development days 4-7
Human development week 2
Human development week 3
Ectoderm
Mesoderm
Endoderm
Development of the placenta
Development of the fetal membranes
Development of twins
Hedgehog signaling pathway
Development of the digestive system and body cavities
Development of the umbilical cord
Development of the cardiovascular system
Fetal circulation
Blood pressure, blood flow, and resistance
Pressures in the cardiovascular system
Resistance to blood flow
Compliance of blood vessels
Microcirculation and Starling forces
Stroke volume, ejection fraction, and cardiac output
Cardiac contractility
Frank-Starling relationship
Cardiac preload
Cardiac afterload
Law of Laplace
Cardiac cycle
Cardiac work
Pressure-volume loops
Changes in pressure-volume loops
Action potentials in myocytes
Action potentials in pacemaker cells
Cardiac conduction system
Cardiac conduction velocity
ECG basics
ECG normal sinus rhythm
ECG intervals
ECG axis
ECG rate and rhythm
ECG cardiac infarction and ischemia
ECG cardiac hypertrophy and enlargement
Baroreceptors
Chemoreceptors
ACE inhibitors, ARBs and direct renin inhibitors
Calcium channel blockers
Adrenergic antagonists: Beta blockers
cGMP mediated smooth muscle vasodilators
Class I antiarrhythmics: Sodium channel blockers
Class II antiarrhythmics: Beta blockers
Class III antiarrhythmics: Potassium channel blockers
Class IV antiarrhythmics: Calcium channel blockers and others
Miscellaneous lipid-lowering medications
Positive inotropic medications
Anatomy of the larynx and trachea
Bones and joints of the thoracic wall
Vessels and nerves of the thoracic wall
Anatomy of the lungs and tracheobronchial tree
Kidney histology
Anatomy of the diaphragm
Anatomy clinical correlates: Thoracic wall
Anatomy clinical correlates: Pleura and lungs
Development of the respiratory system
Nasal cavity and larynx histology
Trachea and bronchi histology
Bronchioles and alveoli histology
Respiratory system anatomy and physiology
Reading a chest X-ray
Lung volumes and capacities
Anatomic and physiologic dead space
Alveolar surface tension and surfactant
Ventilation
Zones of pulmonary blood flow
Regulation of pulmonary blood flow
Pulmonary shunts
Ventilation-perfusion ratios and V/Q mismatch
Airflow, pressure, and resistance
Gas exchange in the lungs, blood and tissues
Alveolar gas equation
Diffusion-limited and perfusion-limited gas exchange
Oxygen binding capacity and oxygen content
Oxygen-hemoglobin dissociation curve
Carbon dioxide transport in blood
Upper respiratory tract infection
Congenital pulmonary airway malformation
Acute respiratory distress syndrome
Emphysema
Asthma
Bronchiectasis
Chronic bronchitis
Cystic fibrosis
Alpha 1-antitrypsin deficiency
Restrictive lung diseases
Pneumonia
Pancoast tumor
Pleural effusion
Pneumothorax
Pulmonary embolism
Pulmonary edema
Pulmonary hypertension
Sleep apnea
Antihistamines for allergies
Bronchodilators: Beta 2-agonists and muscarinic antagonists
Bronchodilators: Leukotriene antagonists and methylxanthines
Acromegaly
Pituitary adenomas and pituitary hyperfunction: Clinical
Protein synthesis inhibitors: Aminoglycosides
Antituberculosis medications
Miscellaneous cell wall synthesis inhibitors
Protein synthesis inhibitors: Tetracyclines
Miscellaneous protein synthesis inhibitors
Mechanisms of antibiotic resistance
Integrase and entry inhibitors
Nucleoside reverse transcriptase inhibitors (NRTIs)
Protease inhibitors
Hepatitis medications
Non-nucleoside reverse transcriptase inhibitors (NNRTIs)
Neuraminidase inhibitors
Herpesvirus medications
Azoles
Echinocandins
Miscellaneous antifungal medications
Antimalarials
Light microscopy and staining methods
Cardiac muscle histology
Artery and vein histology
Arteriole, venule and capillary histology
Pituitary gland histology
Pancreas histology
Eye and ear histology
Gallbladder histology
Esophagus histology
Small intestine histology
Liver histology
Spleen histology
Lymph node histology
Skeletal muscle histology
Ureter, bladder and urethra histology

Transcript

Watch video only

Content Reviewers

There are four basic types of statistical analyses commonly used in epidemiological research, and the analysis you pick depends on two main criteria.

The first criterion is the type of data you have, which can be either individual data or binned data, which is also called group data.

So, for example, let’s say we want to know how many people out of 100 people developed lung cancer the past 5 years.

With individual data, we have information about each person, so we can tell whether or not each of the 100 people developed lung cancer.

So let’s say that 6 people developed lung cancer. If we have individual data, we can look at the individual characteristics for each of those 6 people, like their sex, age, race, or past history of migraines, and we can compare them to the people that didn’t developed lung cancer.

On the other hand, if we have group data, we don’t actually know which specific individuals out of the 100 people developed lung cancer.

So even though we know that 6 people had them, we don’t know which 6 people they were or any of their individual characteristics.

The second criterion is the type of outcome or y-variable you’re measuring, which can be either quantitative, categorical, or time to event.

Quantitative variables have a numeric value, like a person’s forced expiratory volume, which is the total amount of air, in liters, that a person can exhale in a single forced breath.

A very fit person might have an FEV of 5, while a less fit person might have an FEV of 3.

On the other hand, categorical variables have distinct levels.

For example, we could use a categorical variable to characterize if a person was diagnosed with lung cancer in the past five years or if they were not.

And finally, time to event variables describe how long a person was followed before the event or outcome occurred.

For example, if we started following a person at age 50 and they developed lung cancer at age 53, then their time to event would be 3 years.

Now, one of the simplest and most widely used types of analysis is linear regression.

Linear regression uses individual data, and the outcome variable is always quantitative, while the exposure variable can be either categorical or quantitative.

For example, let’s say we want to figure out if there’s an association between the number of cigarettes smoked and FEV, so we ask 100 people how many cigarettes they smoke in a day and then measure each person’s FEV. In this study, the exposure is the number of cigarettes, so it’s quantitative, and the outcome is FEV, which is also quantitative.

Typically, we use statistical software to calculate the linear equation, and the software will provide b0 and b1, which are two numbers we can then plug into the equation y-hat = b0 + b1x1.

Y-hat is the estimated value for the outcome variable, which in this case is FEV, and x1 is the value of the exposure variable, so in this case that’s the number of cigarettes a person smokes.

So let’s say the software gives us a b0 of 4 and a b1 of negative 0.1, so the equation is y-hat equals 4 minus 0.1 times x1.

Now, b1 is the most important number for interpretation because it tells us the effect size, or how much the outcome variable changes for every one-unit increase in the exposure variable.

For example, a b1 of negative 0.1 means that, on average, the FEV will decrease by 0.1 liters per second for every one additional cigarette smoked per day.

One important thing to know is that linear regression can be used in any type of study design as long as the two criteria of individual data and quantitative outcome variable are met.

The next type of statistical analysis is logistic regression. Logistic regression uses individual data, and the outcome variable is always categorical while the exposure variables can be either categorical or quantitative.

For example, let’s say we want to figure out if smoking more cigarettes increases the chance of lung cancer between the ages of 55-64. So, we follow a hundred 55-year-olds that smoke and a hundred 55-year-olds that don’t smoke for 10 years, and compare how many of them develop lung cancer.

In this example, the exposure variable is whether or not a person smokes cigarettes, so it’s categorical; and the outcome variable is whether or not the person develops lung cancer, so it’s also categorical.

And more specifically, because there are only two levels for each variable, they’re called binary categorical variables.

Now, like linear regression, the statistical software will give us b0 and b1, and we can plug them into the same equation of y-hat = b0 + b1x1, but the interpretation of the beta-coefficients are different.

In logistic regression, the beta-coefficients represent the log-odds of the outcome occurring.

For example, let’s say the software gives us a b0 of 0.05 and a b1 of 1.9, so the equation for the line would be y-hat equals 0.05 plus 1.9 times x1.

If we only look at b1, the effect size, it tells us how much the log-odds of the outcome variable changes for the unexposed group, or the non-smokers, versus the exposed group, or the smokers.

So, a b1 of 1.9 means that, on average, the log-odds of developing lung cancer for smokers is 1.9 times the log-odds of developing lung cancer for non-smokers.

Since the log-odds can be a confusing interpretation, we can also convert these numbers to regular odds by exponentiating them by a base of e.

For example, e to the 1.9 equals 6.7, so the odds of developing lung cancer for smokers is 6.7 times the odds of developing lung cancer for non-smokers.

Logistic regression can be used for any type of study, but the interpretation changes slightly depending on the study design.

Our example was a longitudinal cohort study, because we had a group of exposed individuals—those are the ones that smoked—and a group of unexposed individuals—those are the ones that didn’t smoke—and followed them over time.

This type of study design allows you to measure the incidence or the risk, which is the number of new cases that occur over a certain period of time.

Using logistic regression, we then calculate what’s called the risk odds ratio.

On the other hand, logistic regression can also be used in case-control studies, which is where you compare the history of two groups of people—those that have a certain outcome, called cases, and those that don’t have a certain outcome, called controls—to see if they’ve been exposed to different things.

So, for example, we could’ve looked at 100 people that had lung cancer, which would be the cases, and 100 people that don’t have lung cancer, which would be the controls, and then compare how many people in each group smoked cigarettes in the past ten years.

Now, in case-control studies, we can’t measure the incidence, since we’re selecting people that already have the outcome.

Instead, we’re measuring the prevalence, or the number of people that already smoked cigarettes before we started measuring them.

In case-control studies, we can use logistic regression to then calculate the prevalence odds ratio.

Key Takeaways

There are a variety of methods of regression analysis, each with its own strengths and weaknesses. The most commonly used methods are linear regression, logistic regression, and Poisson regression.

Linear regression is used when the data is assumed to be linear in nature. Logistic regression is used when the data is assumed to be binary (e.g., success/failure, yes/no), while Poisson regression is used when the data follows a Poisson distribution, and is used for modeling count data.