← Clinical Trials
Relapsing-Remitting MS Phase 3 Completed NCT00530348

CARE-MS I: Complete Statistical Analysis of Alemtuzumab in Relapsing-Remitting Multiple Sclerosis

An independent statistical analysis of the randomized phase 3 CARE-MS I trial comparing alemtuzumab with interferon beta-1a in participants with relapsing-remitting multiple sclerosis, with emphasis on sustained accumulation of disability, annualized relapse rate, time-to-event analysis, and longitudinal outcomes.

ClinicalTrials.gov record  ·  Phase 3  ·  Enrollment 581  ·  Completed
Scope of this record

This page separates reported trial results from statistical interpretation. Numerical results, endpoint definitions, analysis populations, and statistical methods on this page are restricted to the trial data posted on ClinicalTrials.gov for CARE-MS I.

Registry note: This page provides an independent statistical analysis and educational interpretation of publicly reported results. ClinicalTrials.gov provides the official trial registry record.

1. Trial at a Glance

CARE-MS I was a randomized, parallel, single-masked phase 3 treatment trial comparing alemtuzumab with interferon beta-1a in relapsing-remitting multiple sclerosis. The registry reports 581 participants, two study arms, two primary endpoints, and six posted statistical analyses.

581
Enrollment
Randomized trial
2
Study arms
Parallel design
0.70
SAD HR
95% CI 0.40–1.23
0.45
Relapse-rate ratio
95% CI 0.32–0.63
FeatureCARE-MS I
Trial nameCARE-MS I
Brief titleComparison of Alemtuzumab and Rebif® Efficacy in Multiple Sclerosis, Study One
PhasePhase 3
ConditionMultiple Sclerosis, Relapsing-Remitting
AllocationRandomized
Design modelParallel
MaskingSingle
Primary purposeTreatment
Enrollment581.0
InterventionsAlemtuzumab; Interferon beta-1a
Primary endpoints2
Outcome measures posted6
Statistical analyses posted6
Lead sponsorGenzyme, a Sanofi Company
Sponsor typeIndustry
Trial statusCompleted
Study start2007-08
Primary completion2011-04
ClinicalTrials.govNCT00530348

2. Clinical Question

The trial evaluates two related efficacy questions in relapsing-remitting multiple sclerosis: whether treatment assignment is associated with the occurrence of sustained accumulation of disability over up to 2 years, and whether it is associated with the annualized rate of relapse over up to 2 years.

Population

Participants with Multiple Sclerosis, Relapsing-Remitting, as specified in the registry record.

Intervention

Alemtuzumab, classified in the registry as a biological intervention.

Comparator

Interferon beta-1a, classified in the registry as a biological intervention.

Primary questions

How do alemtuzumab and interferon beta-1a compare with respect to sustained accumulation of disability and annualized relapse rate through up to 2 years?

3. Trial Design

01
Randomize581.0 enrolled
02
Parallel arms2 interventions
03
TreatmentAlemtuzumab vs interferon beta-1a
04
AssessDisability and relapse
05
AnalyzeUp to 2 years
Allocation
Randomized allocation was used to compare the two treatment groups.
Design model
The registry identifies the study as a parallel-design trial.
Masking
The registry describes the trial as single-masked.
Primary purpose
Treatment.
ARM 1

Alemtuzumab

  • Biological intervention.
  • Compared with interferon beta-1a.
  • Primary efficacy analyses used the FAS population.
ARM 2

Interferon beta-1a

  • Biological intervention.
  • Compared with alemtuzumab.
  • Primary efficacy analyses used the FAS population.
What the registry data do not establish: The ClinicalTrials.gov record does not report a crossover design, factorial structure, a non-inferiority margin, Bayesian methodology, or an interim-analysis procedure. Those design features are therefore not treated as components of the CARE-MS I statistical analysis on this page.

4. Endpoints

The registry lists two primary endpoints. Both were evaluated over up to 2 years, but they require different statistical frameworks because one is a time-to-event disability endpoint and the other is a recurrent-event rate endpoint.

EndpointTime frameEndpoint typePrimary effect measure
Percentage of Participants With Sustained Accumulation of Disability (SAD) Up to 2 years Binary / time-to-event analysis Hazard ratio
Annualized Relapse Rate Up to 2 years Count / rate Rate ratio

Sustained Accumulation of Disability

The registry describes EDSS as an ordinal scale in half-point increments that quantifies disability in participants with MS. It assesses 7 functional systems—visual, brainstem, pyramidal, cerebellar, sensory, bowel/bladder, and cerebral—as well as ambulation. The EDSS total score ranges from 0, representing a normal neurological examination, to 10, representing death due to MS.

Annualized Relapse Rate

The registry defines a relapse as new neurological symptoms or worsening of previous neurological symptoms with an objective change on neurological examination, attributable to multiple sclerosis, lasting for at least 48 hours, present at normal body temperature, and preceded by at least 30 days of clinical stability.

The primary endpoint is the Annualized Relapse Rate over up to 2 years. The statistical analysis posted for the endpoint reports a rate ratio estimated using proportional means regression with robust variance estimation and adjustment for geographic region.

5. Statistical Analysis Populations

PopulationDefinition reported in trial dataRole
FAS All randomized participants who received at least 1 dose of study drug. Primary and secondary efficacy analyses
EDSS analysis subset Subset of the FAS population with EDSS assessment at both baseline and end. Change from baseline in EDSS
MSFC analysis subset Subset of the FAS population with MSFC score assessment at baseline; the registry-reported analysis text describes the Year 2 change analysis. Change from baseline in MSFC
MRI-T2 analysis subset Subset of the FAS population with T2 volume assessment at both baseline and Year 2. Percent change in MRI-T2 hyperintense lesion volume

The use of the FAS population means that the principal efficacy comparisons were not restricted to participants who completed the full planned observation period. However, the ClinicalTrials.gov record also show that some continuous or imaging outcomes were evaluated in subsets of the FAS defined by availability of the relevant assessments.

6. Statistical Methodology

Cox proportional-hazards regression for sustained accumulation of disability

The primary SAD analysis used a Cox proportional-hazards regression model with robust variance estimation. Treatment group and geographic region were included as covariates.

Primary SAD model
h(t | X) = h0(t) exp(β1Treatment + β2Region)

The hazard ratio for treatment is obtained by exponentiating the treatment coefficient. The model compares the instantaneous event rate over follow-up rather than simply comparing the percentages observed at one fixed time point.

Proportional means regression for annualized relapse rate

The annualized relapse-rate analysis used a proportional means regression model with robust variance estimation and covariate adjustment for geographic region. The reported effect measure was a rate ratio.

Rate-ratio interpretation
Rate ratio = estimated relapse rate under alemtuzumab relative to interferon beta-1a

A rate ratio below 1 indicates a lower estimated relapse rate in the numerator group under the comparison specified in the registry analysis. It is a comparison of rates, not a percentage of participants who relapsed.

Wei-Lachin analysis for repeated measures

The EDSS and MSFC change analyses were reported as using the Wei-Lachin method for non-parametric analysis of repeated measures. The registry reports the Wei-Lachin method.

This distinction matters. The registry method field should be reported as the actual method used in the analysis record rather than silently replacing it with a generic mixed-model label. The underlying statistical problem is longitudinal: measurements from the same participant are related, so treating every repeated observation as independent would generally be inappropriate.

Ranked ANCOVA for MRI-T2 lesion volume

The MRI-T2 endpoint used ranked ANCOVA with covariate adjustment for geographic region and baseline T2 lesion volume. The registry reports this analysis as a ranked ANCOVA.

ANCOVA combines a treatment-group comparison with adjustment for one or more baseline covariates. Here, baseline T2 lesion volume is particularly relevant because the endpoint is a change or percent change in lesion volume and baseline disease burden can influence the follow-up measurement.

Sequential analysis of secondary endpoints

The analysis notes state that secondary endpoints were analyzed sequentially as: proportion of participants relapse free at Year 2, change from baseline in EDSS, percent change from baseline in MRI-T2 hyperintense lesion volume at Year 2, and acquisition of disability measured by MSFC. The ClinicalTrials.gov record does not specify a formal multiplicity-adjustment procedure for this sequence, so no additional claim about familywise error control is made here.

7. Primary Results: Sustained Accumulation of Disability

The primary SAD endpoint was analyzed in the FAS population using Cox proportional-hazards regression. The comparison was between interferon beta-1a and alemtuzumab, with treatment group and geographic region as covariates and robust variance estimation.

Hazard ratio for sustained accumulation of disability

0.70

95% CI: 0.40–1.23   ·   P = 0.2173

Time frame: up to 2 years  ·  FAS population

Primary endpointAnalysisEffect measureEstimate95% CIP-value
Percentage of Participants With Sustained Accumulation of Disability (SAD) Cox proportional-hazards regression Hazard ratio 0.70 0.40–1.23 0.2173
Clinical Biostats interpretation

The estimated hazard ratio of 0.70 indicates that, under the fitted Cox model, the estimated instantaneous hazard of sustained accumulation of disability for the treatment comparison was 0.70 times the comparator hazard. Expressed as a simple relative interpretation, an HR of 0.70 corresponds to a 30% lower estimated hazard, because 1 − 0.70 = 0.30.

This does not mean that 30% fewer participants developed SAD, and it does not mean that every participant experienced a 30% reduction in personal risk. The hazard ratio is a model-based time-to-event measure.

The 95% confidence interval, 0.40–1.23, is relatively broad and crosses 1.00. It therefore includes values compatible with a lower hazard as well as values compatible with a higher hazard under the statistical model. The P-value of 0.2173 is evidence about compatibility with the tested statistical hypothesis; it is not a measure of the size or clinical importance of the estimated effect.

Because this was a Cox model, interpretation also depends on the proportional-hazards framework. The ClinicalTrials.gov record does not report a formal diagnostic for that assumption, so the HR should be understood as the reported model-based summary rather than as a guarantee that the hazard ratio was constant at every point in time.

8. Primary Results: Annualized Relapse Rate

The second primary endpoint was annualized relapse rate over up to 2 years. The FAS population was analyzed using proportional means regression with robust variance estimation and adjustment for geographic region.

Rate ratio for annualized relapse rate

0.45

95% CI: 0.32–0.63   ·   P < 0.0001

Time frame: up to 2 years  ·  FAS population

Primary endpointAnalysisEffect measureEstimate95% CIP-value
Annualized Relapse Rate Proportional means regression with robust variance estimation Rate ratio 0.45 0.32–0.63 <0.0001
Clinical Biostats interpretation

The rate ratio of 0.45 indicates that the estimated annualized relapse rate under the treatment comparison was 0.45 times the comparator rate. As a simple relative interpretation, this corresponds to a 55% lower estimated relapse rate because 1 − 0.45 = 0.55.

A rate ratio is not the same as a risk ratio. It does not say that 55% of participants avoided relapse, nor does it provide the absolute number of relapses experienced by an individual participant.

The 95% confidence interval of 0.32–0.63 describes uncertainty around the estimated rate ratio. Every value in the interval is below 1.00, indicating that the estimated rate remains below the comparator rate across the reported confidence interval.

The P-value of <0.0001 indicates strong statistical evidence against the null hypothesis specified for the comparison, but the P-value itself does not quantify the size of the treatment effect. The effect size is communicated by the rate ratio and its confidence interval.

The registry reports robust variance estimation and geographic-region adjustment. The ClinicalTrials.gov record does not provide the underlying participant-level relapse counts, exposure times, or model diagnostics, so the page does not attempt to reconstruct an absolute relapse-rate table.

9. Secondary Results: Relapse-Free at Year 2

The first listed secondary endpoint was the Percentage of Participants Who Were Relapse Free at Year 2. It was analyzed in the FAS population using Cox proportional-hazards regression with robust variance estimation and covariate adjustment for geographic region.

Hazard ratio for relapse-free analysis

0.45

95% CI: 0.33–0.61   ·   P < 0.0001

Time frame: Year 2  ·  FAS population

Secondary endpointMethodEstimate95% CIP-value
Percentage of Participants Who Were Relapse Free at Year 2 Cox proportional-hazards regression HR 0.45 0.33–0.61 <0.0001

The analysis note states that the secondary endpoints were analyzed sequentially, beginning with the proportion of participants relapse free at Year 2. The ClinicalTrials.gov record does not provide the corresponding arm-specific percentage values, so the hazard ratio is the principal numerical result reported here.

10. Secondary Results: EDSS Change

The endpoint Change From Baseline in Expanded Disability Status Scale (EDSS) Score at Year 2 was evaluated from baseline to Year 2. The analysis used the Wei-Lachin method for non-parametric analysis of repeated measures.

EDSS change from baseline

P = 0.4188

Time frame: Baseline, Year 2  ·  FAS subset with EDSS assessment at both baseline and end-of-study (Year 2).

Secondary endpointMethod as reportedAnalysis populationP-value
Change From Baseline in EDSS Score at Year 2 Wei-Lachin FAS subset with EDSS assessment at both baseline and end-of-study (Year 2). 0.4188

No effect estimate or confidence interval for this endpoint is included in the ClinicalTrials.gov record. The result therefore should not be converted into an assumed difference in EDSS or an assumed direction of effect.

Clinical Biostats interpretation

The P-value of 0.4188 is the reported inferential result for the repeated-measures EDSS analysis. It should not be interpreted as the probability that the treatments are equal, and it does not tell us the magnitude of any between-group difference.

Because the ClinicalTrials.gov record does not report an effect estimate or confidence interval, the statistically informative description is limited to the reported P-value and the analysis method. The absence of an effect estimate here is a limitation of the posted statistical result, not a reason to infer one.

11. Secondary Results: MSFC Change

The endpoint Change From Baseline in Multiple Sclerosis Functional Composite (MSFC) Score at Year 2 used the Wei-Lachin method for non-parametric analysis of repeated measures. The outcome was expressed as a Z-score.

MSFC change from baseline

P = 0.0115

Time frame: Baseline, Year 2  ·  FAS subset described in the registry analysis

Secondary endpointMethod as reportedOutcome unitP-value
Change From Baseline in MSFC Score at Year 2 Wei-Lachin Z-score 0.0115
Clinical Biostats interpretation

The reported P-value of 0.0115 indicates statistical evidence against the null hypothesis used for this repeated-measures comparison under the analysis framework reported by the registry.

However, the ClinicalTrials.gov record does not provide the estimated between-group change, its confidence interval, or the arm-specific mean or median changes. Therefore, the P-value should not be used to infer the magnitude of the difference or its clinical importance.

The use of a repeated-measures method also means that the statistical result reflects the longitudinal analysis rather than a simple comparison of two independent Year 2 measurements.

12. Secondary Results: MRI-T2 Hyperintense Lesion Volume

The endpoint Percent Change From Baseline in Magnetic Resonance Imaging Time Constant 2 (MRI-T2) Hyperintense Lesion Volume at Year 2 was analyzed using ranked ANCOVA.

MRI-T2 lesion-volume analysis

P = 0.3080

Time frame: Baseline, Year 2  ·  Ranked ANCOVA

Secondary endpointMethodCovariatesP-value
Percent Change From Baseline in MRI-T2 Hyperintense Lesion Volume at Year 2 Ranked ANCOVA Geographic region and baseline T2 lesion volume 0.3080
Clinical Biostats interpretation

The analysis adjusted for geographic region and baseline T2 lesion volume. Baseline adjustment is useful when a continuous outcome can be influenced by the participant's starting disease burden.

The reported P-value of 0.3080 is an inferential result, not an effect-size measure. The ClinicalTrials.gov record does not include an estimated treatment difference or confidence interval for percent change, so no magnitude or direction of treatment effect should be inferred from the P-value alone.

13. Statistical Methods Explained

Why was a Cox proportional-hazards model used for SAD?

Sustained accumulation of disability is naturally represented as a time-to-event outcome: participants can experience the event at different times, while others may not experience it during the observation period. Cox regression uses the timing of events and accommodates right censoring. The CARE-MS I analysis reported a hazard ratio rather than simply comparing the percentage with SAD at a single time point.

What does the SAD hazard ratio of 0.70 mean?

An HR of 0.70 means that the fitted model estimated an instantaneous hazard approximately 30% lower for the treatment comparison because 1 − 0.70 = 0.30. It does not mean that 30% fewer participants necessarily developed SAD, and it does not describe an individual's personal probability of disability accumulation.

Why is the confidence interval important?

The SAD 95% CI is 0.40–1.23. A confidence interval communicates precision around the estimated effect. Because the interval includes 1.00, the statistical uncertainty is compatible with both a lower and a higher hazard under the model. For the relapse-rate ratio, by contrast, the 95% CI is 0.32–0.63, entirely below 1.00.

Why use a rate ratio for annualized relapse rate?

Relapses are recurrent events and can occur at different rates across participants. A rate ratio compares the estimated rate of events between treatment groups. It should not be confused with a risk ratio, which compares the probability of experiencing an event over a specified period.

Why was geographic region included as a covariate?

The analyses posted on ClinicalTrials.gov explicitly report geographic-region adjustment for the primary SAD and annualized relapse-rate analyses, as well as several secondary analyses. Covariate adjustment can account for prespecified variation associated with a geographic factor while estimating the treatment comparison. The ClinicalTrials.gov record does not specify the number or definition of geographic regions, so no further structure is assumed.

Why use Wei-Lachin for EDSS and MSFC?

The registry reports the Wei-Lachin method for non-parametric analysis of repeated measures. Repeated observations from the same participant are correlated, so longitudinal methods account for the within-participant structure rather than treating repeated observations as unrelated data points.

What does a P-value tell us here?

A P-value quantifies how compatible the observed data are with the null hypothesis under the specified statistical model and testing framework. It does not measure the size of the treatment effect, the probability that a treatment works, or clinical importance. Those questions require the effect estimate, confidence interval, absolute outcomes, and clinical context.

14. Confidence Intervals and Effect Measures

The CARE-MS I primary analyses use two different relative effect measures. Understanding the distinction is essential because the numbers have different statistical meanings.

MeasureEndpointEstimate95% CIPlain-language interpretation
Hazard ratio Sustained accumulation of disability 0.70 0.40–1.23 Model-based comparison of instantaneous event hazards
Rate ratio Annualized relapse rate 0.45 0.32–0.63 Comparison of estimated relapse rates
Hazard ratio Relapse free at Year 2 0.45 0.33–0.61 Model-based time-to-event comparison
Relative effects are not absolute effects

A relative measure such as an HR or rate ratio can be statistically compelling while still leaving important questions about absolute event frequency unanswered. For example, an HR of 0.70 does not tell us how many participants experienced SAD without knowing the underlying event and censoring information.

Confidence interval versus P-value

The confidence interval provides both an estimate and information about its precision. The P-value provides a test-based measure of evidence against the null hypothesis. Neither one alone describes clinical importance.

15. Covariate Adjustment

Covariate adjustment appears repeatedly in the CARE-MS I statistical analyses. Geographic region was included in the primary Cox model for SAD and in the proportional means regression for annualized relapse rate. The MRI-T2 ranked ANCOVA additionally adjusted for baseline T2 lesion volume.

Geographic region

Used as a covariate in the primary SAD and annualized relapse-rate analyses and in the reported relapse-free analysis.

Baseline T2 lesion volume

Included in the ranked ANCOVA for percent change in MRI-T2 hyperintense lesion volume at Year 2.

Why adjust?

Adjustment can account for prespecified variation related to baseline or geographic characteristics when estimating the treatment comparison.

What adjustment does not do

Covariate adjustment does not turn a model estimate into an absolute treatment effect, and it does not eliminate all statistical uncertainty.

16. Secondary Endpoint Sequence and Multiplicity

The registry-reported analysis notes state that the secondary endpoints were analyzed sequentially. The reported sequence begins with the proportion of participants relapse free at Year 2, followed by change from baseline in EDSS, percent change from baseline in MRI-T2 hyperintense lesion volume at Year 2, and acquisition of disability measured by MSFC.

SequenceSecondary endpointReported methodReported P-value
1Percentage of Participants Who Were Relapse Free at Year 2Cox proportional-hazards regression<0.0001
2Change From Baseline in EDSS Score at Year 2Wei-Lachin0.4188
3Percent Change From Baseline in MRI-T2 Hyperintense Lesion Volume at Year 2Ranked ANCOVA0.3080
4Change From Baseline in MSFC Score at Year 2Wei-Lachin0.0115
Multiplicity caution: The fact that endpoints were analyzed sequentially does not, by itself, establish exactly how type I error was controlled. The ClinicalTrials.gov record does not specify an alpha-allocation or formal multiplicity-adjustment procedure for these secondary analyses. The reported P-values should therefore be presented as registry-reported results without adding an unsupported claim about familywise error control.

17. Safety

The ClinicalTrials.gov record reports serious adverse events by arm using affected participants divided by participants at risk. These counts are presented exactly as provided in the registry-derived data.

Treatment armSerious adverse events
Interferon Beta-1a 27/187
Alemtuzumab 69/376
Serious adverse events: affected / at risk
Interferon Beta-1a
27/187
Alemtuzumab
69/376

The denominators in this safety summary are the registry-reported at-risk counts and should not be replaced with the overall enrollment of 581.0. The registry-derived ClinicalTrials.gov record do not provide a formal statistical comparison, confidence interval, or P-value for serious adverse events, so none is inferred.

18. Reading the Two Primary Endpoints Together

One of the most useful statistical features of CARE-MS I is that the two primary endpoints represent different dimensions of disease activity.

Disability accumulation

SAD addresses a clinically important time-to-event outcome based on EDSS. Its reported HR was 0.70 with a 95% CI of 0.40–1.23.

Relapse activity

Annualized relapse rate addresses recurrent clinical events. Its reported rate ratio was 0.45 with a 95% CI of 0.32–0.63.

These measures should not be collapsed into a single number. The SAD analysis uses time-to-event methodology and asks when sustained disability accumulation occurs. The annualized relapse-rate analysis uses a rate model and asks about the frequency of relapse events over follow-up.

The difference between their effect measures is therefore not a contradiction. A hazard ratio and a rate ratio summarize different statistical quantities.

19. What the Primary P-values Do — and Do Not — Mean

SAD: P = 0.2173

The reported P-value of 0.2173 does not measure the magnitude of the SAD effect. The relevant magnitude is the HR of 0.70, while the confidence interval of 0.40–1.23 communicates its statistical precision.

Annualized relapse rate: P < 0.0001

The reported P-value of <0.0001 indicates strong statistical evidence against the null hypothesis under the reported model. The size of the estimated effect is communicated by the rate ratio of 0.45 and its 95% CI of 0.32–0.63.

A statistical result is not a clinical recommendation

The statistical analyses quantify differences between randomized treatment groups. Whether a treatment's overall clinical profile is appropriate for an individual patient requires considerations beyond these reported statistical estimates.

20. Limitations

21. Why This Trial Matters Statistically

CARE-MS I is a useful teaching example because the same randomized comparison requires several distinct statistical tools. The primary disability endpoint is analyzed with survival methods, relapse frequency is analyzed as a rate, and secondary continuous or longitudinal outcomes use methods designed for repeated measurements or baseline adjustment.

Statistical conceptHow it appears in CARE-MS I
RandomizationThe trial used randomized allocation to two parallel treatment groups.
FAS analysisThe primary and secondary efficacy analyses used a population consisting of randomized participants who received at least 1 dose of study drug.
Time-to-event analysisSAD and relapse-free-at-Year-2 analyses used Cox proportional-hazards regression.
Hazard ratioUsed for SAD and the relapse-free-at-Year-2 endpoint.
Rate ratioUsed for annualized relapse rate.
Covariate adjustmentGeographic region was included in several analyses; baseline T2 lesion volume was also used in the ranked ANCOVA.
Repeated measuresEDSS and MSFC changes were analyzed using the Wei-Lachin method for non-parametric repeated measures.
ANCOVARanked ANCOVA was used for percent change in MRI-T2 hyperintense lesion volume.
Confidence intervalsPrimary hazard-ratio and rate-ratio analyses included two-sided 95% confidence intervals.
P-valuesReported for all six posted statistical analyses.
Safety analysisSerious adverse events were reported as affected participants divided by participants at risk for each arm.

22. Statistical Concepts in This Trial

Learn more about the methods used in this trial:

23. Related Statistical Calculators

Use these calculator pathways to explore the statistical quantities that appear in CARE-MS I:

24. Trial Timeline

2007-08

Study start

The CARE-MS I trial began in August 2007 according to the ClinicalTrials.gov record.

Phase 3

Randomized treatment comparison

The phase 3 study used randomized allocation, a parallel design, and single masking to compare alemtuzumab with interferon beta-1a.

Up to 2 years

Primary efficacy assessment window

The two registered primary endpoints were evaluated over up to 2 years.

2011-04

Primary completion

The ClinicalTrials.gov record identifies April 2011 as the primary completion date.

25. What the Hazard Ratio Does — and Does Not — Mean

SAD hazard ratio

The reported SAD HR of 0.70 is a relative measure of instantaneous event hazard under the Cox model. A simple mathematical transformation gives 1 − 0.70 = 0.30, so the model estimate corresponds to a 30% lower estimated hazard.

It does not mean that 30% of participants avoided sustained disability accumulation, that 30% of participants were cured, or that every individual participant experienced exactly a 30% reduction in risk.

Rate ratio

The annualized relapse-rate ratio of 0.45 means the estimated relapse rate was 45% of the comparator rate under the reported model. Equivalently, 1 − 0.45 = 0.55, corresponding to a 55% lower estimated rate.

This is a rate comparison rather than a direct comparison of the probability that an individual participant experienced at least one relapse.

Confidence intervals

The SAD 95% CI of 0.40–1.23 communicates substantial uncertainty around the HR estimate. The annualized relapse-rate 95% CI of 0.32–0.63 communicates a narrower range entirely below 1.00. Confidence intervals should be interpreted together with the underlying model, endpoint definition, and analysis population.

26. Clinical Interpretation vs Statistical Interpretation

Statistical interpretation

The randomized comparison produced a SAD hazard ratio of 0.70 and an annualized relapse-rate ratio of 0.45. The two primary endpoints used different statistical models because they measure different types of outcomes.

Endpoint-specific interpretation

The relapse-rate analysis provides a rate comparison with a 95% CI of 0.32–0.63 and P < 0.0001, while the SAD analysis provides an HR of 0.70 with a 95% CI of 0.40–1.23 and P = 0.2173.

These findings should not be compressed into a single overall statistical score. The appropriate interpretation depends on the endpoint, effect measure, confidence interval, analysis population, model assumptions, and the distinction between efficacy outcomes and safety outcomes.

27. Limitations of Statistical Interpretation

28. Sources

Source restriction: The numerical trial results and methodological descriptions on this page are restricted to the registry-reported CARE-MS I trial data. The PubMed links above are provided as the linked publication records reported with the trial dataset; no additional numerical results from those publications have been incorporated into this page.

Continue with the statistical methods behind CARE-MS I

Explore the underlying survival-analysis, rate-ratio, covariate-adjustment, repeated-measures, confidence-interval, and P-value concepts used to interpret randomized clinical-trial results.

29. Record Summary

CARE-MS I provides a useful example of how a randomized phase 3 trial can require multiple statistical frameworks within the same efficacy program. The primary endpoint of sustained accumulation of disability was analyzed as a time-to-event outcome using Cox proportional-hazards regression, producing an HR of 0.70 with a 95% CI of 0.40–1.23 and P = 0.2173. Annualized relapse rate was analyzed using proportional means regression and produced a rate ratio of 0.45 with a 95% CI of 0.32–0.63 and P < 0.0001.

The secondary analyses illustrate additional statistical concepts: Cox regression for relapse-free status at Year 2, Wei-Lachin repeated-measures analysis for EDSS and MSFC, and ranked ANCOVA for MRI-T2 lesion-volume change. The reported P-values were <0.0001, 0.4188, 0.0115, and 0.3080, respectively. Importantly, several of these secondary results are reported without corresponding effect estimates or confidence intervals, so their magnitude should not be reconstructed from P-values alone.

The trial is therefore particularly useful for understanding a central principle of clinical-trial statistics: the endpoint determines the estimand, the estimand determines the appropriate effect measure, and the effect measure must be interpreted together with its confidence interval, analysis population, model assumptions, and design context.

Clinical Biostats methodology: A trial-results page should distinguish reported numerical evidence from statistical explanation. When the ClinicalTrials.gov record contains an effect estimate and confidence interval, both should be interpreted; when it provides only a P-value, the magnitude of an effect should not be invented.