• Graduate Programs
  • Research
  • Browse our Courses
  • Events
    • Events Calendar
    • Events Archive
    • Summer School
      • Applied Public Policy Evaluation
      • Deep Learning
      • Development Economics
      • Economics of Blockchain and Digital Currencies
      • Economics of Climate Change
      • The Economics of Crime
      • Foundations of Machine Learning with Applications in Python
      • From Preference to Choice: The Economic Theory of Decision-Making
      • Inequalities in Health and Healthcare
      • Marketing Research with Purpose
      • Markets with Frictions
      • Modern Toolbox for Spatial and Functional Data
      • Sustainable Finance
      • Tuition Fees and Payment
      • Business Data Science Summer School Program
    • Tinbergen Institute Lectures
    • 2026 Tinbergen Institute Opening Conference
    • Annual Tinbergen Institute Conference
  • News
  • Summer School
    • Applied Public Policy Evaluation
    • Deep Learning
    • Development Economics
    • Economics of Blockchain and Digital Currencies
    • Economics of Climate Change
    • The Economics of Crime
    • Foundations of Machine Learning with Applications in Python
    • From Preference to Choice: The Economic Theory of Decision-Making
    • Inequalities in Health and Healthcare
    • Marketing Research with Purpose
    • Markets with Frictions
    • Modern Toolbox for Spatial and Functional Data
    • Sustainable Finance
    • Tuition Fees and Payment
  • Alumni

Rutten-van Mölken, M.P., van Doorslaer, E.K. and van Vliet, R.C. (1994). Statistical analysis of cost outcomes in a randomized controlled clinical trial. Health Economics, 3(5):333--345.


  • Journal
    Health Economics

This paper suggests an approach to deal with an estimation problem which is often encountered in analyzing the longitudinal cost data gathered in a clinical trial. The source of that estimation problem is twofold: 1) a considerable number of missing data due to treatment-related withdrawal of severely affected patients with high health care costs in only one the treatment groups and 2) a heavily skewed cost distribution due to rare high-cost events. The approach is illustrated using data from a trial comparing 3 different drug regimes. In order to calculate costs per patient-year in case of selectively missing data we extrapolated the costs of patients with incomplete follow-up. Due to the skewness and the associated large variance in costs per patient-year, these costs cannot be analyzed using common parametric statistical methods relying on underlying normal distributions. A logarithmic transformation was performed to approximate a normal distribution, reduce the impact of extreme values and create similar size variances in the treatment groups. An ordinary least squares regression analysis of transformed data then standardized for differences in patient characteristics between the groups. For the retransformation, the so-called smearing estimate was used. This 'transformation-standardization-retransformation' approach enabled us to provide more consistent and efficient estimates of cost differences that were shown to be statistically significant and judged to be important.