Access

You are not currently logged in.

Access your personal account or get JSTOR access through your library or other institution:

login

Log in to your personal account or through your institution.

A Two-Stage Regression Model for Epidemiological Studies with Multivariate Disease Classification Data

Nilanjan Chatterjee
Journal of the American Statistical Association
Vol. 99, No. 465 (Mar., 2004), pp. 127-138
Stable URL: http://www.jstor.org/stable/27590359
Page Count: 12
  • Download ($14.00)
  • Cite this Item
A Two-Stage Regression Model for Epidemiological Studies with Multivariate Disease Classification Data
Preview not available

Abstract

Polytomous logistic regression is commonly used to analyze epidemiological data with disease subtype information. In this approach effects of exposures on different disease subtypes are studied through separate exposure odds ratios comparing different case groups to the common control group. This article considers the situation where disease subtypes can be defined using multiple characteristics of a disease. For efficient analysis of such data, a two-stage modeling approach is proposed. At the first stage, a standard polytomous logistic regression model is considered for all possible distinct disease subtypes that can be defined by the cross-classification of the different disease characteristics. At the second stage, the exposure odds ratio parameters for the first-stage disease subtypes are further modeled in terms of the defining characteristics of the subtypes. When the total number of first-stage disease subtypes is small, standard maximum likelihood methods can be used for inference in the proposed model. For dealing with a large number of disease subtypes, a novel semiparametric pseudo-conditional-likelihood approach is proposed that does not require any model assumption about the baseline probabilities for the different disease subtypes. This article develops the asymptotic theory for the estimator and studies its small-sample properties using simulation experiments. The proposed method is applied to study the effect of fiber on the risk of various forms of colorectal adenoma using data available from a large screening study, the Prostate, Lung, Colorectal and Ovarian Cancer (PLCO) Screening Trial.

Page Thumbnails

  • Thumbnail: Page 
127
    127
  • Thumbnail: Page 
128
    128
  • Thumbnail: Page 
129
    129
  • Thumbnail: Page 
130
    130
  • Thumbnail: Page 
131
    131
  • Thumbnail: Page 
132
    132
  • Thumbnail: Page 
133
    133
  • Thumbnail: Page 
134
    134
  • Thumbnail: Page 
135
    135
  • Thumbnail: Page 
136
    136
  • Thumbnail: Page 
137
    137
  • Thumbnail: Page 
138
    138