Nonparametric Function Estimation for Clustered Da

整理文档很辛苦,赏杯茶钱您下走!

免费阅读已结束,点击下载阅读编辑剩下 ...

阅读已结束,您可以下载文档离线阅读编辑

资源描述

NonparametricFunctionEstimationforClusteredDataWhenthePredictorisMeasuredWithout/WithErrorXihongLinandRaymondJ.Carroll∗October4,1999AbstractWeconsiderlocalpolynomialkernelregressionwithasinglecovariateforclustereddatausingestimatingequations.Weassumethatatmostm∞observationsareavailableoneachcluster.Inthecaseofrandomregressors,withnomeasurementerrorinthepredictor,weshowthatitisgenerallythebeststrategytoignoreentirelythecorrelationstructurewithineachcluster,andinsteadtopretendthatallobservationsareindependent.Inthefurtherspecialcaseoflongitudinaldataonindividualswithfixedcommonobservationtimes,weshowthatequivalenttothepooleddataapproachisthestrategyoffittingseparatenonparametricregressionsateachobservationtimeandconstructinganoptimalweightedaverage.Wealsoconsiderwhathappenswhenthepredictorismeasuredwitherror.UsingtheSIMEXapproachtocorrectformeasurementerror,weconstructanasymptotictheoryforboththepooledandweightedaverageestimators.Surprisingly,forthesameamountofsmoothing,theweightedaverageestimatorstypicallyhavesmallervariancesthanthepoolingstrategy.WeapplytheproposedmethodstotheanalysisoftheAIDSCostsandServicesUtilizationSurvey.KEYWORDS:AIDS;Asymptoticbiasandvariance;Clustereddata;Efficiency;Errorsinvariables;Estimatingequations;Generalizedlinearmodels;Kernelregression;Longitudinaldata;Measure-menterror;Nonparametricregression;Paneldata;SIMEX.Shorttitle.NonparametricRegressionforClusteredData∗XihongLinisAssociateProfessor,DepartmentofBiostatistics,UniversityofMichigan,AnnArbor,MI48109-2029.HerresearchwassupportedbyagrantfromtheNationalCancerInstitute(CA–76404).RaymondJ.CarrollisUniversityDistinguishedProfessor,DepartmentsofStatisticsandBiostatistics&Epidemiology,TexasA&MUni-versity,CollegeStationTX77843–3143.HisresearchwassupportedbyagrantfromNationalCancerInstitute(CA–57030),andbytheTexasA&MCenterforEnvironmentalandRuralHealthviaagrantfromtheNationalInstituteofEnvironmentalHealthSciences(P30-ESO9106).NonparametricFunctionEstimationforClusteredDataWhenthePredictorisMeasuredWithout/WithErrorAbstractWeconsiderlocalpolynomialkernelregressionwithasinglecovariateforclustereddatausingestimatingequations.Weassumethatatmostm∞observationsareavailableoneachcluster.Inthecaseofrandomregressors,withnomeasurementerrorinthepredictor,weshowthatitisgenerallythebeststrategytoignoreentirelythecorrelationstructurewithineachcluster,andinsteadtopretendthatallobservationsareindependent.Inthefurtherspecialcaseoflongitudinaldataonindividualswithfixedcommonobservationtimes,weshowthatequivalenttothepooleddataapproachisthestrategyoffittingseparatenonparametricregressionsateachobservationtimeandconstructinganoptimalweightedaverage.Wealsoconsiderwhathappenswhenthepredictorismeasuredwitherror.UsingtheSIMEXapproachtocorrectformeasurementerror,weconstructanasymptotictheoryforboththepooledandweightedaverageestimators.Surprisingly,forthesameamountofsmoothing,theweightedaverageestimatorstypicallyhavesmallervariancesthanthepoolingstrategy.WeapplytheproposedmethodstotheanalysisoftheAIDSCostsandServicesUtilizationSurvey.KEYWORDS:AIDS;Asymptoticbiasandvariance;Clustereddata;Efficiency;Errorsinvariables;Estimatingequations;Generalizedlinearmodels;Kernelregression;Longitudinaldata;Measure-menterror;Nonparametricregression;Paneldata;SIMEX.Shorttitle.NonparametricRegressionforClusteredData1INTRODUCTIONThereisavastliteraturedevelopedinthepastdecadeonparametricregressionforclustereddatausingestimatingequations(LiangandZeger,1986),wheregeneralizedlinearmodelsareaspecialcase.Suchparametricassumptionsmaynotalwaysbedesirable,sinceappropriatefunctionalformsofthecovariatesmaynotbeknowninadvanceandtheoutcomemaydependonthecovariatesinacomplicatedmanner.Therehasbeensubstantialinterestrecentlyinextendingtheexistingparametricmodelstoallowfornonparametriccovariateeffects(ZegerandDiggle,1994;SeveriniandStaniswalis,1994;WildandYee,1996).Suchnonparametricregressionallowsformoreflexiblefunctionaldependenceoftheoutcomevariableonthecovariatesandcanalsobeusedtoinvestigatewhetheranappropriateparametricfunctioncanbedevelopedtodescribethedatawell.Anothercomplicationintheanalysisofclustereddataisthepresenceofcovariatemeasurementerror.Forexample,ithasbeenwelldocumentedintheliteraturethatcovariatessuchasbloodpressure(Carroll,RuppertandStefanski,1995)andCD4count(Tsiatis,Degruttola,Wulfsohn,1995)areoftensubjecttomeasurementerror.WeconsiderinthispaperdatafromtheAIDSCostsandServicesUtilizationSurvey(ACSUS)(Berk,etal.,1993).TheACSUSsampled2487subjectsin10randomlyselectedUScitieswithhighestAIDSrates.Aseriesofsixinterviewswereconductedforeachrespondenteverythreemonthsfrom1991to1992.Amainoutcomeofinterestwaswhetheranintervieweehadhadhospitaladmissions(yes/no)duringthepastthreemonths.Thecollectedcovariatesincludeddemographicvariables,HIVstatus,CD4count,andtreatments.AquestionofinterestinthisstudyishowCD4countaffectstheriskofhospitalization.Theanalysisofthisdatasethastwomajorcomplications.ThefirstcomplicationisthateventhoughitisbelievedthatalowerCD4countisassociatedwithahigherriskofhospitalization,

1 / 34
下载文档,编辑使用

©2015-2020 m.777doc.com 三七文档.

备案号:鲁ICP备2024069028号-1 客服联系 QQ:2149211541

×
保存成功