Recent Advances in Methods for Quantiles. Matteo Bottai, Sc.D.

Size: px
Start display at page:

Download "Recent Advances in Methods for Quantiles. Matteo Bottai, Sc.D."

Transcription

1 Recent Advances in Methods for Quantiles Matteo Bottai, Sc.D.

2 Many Thanks to Advisees Andrew Ortaglia Huiling Zhen Joe Holbrook Junlong Wu Li Zhou Marco Geraci Nicola Orsini Paolo Frumento Yuan Liu Collaborators Ane Johannessen Bo Cai Jiajia Zheng Jonathan Mitchell Lorenzo Maragoni Monica Chiogna Nicola Salvati Nikos Tzavidis Renee Gardner Robert McKeown International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 2

3 Example 1: Do We Think Percentiles? International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 3

4 Example 2: What s in a Mean? Drug Placebo Log(cholesterol) Log(cholesterol) Log(cholesterol) t test p value = Cholesterol t test p value = Mean log(cholesterol) is significantly smaller on drug than on placebo. We detect no significant difference for mean cholesterol. International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 4

5 Example 3: Logistic Quantile Regression for Bounded Outcomes The Center for Epidemiologic Studies depression scale ranges from 0 to 60. Shape and location of its distribution change across family cohesion groups. 60 CES - Depression Scale International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 5

6 Example 3: Logistic Quantile Regression for Bounded Outcomes We apply a logit transform (Bottai et al 2010) % CES - Depression Scale % 50% 25% 5% Family Cohesion International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 6

7 Example 4: Logistic Quantile Regression for Bounded Outcomes (Miniati et al 2010) International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 7

8 Example 5: Quantiles Can Be Robust to Outliers Body Mass Index (kg/m-squared) Median Regression Linear Regression All sample n = 90,244 14<BMI<60 n = 90,232 All sample n = 90,244 14<BMI<60 n = 90,232 Age P value Removing outliers affects the inference on the mean not on the median. International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 8

9 Example 6: Quantiles Can Be Efficient Bootstrap standard errors for mean and quantiles of IgE levels. Standard Error Ratio Mean quantile quantile quantile quantile quantile The 0.1 quantile with 100 observations has the same power as the mean with 41,200. International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 9

10 Asymptotic Relative Efficiency of Mean vs. Quantiles Quantile Normal(0,1) Uniform(0,1) t Student(3) Exponential(1) Chi square(1) Chi square(4) Lognormal(0,1) International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 10

11 The Asymmetric Laplace Distribution The standard asymmetric Laplace variable has probability density function ; 1 exp The parameter 0,1 drives the asymmetry and 0. The variable has The likelihood maximizer is the least weighted absolute residual estimator argmax The parameters may depend on covariates, and. International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 11

12 Linear Quantile Mixed Effects Models Consider the regression model ~ ~ is an assumed multivariate zero location distribution., International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 12

13 Example 7: Linear Quantile Mixed Effects Model with Longitudinal Data (Liu and Bottai 2009) International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 13

14 Important Related Work This is an incomplete list of related publications in the last few years. Distribution free methods Fixed effects Koenker 2004; Lamarche 2010; Galvao and Montes Rojas 2010; Galvao 2011 Weighted methods Lipsitz et al. 1997; Karlsson 2008; Fu and Wang 2012 Likelihood based approaches Asymmetric Laplace Geraci and Bottai 2007, 2013; Liu and Bottai 2009; Yuan and Yin 2010; Lee and Neocleous 2010; Farcomeni 2012 Bayesian approaches Yu and Moyeed 2001; Yuan and Yin 2010; Reich et al. 2010, 2011; Yu et al (2013) Other approaches Canay 2011; Rigby and Stasinopoulos 2005 International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 14

15 Example 8: Survival in Metastatic Renal Carcinoma There was a 28% reduction in the risk of death (hazard ratio 0.72). The Lancet 1999 International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 15

16 Example 9: Laplace Regression with Censored Data Percentiles of survival Surgery Therapy Years 5% 25% 50% 75% 95% Patients respond to surgery differently. The frailest 5% die within 6 months but the strongest 25% live at least 9.5 years. International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 16

17 Laplace Regression with Censored Data Consider the regression model We observe If ~ min, and 1 then the model is equal to ordinary quantile regression. The maximum likelihood estimator for is biased but has small mean squared error. (Bottai and Zhang 2010) International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 17

18 Important Related Work Early works Uncensored data Koenker and Geling 2001 Fixed censoring Powell 1986 Unconditionally independent censoring Semi parametric estimation Ying et al 1995 Estimating equations Bang and Tsiatis 2002 Random Censoring Global linearity assumption Portnoy 2003; Peng and Huang 2008 Nonparametric smoothing Wang and Wang 2009 Laplace regression Bottai and Zhang 2010 Applications of Laplace Regression Orsini et al 2012; Rizzuto et al 2012; Gigante 2012; Johanssen 2013 International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 18

19 Summary Quantiles can be of research interest provide a full description of continuous distributions are invariant to monotone transformations may be efficient and robust to outliers The asymmetric Laplace distribution offers a convenient likelihood framework. Current research Laplace regression with frailty terms, competing risks Flexible Laplace regression Missing data imputation Robust meta analysis Computation aspects More information at International Workshop, University of Padova, March 21 23, 2013 Matteo Bottai, Sc.D., Karolinska Institutet, Stockholm, Sweden 19

Modern Regression Methods

Modern Regression Methods Modern Regression Methods Second Edition THOMAS P. RYAN Acworth, Georgia WILEY A JOHN WILEY & SONS, INC. PUBLICATION Contents Preface 1. Introduction 1.1 Simple Linear Regression Model, 3 1.2 Uses of Regression

More information

EPI 200C Final, June 4 th, 2009 This exam includes 24 questions.

EPI 200C Final, June 4 th, 2009 This exam includes 24 questions. Greenland/Arah, Epi 200C Sp 2000 1 of 6 EPI 200C Final, June 4 th, 2009 This exam includes 24 questions. INSTRUCTIONS: Write all answers on the answer sheets supplied; PRINT YOUR NAME and STUDENT ID NUMBER

More information

Bayesian Logistic Regression Modelling via Markov Chain Monte Carlo Algorithm

Bayesian Logistic Regression Modelling via Markov Chain Monte Carlo Algorithm Journal of Social and Development Sciences Vol. 4, No. 4, pp. 93-97, Apr 203 (ISSN 222-52) Bayesian Logistic Regression Modelling via Markov Chain Monte Carlo Algorithm Henry De-Graft Acquah University

More information

Performance of Median and Least Squares Regression for Slightly Skewed Data

Performance of Median and Least Squares Regression for Slightly Skewed Data World Academy of Science, Engineering and Technology 9 Performance of Median and Least Squares Regression for Slightly Skewed Data Carolina Bancayrin - Baguio Abstract This paper presents the concept of

More information

Computer Age Statistical Inference. Algorithms, Evidence, and Data Science. BRADLEY EFRON Stanford University, California

Computer Age Statistical Inference. Algorithms, Evidence, and Data Science. BRADLEY EFRON Stanford University, California Computer Age Statistical Inference Algorithms, Evidence, and Data Science BRADLEY EFRON Stanford University, California TREVOR HASTIE Stanford University, California ggf CAMBRIDGE UNIVERSITY PRESS Preface

More information

Ecological Statistics

Ecological Statistics A Primer of Ecological Statistics Second Edition Nicholas J. Gotelli University of Vermont Aaron M. Ellison Harvard Forest Sinauer Associates, Inc. Publishers Sunderland, Massachusetts U.S.A. Brief Contents

More information

Overview. All-cause mortality for males with colon cancer and Finnish population. Relative survival

Overview. All-cause mortality for males with colon cancer and Finnish population. Relative survival An overview and some recent advances in statistical methods for population-based cancer survival analysis: relative survival, cure models, and flexible parametric models Paul W Dickman 1 Paul C Lambert

More information

The Statistical Analysis of Failure Time Data

The Statistical Analysis of Failure Time Data The Statistical Analysis of Failure Time Data Second Edition JOHN D. KALBFLEISCH ROSS L. PRENTICE iwiley- 'INTERSCIENCE A JOHN WILEY & SONS, INC., PUBLICATION Contents Preface xi 1. Introduction 1 1.1

More information

An Introduction to Multiple Imputation for Missing Items in Complex Surveys

An Introduction to Multiple Imputation for Missing Items in Complex Surveys An Introduction to Multiple Imputation for Missing Items in Complex Surveys October 17, 2014 Joe Schafer Center for Statistical Research and Methodology (CSRM) United States Census Bureau Views expressed

More information

Statistical Tolerance Regions: Theory, Applications and Computation

Statistical Tolerance Regions: Theory, Applications and Computation Statistical Tolerance Regions: Theory, Applications and Computation K. KRISHNAMOORTHY University of Louisiana at Lafayette THOMAS MATHEW University of Maryland Baltimore County Contents List of Tables

More information

Original Article Downloaded from jhs.mazums.ac.ir at 22: on Friday October 5th 2018 [ DOI: /acadpub.jhs ]

Original Article Downloaded from jhs.mazums.ac.ir at 22: on Friday October 5th 2018 [ DOI: /acadpub.jhs ] Iranian journal of health sciences 213;1(3):58-7 http://jhs.mazums.ac.ir Original Article Downloaded from jhs.mazums.ac.ir at 22:2 +33 on Friday October 5th 218 [ DOI: 1.18869/acadpub.jhs.1.3.58 ] A New

More information

Score Tests of Normality in Bivariate Probit Models

Score Tests of Normality in Bivariate Probit Models Score Tests of Normality in Bivariate Probit Models Anthony Murphy Nuffield College, Oxford OX1 1NF, UK Abstract: A relatively simple and convenient score test of normality in the bivariate probit model

More information

LIFE SATISFACTION ANALYSIS FOR RURUAL RESIDENTS IN JIANGSU PROVINCE

LIFE SATISFACTION ANALYSIS FOR RURUAL RESIDENTS IN JIANGSU PROVINCE LIFE SATISFACTION ANALYSIS FOR RURUAL RESIDENTS IN JIANGSU PROVINCE Yang Yu 1 and Zhihong Zou 2 * 1 Dr., Beihang University, China, lgyuyang@qq.com 2 Prof. Dr., Beihang University, China, zouzhihong@buaa.edu.cn

More information

Russian Journal of Agricultural and Socio-Economic Sciences, 3(15)

Russian Journal of Agricultural and Socio-Economic Sciences, 3(15) ON THE COMPARISON OF BAYESIAN INFORMATION CRITERION AND DRAPER S INFORMATION CRITERION IN SELECTION OF AN ASYMMETRIC PRICE RELATIONSHIP: BOOTSTRAP SIMULATION RESULTS Henry de-graft Acquah, Senior Lecturer

More information

List of Figures. List of Tables. Preface to the Second Edition. Preface to the First Edition

List of Figures. List of Tables. Preface to the Second Edition. Preface to the First Edition List of Figures List of Tables Preface to the Second Edition Preface to the First Edition xv xxv xxix xxxi 1 What Is R? 1 1.1 Introduction to R................................ 1 1.2 Downloading and Installing

More information

Introduction to Machine Learning. Katherine Heller Deep Learning Summer School 2018

Introduction to Machine Learning. Katherine Heller Deep Learning Summer School 2018 Introduction to Machine Learning Katherine Heller Deep Learning Summer School 2018 Outline Kinds of machine learning Linear regression Regularization Bayesian methods Logistic Regression Why we do this

More information

ISIR: Independent Sliced Inverse Regression

ISIR: Independent Sliced Inverse Regression ISIR: Independent Sliced Inverse Regression Kevin B. Li Beijing Jiaotong University Abstract In this paper we consider a semiparametric regression model involving a p-dimensional explanatory variable x

More information

AP Statistics. Semester One Review Part 1 Chapters 1-5

AP Statistics. Semester One Review Part 1 Chapters 1-5 AP Statistics Semester One Review Part 1 Chapters 1-5 AP Statistics Topics Describing Data Producing Data Probability Statistical Inference Describing Data Ch 1: Describing Data: Graphically and Numerically

More information

Chapter 11: Advanced Remedial Measures. Weighted Least Squares (WLS)

Chapter 11: Advanced Remedial Measures. Weighted Least Squares (WLS) Chapter : Advanced Remedial Measures Weighted Least Squares (WLS) When the error variance appears nonconstant, a transformation (of Y and/or X) is a quick remedy. But it may not solve the problem, or it

More information

LINEAR REGRESSION FOR BIVARIATE CENSORED DATA VIA MULTIPLE IMPUTATION WEI PAN. School of Public Health. 420 Delaware Street SE

LINEAR REGRESSION FOR BIVARIATE CENSORED DATA VIA MULTIPLE IMPUTATION WEI PAN. School of Public Health. 420 Delaware Street SE LINEAR REGRESSION FOR BIVARIATE CENSORED DATA VIA MULTIPLE IMPUTATION Research Report 99-002 WEI PAN Division of Biostatistics School of Public Health University of Minnesota A460 Mayo Building, Box 303

More information

Multivariate dose-response meta-analysis: an update on glst

Multivariate dose-response meta-analysis: an update on glst Multivariate dose-response meta-analysis: an update on glst Nicola Orsini Unit of Biostatistics Unit of Nutritional Epidemiology Institute of Environmental Medicine Karolinska Institutet http://www.imm.ki.se/biostatistics/

More information

16:35 17:20 Alexander Luedtke (Fred Hutchinson Cancer Research Center)

16:35 17:20 Alexander Luedtke (Fred Hutchinson Cancer Research Center) Conference on Causal Inference in Longitudinal Studies September 21-23, 2017 Columbia University Thursday, September 21, 2017: tutorial 14:15 15:00 Miguel Hernan (Harvard University) 15:00 15:45 Miguel

More information

BEST PRACTICES FOR IMPLEMENTATION AND ANALYSIS OF PAIN SCALE PATIENT REPORTED OUTCOMES IN CLINICAL TRIALS

BEST PRACTICES FOR IMPLEMENTATION AND ANALYSIS OF PAIN SCALE PATIENT REPORTED OUTCOMES IN CLINICAL TRIALS BEST PRACTICES FOR IMPLEMENTATION AND ANALYSIS OF PAIN SCALE PATIENT REPORTED OUTCOMES IN CLINICAL TRIALS Nan Shao, Ph.D. Director, Biostatistics Premier Research Group, Limited and Mark Jaros, Ph.D. Senior

More information

Xiaoyan(Iris) Lin. University of South Carolina Office: B LeConte College Fax: Columbia, SC, 29208

Xiaoyan(Iris) Lin. University of South Carolina Office: B LeConte College Fax: Columbia, SC, 29208 Xiaoyan(Iris) Lin Department of Statistics lin9@mailbox.sc.edu University of South Carolina Office: 803-777-3788 209B LeConte College Fax: 803-777-4048 Columbia, SC, 29208 Education Doctor of Philosophy

More information

Contents. Part 1 Introduction. Part 2 Cross-Sectional Selection Bias Adjustment

Contents. Part 1 Introduction. Part 2 Cross-Sectional Selection Bias Adjustment From Analysis of Observational Health Care Data Using SAS. Full book available for purchase here. Contents Preface ix Part 1 Introduction Chapter 1 Introduction to Observational Studies... 3 1.1 Observational

More information

Measurement Error in Nonlinear Models

Measurement Error in Nonlinear Models Measurement Error in Nonlinear Models R.J. CARROLL Professor of Statistics Texas A&M University, USA D. RUPPERT Professor of Operations Research and Industrial Engineering Cornell University, USA and L.A.

More information

A COMPARISON OF IMPUTATION METHODS FOR MISSING DATA IN A MULTI-CENTER RANDOMIZED CLINICAL TRIAL: THE IMPACT STUDY

A COMPARISON OF IMPUTATION METHODS FOR MISSING DATA IN A MULTI-CENTER RANDOMIZED CLINICAL TRIAL: THE IMPACT STUDY A COMPARISON OF IMPUTATION METHODS FOR MISSING DATA IN A MULTI-CENTER RANDOMIZED CLINICAL TRIAL: THE IMPACT STUDY Lingqi Tang 1, Thomas R. Belin 2, and Juwon Song 2 1 Center for Health Services Research,

More information

LINEAR REGRESSION FOR BIVARIATE CENSORED DATA VIA MULTIPLE IMPUTATION

LINEAR REGRESSION FOR BIVARIATE CENSORED DATA VIA MULTIPLE IMPUTATION STATISTICS IN MEDICINE Statist. Med. 18, 3111} 3121 (1999) LINEAR REGRESSION FOR BIVARIATE CENSORED DATA VIA MULTIPLE IMPUTATION WEI PAN * AND CHARLES KOOPERBERG Division of Biostatistics, School of Public

More information

Biostatistics II

Biostatistics II Biostatistics II 514-5509 Course Description: Modern multivariable statistical analysis based on the concept of generalized linear models. Includes linear, logistic, and Poisson regression, survival analysis,

More information

Discrepancy between the nonautomatic

Discrepancy between the nonautomatic Discrepancy between the nonautomatic and semiautomatic measurements of AAA diameter-does it influence outcome? Joy Roy Department of Vascular Surgery Karolinska University Hospital / Karolinska Institutet

More information

Vessel wall differences between middle cerebral artery and basilar artery. plaques on magnetic resonance imaging

Vessel wall differences between middle cerebral artery and basilar artery. plaques on magnetic resonance imaging Vessel wall differences between middle cerebral artery and basilar artery plaques on magnetic resonance imaging Peng-Peng Niu, MD 1 ; Yao Yu, MD 1 ; Hong-Wei Zhou, MD 2 ; Yang Liu, MD 2 ; Yun Luo, MD 1

More information

Quantifying cancer patient survival; extensions and applications of cure models and life expectancy estimation

Quantifying cancer patient survival; extensions and applications of cure models and life expectancy estimation From the Department of Medical Epidemiology and Biostatistics Karolinska Institutet, Stockholm, Sweden Quantifying cancer patient survival; extensions and applications of cure models and life expectancy

More information

Lecture Outline. Biost 590: Statistical Consulting. Stages of Scientific Studies. Scientific Method

Lecture Outline. Biost 590: Statistical Consulting. Stages of Scientific Studies. Scientific Method Biost 590: Statistical Consulting Statistical Classification of Scientific Studies; Approach to Consulting Lecture Outline Statistical Classification of Scientific Studies Statistical Tasks Approach to

More information

A Study on Type 2 Diabetes Mellitus Patients Using Regression Model and Survival Analysis Techniques

A Study on Type 2 Diabetes Mellitus Patients Using Regression Model and Survival Analysis Techniques Available online at www.ijpab.com Shaik et al Int. J. Pure App. Biosci. 6 (1): 514-522 (2018) ISSN: 2320 7051 DOI: http://dx.doi.org/10.18782/2320-7051.5999 ISSN: 2320 7051 Int. J. Pure App. Biosci. 6

More information

For general queries, contact

For general queries, contact Much of the work in Bayesian econometrics has focused on showing the value of Bayesian methods for parametric models (see, for example, Geweke (2005), Koop (2003), Li and Tobias (2011), and Rossi, Allenby,

More information

cloglog link function to transform the (population) hazard probability into a continuous

cloglog link function to transform the (population) hazard probability into a continuous Supplementary material. Discrete time event history analysis Hazard model details. In our discrete time event history analysis, we used the asymmetric cloglog link function to transform the (population)

More information

Xiaoyan (Iris) Lin. University of South Carolina Office: B LeConte College Fax: Columbia, SC, 29208

Xiaoyan (Iris) Lin. University of South Carolina Office: B LeConte College Fax: Columbia, SC, 29208 Xiaoyan (Iris) Lin Department of Statistics lin9@mailbox.sc.edu University of South Carolina Office: 803-777-3788 209B LeConte College Fax: 803-777-4048 Columbia, SC, 29208 Education Ph.D. in Statistics,

More information

Selection and Combination of Markers for Prediction

Selection and Combination of Markers for Prediction Selection and Combination of Markers for Prediction NACC Data and Methods Meeting September, 2010 Baojiang Chen, PhD Sarah Monsell, MS Xiao-Hua Andrew Zhou, PhD Overview 1. Research motivation 2. Describe

More information

investigate. educate. inform.

investigate. educate. inform. investigate. educate. inform. Research Design What drives your research design? The battle between Qualitative and Quantitative is over Think before you leap What SHOULD drive your research design. Advanced

More information

Bayesians methods in system identification: equivalences, differences, and misunderstandings

Bayesians methods in system identification: equivalences, differences, and misunderstandings Bayesians methods in system identification: equivalences, differences, and misunderstandings Johan Schoukens and Carl Edward Rasmussen ERNSI 217 Workshop on System Identification Lyon, September 24-27,

More information

Statistics as a Tool. A set of tools for collecting, organizing, presenting and analyzing numerical facts or observations.

Statistics as a Tool. A set of tools for collecting, organizing, presenting and analyzing numerical facts or observations. Statistics as a Tool A set of tools for collecting, organizing, presenting and analyzing numerical facts or observations. Descriptive Statistics Numerical facts or observations that are organized describe

More information

LOGO. Statistical Modeling of Breast and Lung Cancers. Cancer Research Team. Department of Mathematics and Statistics University of South Florida

LOGO. Statistical Modeling of Breast and Lung Cancers. Cancer Research Team. Department of Mathematics and Statistics University of South Florida LOGO Statistical Modeling of Breast and Lung Cancers Cancer Research Team Department of Mathematics and Statistics University of South Florida 1 LOGO 2 Outline Nonparametric and parametric analysis of

More information

Missing Data and Imputation

Missing Data and Imputation Missing Data and Imputation Barnali Das NAACCR Webinar May 2016 Outline Basic concepts Missing data mechanisms Methods used to handle missing data 1 What are missing data? General term: data we intended

More information

Various Approaches to Szroeter s Test for Regression Quantiles

Various Approaches to Szroeter s Test for Regression Quantiles The International Scientific Conference INPROFORUM 2017, November 9, 2017, České Budějovice, 361-365, ISBN 978-80-7394-667-8. Various Approaches to Szroeter s Test for Regression Quantiles Jan Kalina,

More information

Citation for published version (APA): Ebbes, P. (2004). Latent instrumental variables: a new approach to solve for endogeneity s.n.

Citation for published version (APA): Ebbes, P. (2004). Latent instrumental variables: a new approach to solve for endogeneity s.n. University of Groningen Latent instrumental variables Ebbes, P. IMPORTANT NOTE: You are advised to consult the publisher's version (publisher's PDF) if you wish to cite from it. Please check the document

More information

What you should know before you collect data. BAE 815 (Fall 2017) Dr. Zifei Liu

What you should know before you collect data. BAE 815 (Fall 2017) Dr. Zifei Liu What you should know before you collect data BAE 815 (Fall 2017) Dr. Zifei Liu Zifeiliu@ksu.edu Types and levels of study Descriptive statistics Inferential statistics How to choose a statistical test

More information

Detection of Unknown Confounders. by Bayesian Confirmatory Factor Analysis

Detection of Unknown Confounders. by Bayesian Confirmatory Factor Analysis Advanced Studies in Medical Sciences, Vol. 1, 2013, no. 3, 143-156 HIKARI Ltd, www.m-hikari.com Detection of Unknown Confounders by Bayesian Confirmatory Factor Analysis Emil Kupek Department of Public

More information

Should a Normal Imputation Model Be Modified to Impute Skewed Variables?

Should a Normal Imputation Model Be Modified to Impute Skewed Variables? Sociological Methods and Research, 2013, 42(1), 105-138 Should a Normal Imputation Model Be Modified to Impute Skewed Variables? Paul T. von Hippel Abstract (169 words) Researchers often impute continuous

More information

Sarwar Islam Mozumder 1, Mark J Rutherford 1 & Paul C Lambert 1, Stata London Users Group Meeting

Sarwar Islam Mozumder 1, Mark J Rutherford 1 & Paul C Lambert 1, Stata London Users Group Meeting 2016 Stata London Users Group Meeting stpm2cr: A Stata module for direct likelihood inference on the cause-specific cumulative incidence function within the flexible parametric modelling framework Sarwar

More information

Applied Medical. Statistics Using SAS. Geoff Der. Brian S. Everitt. CRC Press. Taylor Si Francis Croup. Taylor & Francis Croup, an informa business

Applied Medical. Statistics Using SAS. Geoff Der. Brian S. Everitt. CRC Press. Taylor Si Francis Croup. Taylor & Francis Croup, an informa business Applied Medical Statistics Using SAS Geoff Der Brian S. Everitt CRC Press Taylor Si Francis Croup Boca Raton London New York CRC Press is an imprint of the Taylor & Francis Croup, an informa business A

More information

Statistical Analysis with Missing Data. Second Edition

Statistical Analysis with Missing Data. Second Edition Statistical Analysis with Missing Data Second Edition WILEY SERIES IN PROBABILITY AND STATISTICS Established by WALTER A. SHEWHART and SAMUEL S. WILKS Editors: David J. Balding, Peter Bloomfield, Noel

More information

Bayesian Confidence Intervals for Means and Variances of Lognormal and Bivariate Lognormal Distributions

Bayesian Confidence Intervals for Means and Variances of Lognormal and Bivariate Lognormal Distributions Bayesian Confidence Intervals for Means and Variances of Lognormal and Bivariate Lognormal Distributions J. Harvey a,b, & A.J. van der Merwe b a Centre for Statistical Consultation Department of Statistics

More information

breast cancer; relative risk; risk factor; standard deviation; strength of association

breast cancer; relative risk; risk factor; standard deviation; strength of association American Journal of Epidemiology The Author 2015. Published by Oxford University Press on behalf of the Johns Hopkins Bloomberg School of Public Health. All rights reserved. For permissions, please e-mail:

More information

CONDITIONAL REGRESSION MODELS TRANSIENT STATE SURVIVAL ANALYSIS

CONDITIONAL REGRESSION MODELS TRANSIENT STATE SURVIVAL ANALYSIS CONDITIONAL REGRESSION MODELS FOR TRANSIENT STATE SURVIVAL ANALYSIS Robert D. Abbott Field Studies Branch National Heart, Lung and Blood Institute National Institutes of Health Raymond J. Carroll Department

More information

Missing Data in Longitudinal Studies: Strategies for Bayesian Modeling, Sensitivity Analysis, and Causal Inference

Missing Data in Longitudinal Studies: Strategies for Bayesian Modeling, Sensitivity Analysis, and Causal Inference COURSE: Missing Data in Longitudinal Studies: Strategies for Bayesian Modeling, Sensitivity Analysis, and Causal Inference Mike Daniels (Department of Statistics, University of Florida) 20-21 October 2011

More information

Lecture 21. RNA-seq: Advanced analysis

Lecture 21. RNA-seq: Advanced analysis Lecture 21 RNA-seq: Advanced analysis Experimental design Introduction An experiment is a process or study that results in the collection of data. Statistical experiments are conducted in situations in

More information

Linear Regression in SAS

Linear Regression in SAS 1 Suppose we wish to examine factors that predict patient s hemoglobin levels. Simulated data for six patients is used throughout this tutorial. data hgb_data; input id age race $ bmi hgb; cards; 21 25

More information

Selected Topics in Biostatistics Seminar Series. Missing Data. Sponsored by: Center For Clinical Investigation and Cleveland CTSC

Selected Topics in Biostatistics Seminar Series. Missing Data. Sponsored by: Center For Clinical Investigation and Cleveland CTSC Selected Topics in Biostatistics Seminar Series Missing Data Sponsored by: Center For Clinical Investigation and Cleveland CTSC Brian Schmotzer, MS Biostatistician, CCI Statistical Sciences Core brian.schmotzer@case.edu

More information

MEA DISCUSSION PAPERS

MEA DISCUSSION PAPERS Inference Problems under a Special Form of Heteroskedasticity Helmut Farbmacher, Heinrich Kögel 03-2015 MEA DISCUSSION PAPERS mea Amalienstr. 33_D-80799 Munich_Phone+49 89 38602-355_Fax +49 89 38602-390_www.mea.mpisoc.mpg.de

More information

Beyond the intention-to treat effect: Per-protocol effects in randomized trials

Beyond the intention-to treat effect: Per-protocol effects in randomized trials Beyond the intention-to treat effect: Per-protocol effects in randomized trials Miguel Hernán DEPARTMENTS OF EPIDEMIOLOGY AND BIOSTATISTICS Intention-to-treat analysis (estimator) estimates intention-to-treat

More information

Bayesian Inference Bayes Laplace

Bayesian Inference Bayes Laplace Bayesian Inference Bayes Laplace Course objective The aim of this course is to introduce the modern approach to Bayesian statistics, emphasizing the computational aspects and the differences between the

More information

bivariate analysis: The statistical analysis of the relationship between two variables.

bivariate analysis: The statistical analysis of the relationship between two variables. bivariate analysis: The statistical analysis of the relationship between two variables. cell frequency: The number of cases in a cell of a cross-tabulation (contingency table). chi-square (χ 2 ) test for

More information

A Bayesian Nonparametric Model Fit statistic of Item Response Models

A Bayesian Nonparametric Model Fit statistic of Item Response Models A Bayesian Nonparametric Model Fit statistic of Item Response Models Purpose As more and more states move to use the computer adaptive test for their assessments, item response theory (IRT) has been widely

More information

MS&E 226: Small Data

MS&E 226: Small Data MS&E 226: Small Data Lecture 10: Introduction to inference (v2) Ramesh Johari ramesh.johari@stanford.edu 1 / 17 What is inference? 2 / 17 Where did our data come from? Recall our sample is: Y, the vector

More information

From Biostatistics Using JMP: A Practical Guide. Full book available for purchase here. Chapter 1: Introduction... 1

From Biostatistics Using JMP: A Practical Guide. Full book available for purchase here. Chapter 1: Introduction... 1 From Biostatistics Using JMP: A Practical Guide. Full book available for purchase here. Contents Dedication... iii Acknowledgments... xi About This Book... xiii About the Author... xvii Chapter 1: Introduction...

More information

PATCH Analysis Plan v1.2.doc Prophylactic Antibiotics for the Treatment of Cellulitis at Home: PATCH Analysis Plan for PATCH I and PATCH II Authors: Angela Crook, Andrew Nunn, James Mason and Kim Thomas,

More information

Bayesian Analysis of Between-Group Differences in Variance Components in Hierarchical Generalized Linear Models

Bayesian Analysis of Between-Group Differences in Variance Components in Hierarchical Generalized Linear Models Bayesian Analysis of Between-Group Differences in Variance Components in Hierarchical Generalized Linear Models Brady T. West Michigan Program in Survey Methodology, Institute for Social Research, 46 Thompson

More information

PLEASE SCROLL DOWN FOR ARTICLE. Full terms and conditions of use:

PLEASE SCROLL DOWN FOR ARTICLE. Full terms and conditions of use: This article was downloaded by: [CDL Journals Account] On: 26 November 2009 Access details: Access Details: [subscription number 794379768] Publisher Psychology Press Informa Ltd Registered in England

More information

Exploring the Impact of Missing Data in Multiple Regression

Exploring the Impact of Missing Data in Multiple Regression Exploring the Impact of Missing Data in Multiple Regression Michael G Kenward London School of Hygiene and Tropical Medicine 28th May 2015 1. Introduction In this note we are concerned with the conduct

More information

A Comparison of Robust and Nonparametric Estimators Under the Simple Linear Regression Model

A Comparison of Robust and Nonparametric Estimators Under the Simple Linear Regression Model Nevitt & Tam A Comparison of Robust and Nonparametric Estimators Under the Simple Linear Regression Model Jonathan Nevitt, University of Maryland, College Park Hak P. Tam, National Taiwan Normal University

More information

Estimating average treatment effects from observational data using teffects

Estimating average treatment effects from observational data using teffects Estimating average treatment effects from observational data using teffects David M. Drukker Director of Econometrics Stata 2013 Nordic and Baltic Stata Users Group meeting Karolinska Institutet September

More information

Estimands, Missing Data and Sensitivity Analysis: some overview remarks. Roderick Little

Estimands, Missing Data and Sensitivity Analysis: some overview remarks. Roderick Little Estimands, Missing Data and Sensitivity Analysis: some overview remarks Roderick Little NRC Panel s Charge To prepare a report with recommendations that would be useful for USFDA's development of guidance

More information

Characteristics of Patients Initializing Peritoneal Dialysis Treatment From 2007 to 2014 Analysis From Henan Peritoneal Dialysis Registry data

Characteristics of Patients Initializing Peritoneal Dialysis Treatment From 2007 to 2014 Analysis From Henan Peritoneal Dialysis Registry data DIALYSIS Characteristics of Patients Initializing Peritoneal Dialysis Treatment From 7 to 14 Analysis From Henan Peritoneal Dialysis Registry data Xiaoxue Zhang, 1 Ying Chen, 1,2 Yamei Cai, 1 Xing Tian,

More information

Taban Baghfalaki C.V. - 1 CURRICULUM VITAE. Taban Baghfalaki

Taban Baghfalaki C.V. - 1 CURRICULUM VITAE. Taban Baghfalaki Taban Baghfalaki C.V. - 1 CURRICULUM VITAE Taban Baghfalaki 2018 ADDRESS Department of Statistics, Faculty of Mathematical Sciences, Tarbiat Modares University, Tehran, Iran Phone: +98 (0) 21 82884768

More information

Simple Linear Regression

Simple Linear Regression Simple Linear Regression Assoc. Prof Dr Sarimah Abdullah Unit of Biostatistics & Research Methodology School of Medical Sciences, Health Campus Universiti Sains Malaysia Regression Regression analysis

More information

Statistical Models for Censored Point Processes with Cure Rates

Statistical Models for Censored Point Processes with Cure Rates Statistical Models for Censored Point Processes with Cure Rates Jennifer Rogers MSD Seminar 2 November 2011 Outline Background and MESS Epilepsy MESS Exploratory Analysis Summary Statistics and Kaplan-Meier

More information

Statistical Methods for Wearable Technology in CNS Trials

Statistical Methods for Wearable Technology in CNS Trials Statistical Methods for Wearable Technology in CNS Trials Andrew Potter, PhD Division of Biometrics 1, OB/OTS/CDER, FDA ISCTM 2018 Autumn Conference Oct. 15, 2018 Marina del Rey, CA www.fda.gov Disclaimer

More information

Institutional Ranking. VHA Study

Institutional Ranking. VHA Study Statistical Inference for Ranks of Health Care Facilities in the Presence of Ties and Near Ties Minge Xie Department of Statistics Rutgers, The State University of New Jersey Supported in part by NSF,

More information

Comparison And Application Of Methods To Address Confounding By Indication In Non- Randomized Clinical Studies

Comparison And Application Of Methods To Address Confounding By Indication In Non- Randomized Clinical Studies University of Massachusetts Amherst ScholarWorks@UMass Amherst Masters Theses 1911 - February 2014 Dissertations and Theses 2013 Comparison And Application Of Methods To Address Confounding By Indication

More information

NEW METHODS FOR SENSITIVITY TESTS OF EXPLOSIVE DEVICES

NEW METHODS FOR SENSITIVITY TESTS OF EXPLOSIVE DEVICES NEW METHODS FOR SENSITIVITY TESTS OF EXPLOSIVE DEVICES Amit Teller 1, David M. Steinberg 2, Lina Teper 1, Rotem Rozenblum 2, Liran Mendel 2, and Mordechai Jaeger 2 1 RAFAEL, POB 2250, Haifa, 3102102, Israel

More information

Data Analysis Using Regression and Multilevel/Hierarchical Models

Data Analysis Using Regression and Multilevel/Hierarchical Models Data Analysis Using Regression and Multilevel/Hierarchical Models ANDREW GELMAN Columbia University JENNIFER HILL Columbia University CAMBRIDGE UNIVERSITY PRESS Contents List of examples V a 9 e xv " Preface

More information

Can Family Strengths Reduce Risk of Substance Abuse among Youth with SED?

Can Family Strengths Reduce Risk of Substance Abuse among Youth with SED? Can Family Strengths Reduce Risk of Substance Abuse among Youth with SED? Michael D. Pullmann Portland State University Ana María Brannan Vanderbilt University Robert Stephens ORC Macro, Inc. Relationship

More information

Anale. Seria Informatică. Vol. XVI fasc Annals. Computer Science Series. 16 th Tome 1 st Fasc. 2018

Anale. Seria Informatică. Vol. XVI fasc Annals. Computer Science Series. 16 th Tome 1 st Fasc. 2018 HANDLING MULTICOLLINEARITY; A COMPARATIVE STUDY OF THE PREDICTION PERFORMANCE OF SOME METHODS BASED ON SOME PROBABILITY DISTRIBUTIONS Zakari Y., Yau S. A., Usman U. Department of Mathematics, Usmanu Danfodiyo

More information

Bayesian Joint Modelling of Longitudinal and Survival Data of HIV/AIDS Patients: A Case Study at Bale Robe General Hospital, Ethiopia

Bayesian Joint Modelling of Longitudinal and Survival Data of HIV/AIDS Patients: A Case Study at Bale Robe General Hospital, Ethiopia American Journal of Theoretical and Applied Statistics 2017; 6(4): 182-190 http://www.sciencepublishinggroup.com/j/ajtas doi: 10.11648/j.ajtas.20170604.13 ISSN: 2326-8999 (Print); ISSN: 2326-9006 (Online)

More information

Correlation and regression

Correlation and regression PG Dip in High Intensity Psychological Interventions Correlation and regression Martin Bland Professor of Health Statistics University of York http://martinbland.co.uk/ Correlation Example: Muscle strength

More information

CLASSICAL AND. MODERN REGRESSION WITH APPLICATIONS

CLASSICAL AND. MODERN REGRESSION WITH APPLICATIONS - CLASSICAL AND. MODERN REGRESSION WITH APPLICATIONS SECOND EDITION Raymond H. Myers Virginia Polytechnic Institute and State university 1 ~l~~l~l~~~~~~~l!~ ~~~~~l~/ll~~ Donated by Duxbury o Thomson Learning,,

More information

Discontinuation and restarting in patients on statin treatment: prospective open cohort study using a primary care database

Discontinuation and restarting in patients on statin treatment: prospective open cohort study using a primary care database open access Discontinuation and restarting in patients on statin treatment: prospective open cohort study using a primary care database Yana Vinogradova, 1 Carol Coupland, 1 Peter Brindle, 2,3 Julia Hippisley-Cox

More information

A novel approach to estimation of the time to biomarker threshold: Applications to HIV

A novel approach to estimation of the time to biomarker threshold: Applications to HIV A novel approach to estimation of the time to biomarker threshold: Applications to HIV Pharmaceutical Statistics, Volume 15, Issue 6, Pages 541-549, November/December 2016 PSI Journal Club 22 March 2017

More information

What to do with missing data in clinical registry analysis?

What to do with missing data in clinical registry analysis? Melbourne 2011; Registry Special Interest Group What to do with missing data in clinical registry analysis? Rory Wolfe Acknowledgements: James Carpenter, Gerard O Reilly Department of Epidemiology & Preventive

More information

Robust Outbreak Surveillance

Robust Outbreak Surveillance Robust Outbreak Surveillance Marianne Frisén Statistical Research Unit University of Gothenburg Sweden Marianne Frisén Open U 21 1 Outline Aims Method and theoretical background Computer program Application

More information

Fundamental Clinical Trial Design

Fundamental Clinical Trial Design Design, Monitoring, and Analysis of Clinical Trials Session 1 Overview and Introduction Overview Scott S. Emerson, M.D., Ph.D. Professor of Biostatistics, University of Washington February 17-19, 2003

More information

Content. Basic Statistics and Data Analysis for Health Researchers from Foreign Countries. Research question. Example Newly diagnosed Type 2 Diabetes

Content. Basic Statistics and Data Analysis for Health Researchers from Foreign Countries. Research question. Example Newly diagnosed Type 2 Diabetes Content Quantifying association between continuous variables. Basic Statistics and Data Analysis for Health Researchers from Foreign Countries Volkert Siersma siersma@sund.ku.dk The Research Unit for General

More information

Part [1.0] Introduction to Development and Evaluation of Dynamic Predictions

Part [1.0] Introduction to Development and Evaluation of Dynamic Predictions Part [1.0] Introduction to Development and Evaluation of Dynamic Predictions A Bansal & PJ Heagerty Department of Biostatistics University of Washington 1 Biomarkers The Instructor(s) Patrick Heagerty

More information

Differential Item Functioning

Differential Item Functioning Differential Item Functioning Lecture #11 ICPSR Item Response Theory Workshop Lecture #11: 1of 62 Lecture Overview Detection of Differential Item Functioning (DIF) Distinguish Bias from DIF Test vs. Item

More information

Econometric Game 2012: infants birthweight?

Econometric Game 2012: infants birthweight? Econometric Game 2012: How does maternal smoking during pregnancy affect infants birthweight? Case A April 18, 2012 1 Introduction Low birthweight is associated with adverse health related and economic

More information

Introduction to Survival Analysis Procedures (Chapter)

Introduction to Survival Analysis Procedures (Chapter) SAS/STAT 9.3 User s Guide Introduction to Survival Analysis Procedures (Chapter) SAS Documentation This document is an individual chapter from SAS/STAT 9.3 User s Guide. The correct bibliographic citation

More information

Accommodating informative dropout and death: a joint modelling approach for longitudinal and semicompeting risks data

Accommodating informative dropout and death: a joint modelling approach for longitudinal and semicompeting risks data Appl. Statist. (2018) 67, Part 1, pp. 145 163 Accommodating informative dropout and death: a joint modelling approach for longitudinal and semicompeting risks data Qiuju Li and Li Su Medical Research Council

More information

Daniel Boduszek University of Huddersfield

Daniel Boduszek University of Huddersfield Daniel Boduszek University of Huddersfield d.boduszek@hud.ac.uk Introduction to Logistic Regression SPSS procedure of LR Interpretation of SPSS output Presenting results from LR Logistic regression is

More information

Using SAS to Calculate Tests of Cliff s Delta. Kristine Y. Hogarty and Jeffrey D. Kromrey

Using SAS to Calculate Tests of Cliff s Delta. Kristine Y. Hogarty and Jeffrey D. Kromrey Using SAS to Calculate Tests of Cliff s Delta Kristine Y. Hogarty and Jeffrey D. Kromrey Department of Educational Measurement and Research, University of South Florida ABSTRACT This paper discusses a

More information

A Comparison of Analytic Models for the Costs of the Hospitalized Diabetic Patients

A Comparison of Analytic Models for the Costs of the Hospitalized Diabetic Patients Metodološki zvezki, Vol. 2, No. 1, 2005, 135-145 A Comparison of Analytic Models for the Costs of the Hospitalized Diabetic Patients Giulia Zigon 1, Rosalba Rosato 2, Simona Bò 3 and Dario Gregori 4 Abstract

More information