GETTING STARTED: STATA & R BASIC COMMANDS ECONOMETRICS II. Stata Output Regression of wages on education


 Thomasina Hall
 2 years ago
 Views:
Transcription
1 GETTING STARTED: STATA & R BASIC COMMANDS ECONOMETRICS II Stata Output Regression of wages on education. sum wage educ Variable Obs Mean Std. Dev. Min Max wage educ reg wage educ Source SS df MS Number of obs F( 1, 524) Model Prob > F 0 Residual Rsquared Adj Rsquared Total Root MSE wage Coef. Std. Err. t P> t [95% Conf. Interval] educ _cons Do File Regression of wages on education * clear log using lecture1.log, replace use Wage1.dta sum wage educ reg wage educ log close * R Output Regression of wages on education > data < read.csv( YOUR PATH, headert) > data < data.frame(data) > summary(data)[,1:2] wage educ "Min. : " "Min. : 0.00 " "1st Qu.: " "1st Qu.:12.00 " "Median : " "Median :12.00 " "Mean : " "Mean :12.56 " "3rd Qu.: " "3rd Qu.:14.00 " "Max. : " "Max. :18.00 " > fit < lm(data$wage ~ data$educ)
2 > summary(fit) Call: lm(formula data$wage ~ data$educ) Residuals: Min 1Q Median 3Q Max Coefficients: Estimate Std. Error t value Pr(> t ) (Intercept) data$educ <2e16 ***  Signif. codes: 0 `***' `**' 0.01 `*' 0.05 `.' 0.1 ` ' 1 Residual standard error: on 524 degrees of freedom Multiple RSquared: , Adjusted Rsquared: Fstatistic: on 1 and 524 DF, pvalue: < 2.2e16 R text File Regression of wages on education # data < read.csv("your PATH", headert) data < data.frame(data) # # Figure 1, Lecture 1: Scartterplot # postscript("wage_fig1.ps", horizontal FALSE, height6.5,width6.5) par(mfrowc(1,1)) plot(data$educ,data$wage,,type"n",xlab "years of education", ylab"average hourly earnings") points(data$educ,data$wage,pch19,cex0.5) dev.off() # summary(data)[,1:2] fit < lm(data$wage ~ data$educ) summary(fit) #
3 Stata Output Regression of wages on education, experience and tenure. use Wage1.dta. reg lwage educ exper expersq tenure Source SS df MS Number of obs F( 4, 521) Model Prob > F 0 Residual Rsquared Adj Rsquared Total Root MSE educ exper expersq tenure _cons display "Number of Observations " _result(1) Number of Observations 526. display "R2 " _result(7) R vce educ exper expersq tenure _cons educ exper 1.7e expersq 1.2e e e08 tenure 2.2e e e e06 _cons e predict yhat (option xb assumed; fitted values). predict uhat, resid. sum uhat Variable Obs Mean Std. Dev. Min Max uhat e R Output Regression of wages on education, experience and tenure > data < read.csv( YOUR PATH", headert) > data < data.frame(data) > fit < lm(data$lwage ~ data$educ + data$exper + data$expersq + data$tenure) > summary(fit)
4 Call: lm(formula data$lwage ~ data$educ + data$exper + data$expersq + data$tenure) Residuals: Min 1Q Median 3Q Max Coefficients: Estimate Std. Error t value Pr(> t ) (Intercept) data$educ < 2e16 *** data$exper e10 *** data$expersq e09 *** data$tenure e11 ***  Signif. codes: 0 `***' `**' 0.01 `*' 0.05 `.' 0.1 ` ' 1 Residual standard error: on 521 degrees of freedom Multiple RSquared: , Adjusted Rsquared: Fstatistic: on 4 and 521 DF, pvalue: < 2.2e16 > length(data$educ) [1] 526 > vcov(fit) (Intercept) data$educ data$exper data$expersq (Intercept) data$educ e e e e e e e e07 data$exper e e e e07 data$expersq e e e e08 data$tenure e e e e08 data$tenure (Intercept) data$educ e e06 data$exper e06 data$expersq e08 data$tenure e06 > fit$fitted > yhat #or > fitted(fit) > yhat > fit$resid > uhat
5 . use Wage1.dta. reg lwage educ exper tenure Source SS df MS Number of obs F( 3, 522) Model Prob > F 0 Residual Rsquared Adj Rsquared Total Root MSE educ exper tenure _cons test educ ( 1) educ 0 F( 1, 522) Prob > F 0. test educ 0.1 ( 1) educ.1 F( 1, 522) 1.18 Prob > F test exper tenure ( 1) exper 0 ( 2) tenure 0 F( 2, 522) Prob > F 0. test exper tenure ( 1) exper  tenure 0 F( 1, 522) Prob > F 0 R Output > fit < lm(data$lwage ~ data$educ + data$exper + data$tenure) > summary(fit) Call: lm(formula data$lwage ~ data$educ + data$exper + data$tenure) Residuals: Min 1Q Median 3Q Max Coefficients: Estimate Std. Error t value Pr(> t ) (Intercept) data$educ ** < 2e16 *** data$exper * data$tenure e12 ***
6  Signif. codes: 0 `***' `**' 0.01 `*' 0.05 `.' 0.1 ` ' 1 Residual standard error: on 522 degrees of freedom Multiple RSquared: 0.316, Adjusted Rsquared: Fstatistic: on 3 and 522 DF, pvalue: < 2.2e16 > # > fit.u < lm(data$lwage ~ data$educ + data$exper + data$tenure) > fit.r < lm(data$lwage ~ data$exper + data$tenure) > F.test(fit.u,fit.r) $F [1] $Prob [1] 0 > # > fit.u < lm(data$lwage ~ data$educ + data$exper + data$tenure) > fit.r < lm((data$lwage  0.1*data$educ) ~ data$exper + data$tenure) > F.test(fit.u,fit.r) $F [1] $Prob [1] > # > fit.u < lm(data$lwage ~ data$educ + data$exper + data$tenure) > fit.r < lm(data$lwage ~ data$educ) > F.test(fit.u,fit.r) $F [1] $Prob [1] 0 > # > fit.u < lm(data$lwage ~ data$educ + data$exper + data$tenure) > x < (data$exper + data$tenure) > fit.r < lm(data$lwage ~ data$educ + x) > F.test(fit.u,fit.r) $F [1] $Prob [1] e05
7 Panel Data Methods  Stata Output. clear. use Jtrain.dta. tis year. iis fcode. sort fcode. quietly by fcode: gen lscrap1 lscrap[_n1]. gen lscrapd lscrap  lscrap1 (363 missing values generated). quietly by fcode: gen grant1 grant[_n1]. gen grantd grant  grant1 (157 missing values generated). quietly by fcode: gen grant1_1 grant_1[_n1]. gen grant_1d lscrap  grant1_1 (363 missing values generated) First Difference Estimator:. reg lscrapd grantd grant_1d Source SS df MS Number of obs F( 2, 105) 3.54 Model Prob > F Residual Rsquared Adj Rsquared Total Root MSE lscrapd Coef. Std. Err. t P> t [95% Conf. Interval] grantd grant_1d _cons xtreg lscrap grant grant_1, fe Fixedeffects (within) regression Number of obs 162 Group variable (i): fcode Number of groups 54 Rsq: within Obs per group: min 3 between overall avg max F(2,106) corr(u_i, Xb) Prob > F 0 lscrap Coef. Std. Err. t P> t [95% Conf. Interval] grant grant_ _cons
8 sigma_u sigma_e rho (fraction of variance due to u_i) F test that all u_i0: F(53, 106) Prob > F 0. xtreg lscrap grant grant_1 union, fe Fixedeffects (within) regression Number of obs 162 Group variable (i): fcode Number of groups 54 Rsq: within Obs per group: min 3 between overall avg max F(2,106) corr(u_i, Xb) Prob > F 0 lscrap Coef. Std. Err. t P> t [95% Conf. Interval] grant grant_ union (dropped) _cons sigma_u sigma_e rho (fraction of variance due to u_i) F test that all u_i0: F(53, 106) Prob > F 0. xtreg lscrap grant grant_1, re Randomeffects GLS regression Number of obs 162 Group variable (i): fcode Number of groups 54 Rsq: within between Obs per group: min avg overall max 3 Random effects u_i ~ Gaussian Wald chi2(2) corr(u_i, X) 0 (assumed) Prob > chi2 0 lscrap Coef. Std. Err. z P> z [95% Conf. Interval] grant grant_ _cons sigma_u sigma_e rho (fraction of variance due to u_i). xtreg lscrap grant grant_1 d88 d89, fe Fixedeffects (within) regression Number of obs 162 Group variable (i): fcode Number of groups 54
9 Rsq: within between Obs per group: min avg overall max 3 F(4,104) 6.54 corr(u_i, Xb) Prob > F 1 lscrap Coef. Std. Err. t P> t [95% Conf. Interval] grant grant_ d d _cons sigma_u sigma_e rho (fraction of variance due to u_i) F test that all u_i0: F(53, 104) Prob > F 0. hausman, save. xtreg lscrap grant grant_1 d88 d89, re Randomeffects GLS regression Number of obs 162 Group variable (i): fcode Number of groups 54 Rsq: within Obs per group: min 3 between overall avg max Random effects u_i ~ Gaussian Wald chi2(4) corr(u_i, X) 0 (assumed) Prob > chi2 0 lscrap Coef. Std. Err. z P> z [95% Conf. Interval] grant grant_ d d _cons sigma_u sigma_e rho (fraction of variance due to u_i). hausman  Coefficients  (b) (B) (bb) sqrt(diag(v_bv_b)) Consistent Efficient Difference S.E. grant grant_1 d d b consistent under Ho and Ha; obtained from xtreg B inconsistent under Ha, efficient under Ho; obtained from xtreg
10 Test: Ho: difference in coefficients not systematic chi2(4) (bb)'[(v_bv_b)^(1)](bb) 2.14 Prob>chi
11 IV & 2SLS Methods  Stata Output Do file for IV/2SLS lecture.. clear. use Mroz.dta IV. reg lwage educ Source SS df MS Number of obs F( 1, 426) Model Prob > F 0 Residual Rsquared Adj Rsquared Total Root MSE educ _cons ivreg lwage (educfatheduc) Instrumental variables (2SLS) regression Source SS df MS Number of obs F( 1, 426) 2.84 Model Prob > F Residual Rsquared Adj Rsquared Total Root MSE educ _cons Instrumented: educ Instruments: fatheduc 2SLS. regress lwage educ exper expersq Source SS df MS Number of obs F( 3, 424) Model Prob > F 0 Residual Rsquared Adj Rsquared Total Root MSE educ
12 exper expersq _cons regress lwage educ exper expersq (exper expersq motheduc fatheduc) Instrumental variables (2SLS) regression Source SS df MS Number of obs F( 3, 424) 8.14 Model Prob > F 0 Residual Rsquared Adj Rsquared Total Root MSE educ exper expersq _cons Tests: (1) poor instruments:. reg educ exper expersq motheduc fatheduc Source SS df MS Number of obs F( 4, 748) Model Prob > F 0 Residual Rsquared Adj Rsquared Total Root MSE educ Coef. Std. Err. t P> t [95% Conf. Interval] exper expersq motheduc fatheduc _cons test motheduc fatheduc ( 1) motheduc 0 ( 2) fatheduc 0 F( 2, 748) Prob > F 0 (2) over identifying restrictions:. regress lwage educ exper expersq (exper expersq motheduc fatheduc) Instrumental variables (2SLS) regression Source SS df MS Number of obs 428
13 F( 3, 424) 8.14 Model Prob > F 0 Residual Rsquared Adj Rsquared Total Root MSE educ exper expersq _cons predict u2sls, resid (325 missing values generated). regress u2sls exper expersq motheduc fatheduc Source SS df MS Number of obs F( 4, 423) 0.09 Model Prob > F Residual Rsquared Adj Rsquared Total Root MSE u2sls Coef. Std. Err. t P> t [95% Conf. Interval] exper expersq 7.34e07 motheduc fatheduc _cons test motheduc fatheduc ( 1) motheduc 0 ( 2) fatheduc 0 F( 2, 423) 0.19 Prob > F
14 . (3) Hausman test:. regress lwage educ exper expersq (exper expersq motheduc fatheduc) Instrumental variables (2SLS) regression Source SS df MS Number of obs F( 3, 424) 8.14 Model Prob > F 0 Residual Rsquared Adj Rsquared Total Root MSE educ exper expersq _cons hausman, save. regress lwage educ exper expersq Source SS df MS Number of obs F( 3, 424) Model Prob > F 0 Residual Rsquared Adj Rsquared Total Root MSE educ exper expersq _cons hausman  Coefficients  (b) Consistent (B) Efficient (bb) Difference sqrt(diag(v_bv_b)) S.E. educ exper expersq b consistent under Ho and Ha; obtained from regress B inconsistent under Ha, efficient under Ho; obtained from regress Test: Ho: difference in coefficients not systematic chi2(3) (bb)'[(v_bv_b)^(1)](bb) 2.70 Prob>chi
15 Probit, Logit & Tobit Methods Stata Output. clear. use Mroz.dta Linear Probability Model. reg inlf nwifeinc educ exper expersq Source SS df MS Number of obs F( 4, 748) Model Prob > F 0 Residual Rsquared Adj Rsquared Total Root MSE inlf Coef. Std. Err. t P> t [95% Conf. Interval] nwifeinc educ exper expersq _cons mfx Marginal effects after regress y Fitted values (predict) variable dy/dx Std. Err. z P> z [ 95% C.I. ] X nwifeinc educ exper expersq Probit.. probit inlf nwifeinc educ exper expersq Iteration 0: log likelihood Iteration 1: Iteration 2: log likelihood log likelihood Iteration 3: log likelihood Probit estimates Number of obs 753 LR chi2(4) Log likelihood Prob > chi2 Pseudo R inlf Coef. Std. Err. z P> z [95% Conf. Interval]
16 nwifeinc educ exper expersq _cons mfx Marginal effects after probit y Pr(inlf) (predict) variable dy/dx Std. Err. z P> z [ 95% C.I. ] X nwifeinc educ exper expersq Logit.. logit inlf nwifeinc educ exper expersq Iteration 0: log likelihood Iteration 1: log likelihood Iteration 2: Iteration 3: log likelihood log likelihood Logit estimates Number of obs LR chi2(4) Prob > chi2 0 Log likelihood Pseudo R inlf Coef. Std. Err. z P> z [95% Conf. Interval] nwifeinc educ exper expersq _cons mfx Marginal effects after logit y Pr(inlf) (predict) variable dy/dx Std. Err. z P> z [ 95% C.I. ] X nwifeinc educ exper expersq tobit faminc educ exper expersq, ul(100000) Tobit estimates Number of obs 753
17 LR chi2(3) Log likelihood Prob > chi2 Pseudo R faminc Coef. Std. Err. t P> t [95% Conf. Interval] educ exper expersq _cons _se (Ancillary parameter) Obs. summary: 753 uncensored observations Heckman Correction  Stata Output. logit inlf nwifeinc educ exper expersq age kidslt6 kidsge6 Iteration 0: log likelihood Iteration 1: log likelihood Iteration 2: Iteration 3: log likelihood log likelihood Iteration 4: log likelihood Logit estimates Number of obs 753 LR chi2(7) Log likelihood Prob > chi2 Pseudo R inlf Coef. Std. Err. z P> z [95% Conf. Interval] nwifeinc educ exper expersq age kidslt kidsge6 _cons predict xb, xb. gen smallphinormd(xb). gen largephinormprob(xb). gen lambdasmallphi/largephi. reg lwage educ exper expersq lambda if inlf1 Source SS df MS Number of obs F( 4, 423) Model Prob > F 0 Residual Rsquared Adj Rsquared Total Root MSE.66708
18 educ exper expersq lambda _cons heckman lwage educ exper expersq, select(inlf nwifeinc educ exper expersq age kidslt6 kidsge > 6) twostep Heckman selection model  twostep estimates (regression model with sample selection) Number of obs Censored obs Uncensored obs 428 Wald chi2(6) Prob > chi2 0 Coef. Std. Err. z P> z [95% Conf. Interval] lwage educ exper expersq e06 _cons inlf nwifeinc educ exper expersq age kidslt kidsge _cons mills lambda rho sigma lambda
ESTIMATING AVERAGE TREATMENT EFFECTS: IV AND CONTROL FUNCTIONS, II Jeff Wooldridge Michigan State University BGSE/IZA Course in Microeconometrics
ESTIMATING AVERAGE TREATMENT EFFECTS: IV AND CONTROL FUNCTIONS, II Jeff Wooldridge Michigan State University BGSE/IZA Course in Microeconometrics July 2009 1. Quantile Treatment Effects 2. Control Functions
More informationHURDLE AND SELECTION MODELS Jeff Wooldridge Michigan State University BGSE/IZA Course in Microeconometrics July 2009
HURDLE AND SELECTION MODELS Jeff Wooldridge Michigan State University BGSE/IZA Course in Microeconometrics July 2009 1. Introduction 2. A General Formulation 3. Truncated Normal Hurdle Model 4. Lognormal
More informationDETERMINANTS OF CAPITAL ADEQUACY RATIO IN SELECTED BOSNIAN BANKS
DETERMINANTS OF CAPITAL ADEQUACY RATIO IN SELECTED BOSNIAN BANKS Nađa DRECA International University of Sarajevo nadja.dreca@students.ius.edu.ba Abstract The analysis of a data set of observation for 10
More informationLab 5 Linear Regression with Withinsubject Correlation. Goals: Data: Use the pig data which is in wide format:
Lab 5 Linear Regression with Withinsubject Correlation Goals: Data: Fit linear regression models that account for withinsubject correlation using Stata. Compare weighted least square, GEE, and random
More informationDepartment of Economics Session 2012/2013. EC352 Econometric Methods. Solutions to Exercises from Week 10 + 0.0077 (0.052)
Department of Economics Session 2012/2013 University of Essex Spring Term Dr Gordon Kemp EC352 Econometric Methods Solutions to Exercises from Week 10 1 Problem 13.7 This exercise refers back to Equation
More informationMarginal Effects for Continuous Variables Richard Williams, University of Notre Dame, http://www3.nd.edu/~rwilliam/ Last revised February 21, 2015
Marginal Effects for Continuous Variables Richard Williams, University of Notre Dame, http://www3.nd.edu/~rwilliam/ Last revised February 21, 2015 References: Long 1997, Long and Freese 2003 & 2006 & 2014,
More informationHow Do We Test Multiple Regression Coefficients?
How Do We Test Multiple Regression Coefficients? Suppose you have constructed a multiple linear regression model and you have a specific hypothesis to test which involves more than one regression coefficient.
More informationEconometrics II. Lecture 9: Sample Selection Bias
Econometrics II Lecture 9: Sample Selection Bias Måns Söderbom 5 May 2011 Department of Economics, University of Gothenburg. Email: mans.soderbom@economics.gu.se. Web: www.economics.gu.se/soderbom, www.soderbom.net.
More informationCorrelated Random Effects Panel Data Models
INTRODUCTION AND LINEAR MODELS Correlated Random Effects Panel Data Models IZA Summer School in Labor Economics May 1319, 2013 Jeffrey M. Wooldridge Michigan State University 1. Introduction 2. The Linear
More informationEcon 371 Problem Set #3 Answer Sheet
Econ 371 Problem Set #3 Answer Sheet 4.1 In this question, you are told that a OLS regression analysis of third grade test scores as a function of class size yields the following estimated model. T estscore
More informationColombian industrial structure behavior and its regions between 1974 and 2005.
Colombian industrial structure behavior and its regions between 1974 and 2005. Luis Fernando López Pineda Director of Center of Research to Development and Compeveness Chief Economic Research Columbus,
More informationPanel Data Analysis Fixed and Random Effects using Stata (v. 4.2)
Panel Data Analysis Fixed and Random Effects using Stata (v. 4.2) Oscar TorresReyna otorres@princeton.edu December 2007 http://dss.princeton.edu/training/ Intro Panel data (also known as longitudinal
More informationIntroduction to Stata
Introduction to Stata September 23, 2014 Stata is one of a few statistical analysis programs that social scientists use. Stata is in the midrange of how easy it is to use. Other options include SPSS,
More informationFailure to take the sampling scheme into account can lead to inaccurate point estimates and/or flawed estimates of the standard errors.
Analyzing Complex Survey Data: Some key issues to be aware of Richard Williams, University of Notre Dame, http://www3.nd.edu/~rwilliam/ Last revised January 24, 2015 Rather than repeat material that is
More informationApplied Econometrics. Lecture 15: Sample Selection Bias. Estimation of Nonlinear Models with Panel Data
Applied Econometrics Lecture 15: Sample Selection Bias Estimation of Nonlinear Models with Panel Data Måns Söderbom 13 October 2009 University of Gothenburg. Email: mans:soderbom@economics:gu:se. Web:
More informationDiscussion Section 4 ECON 139/239 2010 Summer Term II
Discussion Section 4 ECON 139/239 2010 Summer Term II 1. Let s use the CollegeDistance.csv data again. (a) An education advocacy group argues that, on average, a person s educational attainment would increase
More informationECON Introductory Econometrics Seminar 9
ECON4150  Introductory Econometrics Seminar 9 Stock and Watson EE13.1 April 28, 2015 Stock and Watson EE13.1 ECON4150  Introductory Econometrics Seminar 9 April 28, 2015 1 / 15 Empirical exercise E13.1:
More informationRegression Analysis. Data Calculations Output
Regression Analysis In an attempt to find answers to questions such as those posed above, empirical labour economists use a useful tool called regression analysis. Regression analysis is essentially a
More informationHandling missing data in Stata a whirlwind tour
Handling missing data in Stata a whirlwind tour 2012 Italian Stata Users Group Meeting Jonathan Bartlett www.missingdata.org.uk 20th September 2012 1/55 Outline The problem of missing data and a principled
More informationExam and Solution. Please discuss each problem on a separate sheet of paper, not just on a separate page!
Econometrics  Exam 1 Exam and Solution Please discuss each problem on a separate sheet of paper, not just on a separate page! Problem 1: (20 points A health economist plans to evaluate whether screening
More informationSTATA FUNDAMENTALS FOR MIDDLEBURY COLLEGE ECONOMICS STUDENTS
STATA FUNDAMENTALS FOR MIDDLEBURY COLLEGE ECONOMICS STUDENTS BY EMILY FORREST AUGUST 2008 CONTENTS INTRODUCTION STATA SYNTAX DATASET FILES OPENING A DATASET FROM EXCEL TO STATA WORKING WITH LARGE DATASETS
More informationECON 142 SKETCH OF SOLUTIONS FOR APPLIED EXERCISE #2
University of California, Berkeley Prof. Ken Chay Department of Economics Fall Semester, 005 ECON 14 SKETCH OF SOLUTIONS FOR APPLIED EXERCISE # Question 1: a. Below are the scatter plots of hourly wages
More informationUsing Stata 11 & higher for Logistic Regression Richard Williams, University of Notre Dame, Last revised March 28, 2015
Using Stata 11 & higher for Logistic Regression Richard Williams, University of Notre Dame, http://www3.nd.edu/~rwilliam/ Last revised March 28, 2015 NOTE: The routines spost13, lrdrop1, and extremes are
More informationREGRESSION LINES IN STATA
REGRESSION LINES IN STATA THOMAS ELLIOTT 1. Introduction to Regression Regression analysis is about eploring linear relationships between a dependent variable and one or more independent variables. Regression
More informationQuantitative Methods for Economics Tutorial 9. Katherine Eyal
Quantitative Methods for Economics Tutorial 9 Katherine Eyal TUTORIAL 9 4 October 2010 ECO3021S Part A: Problems 1. In Problem 2 of Tutorial 7, we estimated the equation ŝleep = 3, 638.25 0.148 totwrk
More informationxtmixed & denominator degrees of freedom: myth or magic
xtmixed & denominator degrees of freedom: myth or magic 2011 Chicago Stata Conference Phil Ender UCLA Statistical Consulting Group July 2011 Phil Ender xtmixed & denominator degrees of freedom: myth or
More informationStatistical Modelling in Stata 5: Linear Models
Statistical Modelling in Stata 5: Linear Models Mark Lunt Arthritis Research UK Centre for Excellence in Epidemiology University of Manchester 08/11/2016 Structure This Week What is a linear model? How
More informationStatistics 104 Final Project A Culture of Debt: A Study of Credit Card Spending in America TF: Kevin Rader Anonymous Students: LD, MH, IW, MY
Statistics 104 Final Project A Culture of Debt: A Study of Credit Card Spending in America TF: Kevin Rader Anonymous Students: LD, MH, IW, MY ABSTRACT: This project attempted to determine the relationship
More informationAnalysis of Longitudinal Data in Stata, Splus and SAS
Analysis of Longitudinal Data in Stata, Splus and SAS Rino Bellocco, Sc.D. Department of Medical Epidemiology Karolinska Institutet Stockholm, Sweden rino@mep.ki.se March 12, 2001 NASUGS, 2001 OUTLINE
More informationIn Chapter 2, we used linear regression to describe linear relationships. The setting for this is a
Math 143 Inference on Regression 1 Review of Linear Regression In Chapter 2, we used linear regression to describe linear relationships. The setting for this is a bivariate data set (i.e., a list of cases/subjects
More informationOutline. Topic 4  Analysis of Variance Approach to Regression. Partitioning Sums of Squares. Total Sum of Squares. Partitioning sums of squares
Topic 4  Analysis of Variance Approach to Regression Outline Partitioning sums of squares Degrees of freedom Expected mean squares General linear test  Fall 2013 R 2 and the coefficient of correlation
More informationInteraction effects between continuous variables (Optional)
Interaction effects between continuous variables (Optional) Richard Williams, University of Notre Dame, http://www.nd.edu/~rwilliam/ Last revised February 0, 05 This is a very brief overview of this somewhat
More informationECON Introductory Econometrics. Lecture 15: Binary dependent variables
ECON4150  Introductory Econometrics Lecture 15: Binary dependent variables Monique de Haan (moniqued@econ.uio.no) Stock and Watson Chapter 11 Lecture Outline 2 The linear probability model Nonlinear probability
More informationLecture 15. Endogeneity & Instrumental Variable Estimation
Lecture 15. Endogeneity & Instrumental Variable Estimation Saw that measurement error (on right hand side) means that OLS will be biased (biased toward zero) Potential solution to endogeneity instrumental
More informationMilk Data Analysis. 1. Objective Introduction to SAS PROC MIXED Analyzing protein milk data using STATA Refit protein milk data using PROC MIXED
1. Objective Introduction to SAS PROC MIXED Analyzing protein milk data using STATA Refit protein milk data using PROC MIXED 2. Introduction to SAS PROC MIXED The MIXED procedure provides you with flexibility
More informationStandard errors of marginal effects in the heteroskedastic probit model
Standard errors of marginal effects in the heteroskedastic probit model Thomas Cornelißen Discussion Paper No. 320 August 2005 ISSN: 0949 9962 Abstract In nonlinear regression models, such as the heteroskedastic
More informationMULTIPLE REGRESSION EXAMPLE
MULTIPLE REGRESSION EXAMPLE For a sample of n = 166 college students, the following variables were measured: Y = height X 1 = mother s height ( momheight ) X 2 = father s height ( dadheight ) X 3 = 1 if
More informationPanel Data Analysis Josef Brüderl, University of Mannheim, March 2005
Panel Data Analysis Josef Brüderl, University of Mannheim, March 2005 This is an introduction to panel data analysis on an applied level using Stata. The focus will be on showing the "mechanics" of these
More informationSample Size Calculation for Longitudinal Studies
Sample Size Calculation for Longitudinal Studies Phil Schumm Department of Health Studies University of Chicago August 23, 2004 (Supported by National Institute on Aging grant P01 AG1891101A1) Introduction
More informationA Panel Data Analysis of Corporate Attributes and Stock Prices for Indian Manufacturing Sector
Journal of Modern Accounting and Auditing, ISSN 15486583 November 2013, Vol. 9, No. 11, 15191525 D DAVID PUBLISHING A Panel Data Analysis of Corporate Attributes and Stock Prices for Indian Manufacturing
More informationGroup Comparisons: Differences in Composition Versus Differences in Models and Effects
Group Comparisons: Differences in Composition Versus Differences in Models and Effects Richard Williams, University of Notre Dame, http://www3.nd.edu/~rwilliam/ Last revised February 15, 2015 Overview.
More informationSoci708 Statistics for Sociologists
Soci708 Statistics for Sociologists Module 11 Multiple Regression 1 François Nielsen University of North Carolina Chapel Hill Fall 2009 1 Adapted from slides for the course Quantitative Methods in Sociology
More informationInteraction Terms Vs. Interaction Effects in Logistic and Probit Regression
 Background: In probit or logistic regressions, one can not base statistical inferences based on simply looking at the coefficient and statistical significance of
More informationC2.1. (i) (5 marks) The average participation rate is , the average match rate is summ prate
BOSTON COLLEGE Department of Economics EC 228 01 Econometric Methods Fall 2008, Prof. Baum, Ms. Phillips (tutor), Mr. Dmitriev (grader) Problem Set 2 Due at classtime, Thursday 2 Oct 2008 2.4 (i)(5 marks)
More informationRegression in ANOVA. James H. Steiger. Department of Psychology and Human Development Vanderbilt University
Regression in ANOVA James H. Steiger Department of Psychology and Human Development Vanderbilt University James H. Steiger (Vanderbilt University) 1 / 30 Regression in ANOVA 1 Introduction 2 Basic Linear
More informationModern Methods for Missing Data
Modern Methods for Missing Data Paul D. Allison, Ph.D. Statistical Horizons LLC www.statisticalhorizons.com 1 Introduction Missing data problems are nearly universal in statistical practice. Last 25 years
More informationDepartment of Economics, Session 2012/2013. EC352 Econometric Methods. Exercises from Week 03
Department of Economics, Session 01/013 University of Essex, Autumn Term Dr Gordon Kemp EC35 Econometric Methods Exercises from Week 03 1 Problem P3.11 The following equation describes the median housing
More informationMODEL I: DRINK REGRESSED ON GPA & MALE, WITHOUT CENTERING
Interpreting Interaction Effects; Interaction Effects and Centering Richard Williams, University of Notre Dame, http://www3.nd.edu/~rwilliam/ Last revised February 20, 2015 Models with interaction effects
More informationEcon 371 Problem Set #3 Answer Sheet
Econ 371 Problem Set #3 Answer Sheet 4.3 In this question, you are told that a OLS regression analysis of average weekly earnings yields the following estimated model. AW E = 696.7 + 9.6 Age, R 2 = 0.023,
More informationMultiple Linear Regression
Multiple Linear Regression A regression with two or more explanatory variables is called a multiple regression. Rather than modeling the mean response as a straight line, as in simple regression, it is
More informationDoes corporate performance predict the cost of equity capital?
AMERICAN JOURNAL OF SOCIAL AND MANAGEMENT SCIENCES ISSN Print: 21561540, ISSN Online: 21511559, doi:10.5251/ajsms.2011.2.1.26.33 2010, ScienceHuβ, http://www.scihub.org/ajsms Does corporate performance
More informationRegression in Stata. Alicia Doyle Lynch HarvardMIT Data Center (HMDC)
Regression in Stata Alicia Doyle Lynch HarvardMIT Data Center (HMDC) Documents for Today Find class materials at: http://libraries.mit.edu/guides/subjects/data/ training/workshops.html Several formats
More informationNonlinear relationships Richard Williams, University of Notre Dame, http://www3.nd.edu/~rwilliam/ Last revised February 20, 2015
Nonlinear relationships Richard Williams, University of Notre Dame, http://www.nd.edu/~rwilliam/ Last revised February, 5 Sources: Berry & Feldman s Multiple Regression in Practice 985; Pindyck and Rubinfeld
More informationFrom this it is not clear what sort of variable that insure is so list the first 10 observations.
MNL in Stata We have data on the type of health insurance available to 616 psychologically depressed subjects in the United States (Tarlov et al. 1989, JAMA; Wells et al. 1989, JAMA). The insurance is
More informationPaired Differences and Regression
Paired Differences and Regression Students sometimes have difficulty distinguishing between paired data and independent samples when comparing two means. One can return to this topic after covering simple
More informationAugust 2012 EXAMINATIONS Solution Part I
August 01 EXAMINATIONS Solution Part I (1) In a random sample of 600 eligible voters, the probability that less than 38% will be in favour of this policy is closest to (B) () In a large random sample,
More informationLecture 13. Use and Interpretation of Dummy Variables. Stop worrying for 1 lecture and learn to appreciate the uses that dummy variables can be put to
Lecture 13. Use and Interpretation of Dummy Variables Stop worrying for 1 lecture and learn to appreciate the uses that dummy variables can be put to Using dummy variables to measure average differences
More informationLecture 16. Endogeneity & Instrumental Variable Estimation (continued)
Lecture 16. Endogeneity & Instrumental Variable Estimation (continued) Seen how endogeneity, Cov(x,u) 0, can be caused by Omitting (relevant) variables from the model Measurement Error in a right hand
More informationPlease follow the directions once you locate the Stata software in your computer. Room 114 (Business Lab) has computers with Stata software
STATA Tutorial Professor Erdinç Please follow the directions once you locate the Stata software in your computer. Room 114 (Business Lab) has computers with Stata software 1.Wald Test Wald Test is used
More informationLinear Regression with One Regressor
Linear Regression with One Regressor Michael Ash Lecture 10 Analogy to the Mean True parameter µ Y β 0 and β 1 Meaning Central tendency Intercept and slope E(Y ) E(Y X ) = β 0 + β 1 X Data Y i (X i, Y
More informationQuick Stata Guide by Liz Foster
by Liz Foster Table of Contents Part 1: 1 describe 1 generate 1 regress 3 scatter 4 sort 5 summarize 5 table 6 tabulate 8 test 10 ttest 11 Part 2: Prefixes and Notes 14 by var: 14 capture 14 use of the
More informationGeneralized Linear Models
Generalized Linear Models We have previously worked with regression models where the response variable is quantitative and normally distributed. Now we turn our attention to two types of models where the
More informationInteraction effects and group comparisons Richard Williams, University of Notre Dame, http://www3.nd.edu/~rwilliam/ Last revised February 20, 2015
Interaction effects and group comparisons Richard Williams, University of Notre Dame, http://www3.nd.edu/~rwilliam/ Last revised February 20, 2015 Note: This handout assumes you understand factor variables,
More informationTitle. Syntax. nlcom Nonlinear combinations of estimators. Nonlinear combination of estimators one expression. nlcom [ name: ] exp [, options ]
Title nlcom Nonlinear combinations of estimators Syntax Nonlinear combination of estimators one expression nlcom [ name: ] exp [, options ] Nonlinear combinations of estimators more than one expression
More informationis paramount in advancing any economy. For developed countries such as
Introduction The provision of appropriate incentives to attract workers to the health industry is paramount in advancing any economy. For developed countries such as Australia, the increasing demand for
More informationQuantitative Methods for Economics Tutorial 12. Katherine Eyal
Quantitative Methods for Economics Tutorial 12 Katherine Eyal TUTORIAL 12 25 October 2010 ECO3021S Part A: Problems 1. State with brief reason whether the following statements are true, false or uncertain:
More informationA Simple Feasible Alternative Procedure to Estimate Models with HighDimensional Fixed Effects
DISCUSSION PAPER SERIES IZA DP No. 3935 A Simple Feasible Alternative Procedure to Estimate Models with HighDimensional Fixed Effects Paulo Guimarães Pedro Portugal January 2009 Forschungsinstitut zur
More informationECON Introductory Econometrics. Lecture 17: Experiments
ECON4150  Introductory Econometrics Lecture 17: Experiments Monique de Haan (moniqued@econ.uio.no) Stock and Watson Chapter 13 Lecture outline 2 Why study experiments? The potential outcome framework.
More informationWe extended the additive model in two variables to the interaction model by adding a third term to the equation.
Quadratic Models We extended the additive model in two variables to the interaction model by adding a third term to the equation. Similarly, we can extend the linear model in one variable to the quadratic
More informationResiduals. Residuals = ª Department of ISM, University of Alabama, ST 260, M23 Residuals & Minitab. ^ e i = y i  y i
A continuation of regression analysis Lesson Objectives Continue to build on regression analysis. Learn how residual plots help identify problems with the analysis. M231 M232 Example 1: continued Case
More informationRockefeller College University at Albany
Rockefeller College University at Albany PAD 705 Handout:, the DurbinWatson Statistic, and the CochraneOrcutt Procedure Serial correlation (also called autocorrelation ) is said to exist when the error
More informationImpact of working capital management on profitability: the case of Canadian firms
Impact of working capital management on profitability: the case of Canadian firms By Ruichao Lu A00320698 A Research Project Submitted in Partial fulfillment of the Requirements Of The Degree of Master
More informationCorrelation and Regression
Correlation and Regression Scatterplots Correlation Explanatory and response variables Simple linear regression General Principles of Data Analysis First plot the data, then add numerical summaries Look
More information1. What is the critical value for this 95% confidence interval? CV = z.025 = invnorm(0.025) = 1.96
1 Final Review 2 Review 2.1 CI 1propZint Scenario 1 A TV manufacturer claims in its warranty brochure that in the past not more than 10 percent of its TV sets needed any repair during the first two years
More informationIAPRI Quantitative Analysis Capacity Building Series. Multiple regression analysis & interpreting results
IAPRI Quantitative Analysis Capacity Building Series Multiple regression analysis & interpreting results How important is Rsquared? Rsquared Published in Agricultural Economics 0.45 Best article of the
More informationUsing Minitab for Regression Analysis: An extended example
Using Minitab for Regression Analysis: An extended example The following example uses data from another text on fertilizer application and crop yield, and is intended to show how Minitab can be used to
More informationEcon 371 Problem Set #4 Answer Sheet. P rice = (0.485)BDR + (23.4)Bath + (0.156)Hsize + (0.002)LSize + (0.090)Age (48.
Econ 371 Problem Set #4 Answer Sheet 6.5 This question focuses on what s called a hedonic regression model; i.e., where the sales price of the home is regressed on the various attributes of the home. The
More informationRegression of Systolic Blood Pressure on Age, Weight & Cholesterol
Regression of Systolic Blood Pressure on Age, Weight & Cholesterol 1 * bp.sas; 2 options ls=120 ps=75 nocenter nodate; 3 title Regression of Systolic Blood Pressure on Age, Weight & Cholesterol ; 4 * BP
More informationData and Regression Analysis. Lecturer: Prof. Duane S. Boning. Rev 10
Data and Regression Analysis Lecturer: Prof. Duane S. Boning Rev 10 1 Agenda 1. Comparison of Treatments (One Variable) Analysis of Variance (ANOVA) 2. Multivariate Analysis of Variance Model forms 3.
More informationoutreg help pages Write formatted regression output to a text file After any estimation command: (Textrelated options)
outreg help pages OUTREG HELP PAGES... 1 DESCRIPTION... 2 OPTIONS... 3 1. Textrelated options... 3 2. Coefficient options... 4 3. Options for t statistics, standard errors, etc... 5 4. Statistics options...
More informationBinary Dependent Variables. In some cases the outcome of interest rather than one of the right hand side variables is discrete rather than continuous
Bnary Dependent Varables In some cases the outcome of nterest rather than one of the rght hand sde varables s dscrete rather than contnuous The smplest example of ths s when the Y varable s bnary so that
More informationLectures 8, 9 & 10. Multiple Regression Analysis
Lectures 8, 9 & 0. Multiple Regression Analysis In which you learn how to apply the principles and tests outlined in earlier lectures to more realistic models involving more than explanatory variable and
More informationStat 5303 (Oehlert): Tukey One Degree of Freedom 1
Stat 5303 (Oehlert): Tukey One Degree of Freedom 1 > catch
More informationLogistic Regression, Part III: Hypothesis Testing, Comparisons to OLS
Logistic Regression, Part III: Hypothesis Testing, Comparisons to OLS Richard Williams, University of Notre Dame, http://www3.nd.edu/~rwilliam/ Last revised February 22, 2015 This handout steals heavily
More informationFrom the help desk: Swamy s randomcoefficients model
The Stata Journal (2003) 3, Number 3, pp. 302 308 From the help desk: Swamy s randomcoefficients model Brian P. Poi Stata Corporation Abstract. This article discusses the Swamy (1970) randomcoefficients
More informationTesting for serial correlation in linear paneldata models
The Stata Journal (2003) 3, Number 2, pp. 168 177 Testing for serial correlation in linear paneldata models David M. Drukker Stata Corporation Abstract. Because serial correlation in linear paneldata
More informationMulticollinearity Richard Williams, University of Notre Dame, http://www3.nd.edu/~rwilliam/ Last revised January 13, 2015
Multicollinearity Richard Williams, University of Notre Dame, http://www3.nd.edu/~rwilliam/ Last revised January 13, 2015 Stata Example (See appendices for full example).. use http://www.nd.edu/~rwilliam/stats2/statafiles/multicoll.dta,
More informationBIOS 312: MODERN REGRESSION ANALYSIS
BIOS 312: MODERN REGRESSION ANALYSIS James C (Chris) Slaughter Department of Biostatistics Vanderbilt University School of Medicine james.c.slaughter@vanderbilt.edu biostat.mc.vanderbilt.edu/coursebios312
More information1.1. Simple Regression in Excel (Excel 2010).
.. Simple Regression in Excel (Excel 200). To get the Data Analysis tool, first click on File > Options > AddIns > Go > Select Data Analysis Toolpack & Toolpack VBA. Data Analysis is now available under
More informationInference for Regression
Simple Linear Regression Inference for Regression The simple linear regression model Estimating regression parameters; Confidence intervals and significance tests for regression parameters Inference about
More informationBasic Statistical and Modeling Procedures Using SAS
Basic Statistical and Modeling Procedures Using SAS OneSample Tests The statistical procedures illustrated in this handout use two datasets. The first, Pulse, has information collected in a classroom
More informationRockefeller College University at Albany
Rockefeller College University at Albany PAD 705 Handout: Hypothesis Testing on Multiple Parameters In many cases we may wish to know whether two or more variables are jointly significant in a regression.
More informationSCHOOL OF MATHEMATICS AND STATISTICS
RESTRICTED OPEN BOOK EXAMINATION (Not to be removed from the examination hall) Data provided: Statistics Tables by H.R. Neave MAS5052 SCHOOL OF MATHEMATICS AND STATISTICS Basic Statistics Spring Semester
More informationFrom the help desk: hurdle models
The Stata Journal (2003) 3, Number 2, pp. 178 184 From the help desk: hurdle models Allen McDowell Stata Corporation Abstract. This article demonstrates that, although there is no command in Stata for
More informationLecture 16: Logistic regression diagnostics, splines and interactions. Sandy Eckel 19 May 2007
Lecture 16: Logistic regression diagnostics, splines and interactions Sandy Eckel seckel@jhsph.edu 19 May 2007 1 Logistic Regression Diagnostics Graphs to check assumptions Recall: Graphing was used to
More information121 Multiple Linear Regression Models
121.1 Introduction Many applications of regression analysis involve situations in which there are more than one regressor variable. A regression model that contains more than one regressor variable is
More informationespecially with continuous
Handling interactions in Stata, especially with continuous predictors Patrick Royston & Willi Sauerbrei German Stata Users meeting, Berlin, 1 June 2012 Interactions general concepts General idea of a (twoway)
More informationI n d i a n a U n i v e r s i t y U n i v e r s i t y I n f o r m a t i o n T e c h n o l o g y S e r v i c e s
I n d i a n a U n i v e r s i t y U n i v e r s i t y I n f o r m a t i o n T e c h n o l o g y S e r v i c e s Linear Regression Models for Panel Data Using SAS, Stata, LIMDEP, and SPSS * Hun Myoung Park,
More informationStata Walkthrough 4: Regression, Prediction, and Forecasting
Stata Walkthrough 4: Regression, Prediction, and Forecasting Over drinks the other evening, my neighbor told me about his 25yearold nephew, who is dating a 35yearold woman. God, I can t see them getting
More informationHeteroskedasticity Richard Williams, University of Notre Dame, Last revised January 30, 2015
Heteroskedasticity Richard Williams, University of Notre Dame, http://www3.nd.edu/~rwilliam/ Last revised January 30, 2015 These notes draw heavily from Berry and Feldman, and, to a lesser extent, Allison,
More information