Link functions, Advanced Statistics

Assignment Help:

Link functions:

The link function relates the linear predictor ηi to the expected value of the data. In classical linear models the mean and the linear predictor are identical. However, when dealing with counts and the distribution is Poisson, we must have the Poisson distribution parameter satisfy µi > 0 so that the identity link is less attractive, partly because ηi may be negative while µi > 0 must not be. It is advisable to utilize the log link η = log(µ) in this case. Similarly, when dealing with binomial distribution, the parameter p of probability of success in a single trial is restricted to be in (0,1) and the link function serves the purpose to map the interval (0,1) on to R1 . Therefore, links like the following (here µ is replaced by p):

614_Link functions6.png

have been suggested and widely used.

Let us illustrate the most commonly met examples of GLM together with describing the types of response variable, distribution, and the link function:

- Traditional Linear Model:

i) response variable: continuous

ii) distribution: normal

iii) link function: identity : η = µ

- Logistic Regression:

i) response variable: probability ( µ)

ii) distribution: binomial

iii) link function: logit: η = log( µ/1-µ)

- Poisson Regression in Log Linear Model:

i) response variable: count

ii) distribution: Poisson
iii) link function: η = log(µ)
- Gamma model with Log Link:
i) response variable: a positive continuous variable
ii) distribution: Gamma
iii) link function: η = log(µ)

Intermezzo and history. If you read di?erent references, you may get confused about the terminology. You may have already come across the term "general linear model" in your introductory Statistics courses or in some reference books. Note, however, that this term refers to a conventional linear regression model for a continuous response variables given continuous and/or categorical predictors. It includes multiple linear regression, as well as ANOVA and ANCOVA. In SAS, such models are ?t by least squares and weighted least squares using (typically) proc glm. HOWEVER, the "generalized linear model" we are speaking about here, refers to the larger class discussed in this section. The ?rst widely used software package for ?tting these models was called GLIM. Because of this program, "GLIM" became a well-accepted abbreviation for generalized linear models, as opposed to "GLM". Since we clari?ed the confusion though, we will continue using "GLM" for generalized linear models since many recent references use it. Today, generalized linear models are ?t by many packages, notably by the SAS proc genmod. (End of intermezzo).

One of the advantages of the full probabilistic speci?cation of the GLM model is that ML Estimation suggests itself as a natural general estimation method. We have to maximize the log-likelihood

2487_Link functions2.png

where β is linked to θ through the link function. Recall that the main parameter- vector of interest is β, the vector of regression coeffcients in the relation ηi = g(µi) = x0

1971_Link functions3.png

There is nowadays, with the availability of modern computing power, seldom any reason to consider estimators of β that are di?erent from the MLE. By using the chain rule, we get for the components of the score function:

1533_Link functions4.png

The (expected) Fisher information matrix is given then by

687_Link functions5.png

The ML Estimator is de?ned by equating the score function to zero. Numerically, the equation is solved by applying iterative procedures which we discuss next.


Related Discussions:- Link functions

Queuing theory, 1) Let N1(t) and N2(t) be independent Poisson processes wit...

1) Let N1(t) and N2(t) be independent Poisson processes with rates, ?1 and ?2, respectively. Let N (t) = N1(t) + N2(t). a) What is the distribution of the time till the next epoch

Calculate the standard deviation, Q. A toothpaste company want to know if i...

Q. A toothpaste company want to know if its new product increases the length of time in-between dentist visit to its user. The company sets a target for 180 days to determine if it

Huffman coding based compression, Huffman code is used to compress data fil...

Huffman code is used to compress data file, where the data is represented as a sequence of characters. Huffman's greedy algorithm uses a table giving how often each character occur

Student, the problem that demonstrates inference from two dependent samples...

the problem that demonstrates inference from two dependent samples uses hypothetical data from TB vaccinations and the number of new cases before and after vaccinations for cases o

Construct a stem-and-leaf diagram, The number of employees absent from work...

The number of employees absent from work at a large electronics manufacturing plant over aperiod of 106 days is given in the table below. 146 141 139 140 145 141 142 131 142 140

Statistical methods with financial applications, The marketing manager of H...

The marketing manager of Handy Foods Ltd. is concerned with the sales appeal of one of the company's present label for one of its products. Market research indicates that supermark

Em algorithm, The method or technique for producing the sequence of paramet...

The method or technique for producing the sequence of parameter estimates that, under the mild regularity conditions, converges to maximum likelihood estimator. Of particular signi

Please answer this question, How large would the sample need to be if we ar...

How large would the sample need to be if we are to pick a 95% confidence level sample: (i) From a population of 70; (ii) From a population of 450; (iii) From a population of 1000;

Leaps-and-bounds algorithm, Leaps-and-bounds algorithm is an algorithm whi...

Leaps-and-bounds algorithm is an algorithm which is used to ?nd the optimal solution in problems which might have a large number of possible solutions. Begins by dividing the poss

Mortality odds ratio, Mortality odds ratio  is the ratio equivalent to the ...

Mortality odds ratio  is the ratio equivalent to the odds ratio used in case-control studies where the equivalent of the cases are deaths from the cause of interest and the equival

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd