a table showing the average age of the survivors and the average age of the people who died.
titanic.df <- read.csv(paste("Titanic Data.csv", sep=""))
aggregate((titanic.df$Age~titanic.df$Survived), FUN=mean)
## titanic.df$Survived titanic.df$Age
## 1 0 30.41530
## 2 1 28.42382
a t-test to test the following hypothesis:
H2: The Titanic survivors were younger than the passengers who died.
t.test(Age~Survived,data=titanic.df)
##
## Welch Two Sample t-test
##
## data: Age by Survived
## t = 2.1816, df = 667.56, p-value = 0.02949
## alternative hypothesis: true difference in means is not equal to 0
## 95 percent confidence interval:
## 0.1990628 3.7838912
## sample estimates:
## mean in group 0 mean in group 1
## 30.41530 28.42382
from the above t-test performed on the data we can observe that the p-value is 0.02449(<0.05) therefore we reject null hypothesis and accept the hypothesis H2 that is titanic survivors were younger then the passengers who died.