setwd("C:/Users/Leo Tolstoy/Downloads")
titanic.df<-read.csv(paste("Titanic Data.csv",sep=""))
View(titanic.df)
aggregate(Age~Survived,data=titanic.df,mean)
## Survived Age
## 1 0 30.41530
## 2 1 28.42382
Average age of survivors is less than that of non-survivors
TASK 4c Use R to run a t-test to test the following hypothesis: H2: The Titanic survivors were younger than the passengers who died.
t.test(titanic.df$Age ~ titanic.df$Survived)
##
## Welch Two Sample t-test
##
## data: titanic.df$Age by titanic.df$Survived
## t = 2.1816, df = 667.56, p-value = 0.02949
## alternative hypothesis: true difference in means is not equal to 0
## 95 percent confidence interval:
## 0.1990628 3.7838912
## sample estimates:
## mean in group 0 mean in group 1
## 30.41530 28.42382
Since the p-value is less than 0.05, thus we can reject the null hypothesis that the ages of the survivors and the dead people are the same. Hence, there is a significant difference between age of survivors and non-survivors of RMS Titanic.