titanic.df <- read.csv(paste("Titanic Data.csv"), sep= ",")

task 4b

create a table showing the average age of the survivors and the average age of the people who died.

aggregate(titanic.df$Age,by=list(titanic.df$Survived),mean)
##   Group.1        x
## 1       0 30.41530
## 2       1 28.42382

task 4(c)

run a t-test to test the following hypothesis: H2: The Titanic survivors were younger than the passengers who died.

null hypothesis : there is no significant difference between the average age of people who survived and those do not.

 t.test(Age~Survived,data=titanic.df)
## 
##  Welch Two Sample t-test
## 
## data:  Age by Survived
## t = 2.1816, df = 667.56, p-value = 0.02949
## alternative hypothesis: true difference in means is not equal to 0
## 95 percent confidence interval:
##  0.1990628 3.7838912
## sample estimates:
## mean in group 0 mean in group 1 
##        30.41530        28.42382

since the p-value >0.05 we cannot reject null hypothesis. thus there is no significant difference between the average age of people who survived and who didn’t.