Revisiting Titanic Case-study

This is an R Markdown document which gives the extended analysis of the survivors of the Titanic dataset case-study.

Create Dataframe of Titanic

setwd("~/Desktop/Data Analytics Internship/Titanic Case study")
titanic <- read.csv(file="Titanic Data.csv")

Use R to create a table showing the average age of the survivors and the average age of the people who died.

aggregate(Age~Survived, data = titanic, FUN= mean)
##   Survived      Age
## 1        0 30.41530
## 2        1 28.42382

Hence, the average age of survivors is approx. 28 and average age of people who died is 31.

Use R to run a t-test to test the following hypothesis:

H2: The Titanic survivors were younger than the passengers who died.

t.test(titanic$Age~titanic$Survived, var.equal = TRUE)
## 
##  Two Sample t-test
## 
## data:  titanic$Age by titanic$Survived
## t = 2.2302, df = 887, p-value = 0.02599
## alternative hypothesis: true difference in means is not equal to 0
## 95 percent confidence interval:
##  0.238890 3.744064
## sample estimates:
## mean in group 0 mean in group 1 
##        30.41530        28.42382

P Value=0.02599, which is P<0.05, hence the hypothesis serves well on the data, and the titanic survivors were younger than the passengers who died.