12. Exploratory Data Analysis (EDA)
12.2 First Pass Checklist
EDA steps in Jupyter
df.shape,df.head(),df.tail().df.info()anddf.isna().sum().df["order_id"].duplicated().sum().df.describe(include="all")— or split numeric / object.df["city"].value_counts()andnormalize=True.df["status"].value_counts().pd.crosstab(df["city"], df["status"], margins=True).df.groupby("festival")["amount"].agg(["sum", "mean", "count"]).- Write 3 bullet findings in a markdown cell.
What you should see. Pune appears most; Diwali rows show higher amounts in this fictional sample; one Cancelled (Kolhapur) and one Returned (Pune).