All work

02Project work

Data Science Project Work

Data Seekho | DSMP Fellowship

Angled view of an illustrative notebook with PySpark and SQL cells, a result table and a distribution chart.

Overview

Worked with complex datasets and practical data challenges using Python, SQL, and PySpark while strengthening analytical thinking, visualization, and project execution.

Key areas

  • Data manipulation
  • Data cleaning
  • Exploratory analysis
  • SQL analysis
  • PySpark
  • Data visualization
  • Analytical problem solving
Role
DSMP Fellow
Organisation
Data Seekho
Period
Oct 2024 – Mar 2025
Location
Lahore
NOTEBOOK — PYTHON / SQL / PYSPARKFIG. 02 — ILLUSTRATIVE[1]# clean, group, summarisefrom pyspark.sql import functions as Fsummary = (df.dropna(subset=["value"]) .groupBy("segment") .agg(F.avg("value").alias("avg_value")))[2]SELECT segment, COUNT(*) AS nFROM recordsGROUP BY segmentORDER BY n DESC;OUT — TABLEOUT — DISTRIBUTION
Illustrative diagram — not a screenshot of a product or deliverable.

Let's turn data into intelligence.

Have an AI, analytics, business-intelligence, or research problem worth solving?