Skip to content
View SamiraSiavash's full-sized avatar

Block or report SamiraSiavash

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
SamiraSiavash/README.md

Hi, I'm Samira Siavash

Data Engineer | Data Warehousing | BI & Data Architecture

I have a background in financial software implementation and support, with hands-on experience working with business processes, databases, and real-world data.

My focus is on building reliable and maintainable data solutions across Data Engineering, Data Warehousing, Business Intelligence, and Data Architecture. I am particularly interested in designing data pipelines, integrating data from multiple sources, transforming and modeling data, and building analytical solutions that support data-driven decision-making.

My technical experience includes SQL Server, PostgreSQL, Python, Apache Airflow, Apache Spark, Apache Kafka, Docker, ClickHouse, Power BI, and Grafana.

I have built practical data projects involving ETL/ELT pipelines, data integration, workflow orchestration, streaming data, data visualization, and analytical dashboards.

I am currently deepening my expertise in Data Warehouse architecture, dimensional modeling, data quality, and modern data platforms, with a particular interest in scalable data architectures and Lakehouse solutions.


πŸ”§ Tech Stack

Programming & Data Processing

  • Python
  • SQL
  • Scala (Apache Spark Fundamentals)
  • Linux

Data Engineering

  • Apache Kafka
  • Apache Airflow
  • Apache Spark
  • ETL / ELT Development
  • Data Cleaning & Transformation
  • Data Integration
  • Data Pipelines

Databases

  • PostgreSQL
  • SQL Server
  • MongoDB
  • ClickHouse

Analytics & Visualization

  • Power BI
  • Grafana

Search & Log Management

  • Elasticsearch
  • Logstash
  • Kibana

Tools

  • Git & GitHub
  • Docker
  • REST APIs
  • Web Scraping
  • JSON / API Data Processing

πŸ”„ SMS ETL Pipeline

End-to-end ETL workflow for extracting, transforming, and loading SMS datasets.

Tech: Python, PostgreSQL, ETL


🌐 Divar API to PostgreSQL

Automated data ingestion pipeline that extracts data from the Divar API and stores it in PostgreSQL.

Tech: Python, REST API, PostgreSQL


βš™οΈ Digikala Price Pipeline

Workflow orchestration and scheduling using Apache Airflow.

Tech: Apache Airflow, Python, PostgreSQL


🧹 Data Cleaning Project

Data preprocessing, cleaning, and transformation workflow for analytics-ready datasets.

Tech: Python, Pandas


πŸ”₯ Spark Data Processing Project

Large-scale dataset processing, cleaning, and transformation using Apache Spark.

Tech: Spark, Python, Scala, Parquet


πŸ›οΈ ClickHouse Analytics Project

Analytical processing and reporting using ClickHouse, PostgreSQL, and Grafana.

Tech: ClickHouse, PostgreSQL, Grafana


πŸ“Š Grafana Analytics Dashboard

Interactive dashboards and KPI monitoring for business analytics.

Tech: Grafana, PostgreSQL


πŸ“ˆ Students Dashboard

Interactive Power BI dashboard for analyzing student performance metrics.

Tech: Power BI, DAX, Data Modeling


🌸 Perfume Scraper

Web scraping project for collecting and storing perfume product data.

Tech: Python, Requests, BeautifulSoup, SQLite


πŸ“š Currently Learning

  • Databricks
  • Lakehouse Architecture
  • Data Warehousing
  • Data Modeling
  • Distributed Data Processing
  • Apache Spark
  • Real-Time Data Pipelines
  • Data Architecture & Design

🎯 Career Focus

  • Data Engineering
  • Data Platforms & Data Pipelines
  • ETL / ELT Development
  • Data Architecture
  • Big Data Processing
  • Analytics Engineering
  • Workflow Orchestration

πŸ“« Connect With Me

LinkedIn: https://linkedin.com/in/samira-siavash

GitHub: https://github.com/SamiraSiavash

Email: [email protected]

Let’s connect and talk about data, BI, and engineering solutions!

Pinned Loading

  1. digikala-price-pipeline digikala-price-pipeline Public

    Daily ETL pipeline that collects Digikala mobile phone prices using Python, Apache Airflow, and PostgreSQL.

    Python

  2. sms-etl-pipeline sms-etl-pipeline Public

    ETL pipeline that processes SMSSpamCollection dataset using Python, Pandas, PostgreSQL, and MongoDB with structured and unstructured data storage.

    Python 1

  3. Perfume_Scraper_Multi_DB Perfume_Scraper_Multi_DB Public

    A complete web scraping pipeline built with Python, Requests, BeautifulSoup, and SQLite, SQL Server, PostgreSQL and MongoDB to collect and store perfume product details.

    Python

  4. divar-api-to-postgres divar-api-to-postgres Public

    A simple end-to-end ETL pipeline that extracts real-world data from Divar API, transforms it using Pandas, and loads it into a PostgreSQL database.

    Python 1

  5. Grafana-Northwind-Dashboard Grafana-Northwind-Dashboard Public

    Grafana dashboard built using the Northwind dataset (JSON export included)

  6. Students-dashboard Students-dashboard Public

    Interactive Power BI Dashboard for Student Performance