Skip to content
View christiane-bacani's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report christiane-bacani

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
christiane-bacani/README.md

Hi, I'm Christiane 👋

Data Engineering & Analytics · Philippines 🇵🇭

I design and build data pipelines that extract, transform, and load data from whatever sources I find interesting, then turn that data into structured, insight-ready outputs. I'm a Computer Science graduate with strong hands-on experience in Python, SQL, and data handling, and I'm expanding into machine learning, statistics, and data visualization as I work toward covering the full data lifecycle from raw ingestion to business insight. Currently building end-to-end data projects that connect live data pipelines to analytics dashboards and recommendation system that bridge data engineering with real analytical output and machine learning capabilities.


🛠 Tech Stack

  • Languages: Python · SQL
  • Data & Analytics: Pandas · NumPy · Matplotlib · Seaborn · scikit-learn
  • Databases & Warehouses: PostgreSQL · MySQL · SQLite · Snowflake
  • Backend & APIs: SQLAlchemy · FastAPI · Requests · BeautifulSoup
  • Data Visualization: Power BI (in progress) · Matplotlib · Seaborn
  • Tools & Environments: Git · GitHub · Jupyter Notebook · Docker · Ubuntu

🔭 What I'm Working On

  • Extending Steam Charts Tracker: an end-to-end data engineering project tracking game popularity trends, connected to an analytics dashboard using Power BI and Matplotlib/Seaborn Python Library and a game recommendation system powered by machine learning model.
  • Exploring the ELK Stack (Elasticsearch · Logstash · Kibana) for log analytics and data observability via Docker.
  • Deepening applied statistics, machine learning, and data visualization skills through end-to-end projects.
  • Open to Data Engineer, AI/ML Engineer, and Developer roles.

📌 Featured Projects

Designed an enterprise-grade analytical data repository for a real government client, centralizing performance measurements across 37+ departments and offices. The system tracks the Governor's Joint Strategic Goals (GJSG) through a galaxy schema with shared dimensions, a full data catalog, and data lineage documentation that was built with careful attention to referential integrity and cross-departmental query performance.

Real-world client · 37+ departments · Galaxy schema · Full data catalog · 100% referential integrity

Data Modeling PostgreSQL SQL Dimensional Modeling Entity-Relationship Diagram (ERD) Data Catalog Data Lineage Government


Exploratory data analysis of public tweets on Philippine flood control issues. Applied Natural-Languange Processing (NLP), statistical pattern analysis, and composite engagement scoring using MinMaxScaler (scikit-learn) to surface and visualize public sentiment trends. Demonstrates end-to-end analytical thinking from raw text data to interpretable, visualized insight.

Python Pandas NumPy scikit-learn Matplotlib Seaborn Statistics NLP EDA Data Visualization Jupyter Notebook


Steam Charts Tracker (in progress)

End-to-end data project that automatically scapes live gaming metrics from Steam Charts Website across Bronze-Silver-Gold Medallion Architecture for tracking trending games, top games, and peak game records. Being extended with a Power BI analytics dashboard and a game recommendation web-application system applying machine learning on processed pipeline data.

Python Pandas NumPy PostgreSQL Snowflake SQLAlchemy BeautifulSoup ETL Medallion Architecture Machine Learning Data Visualization Power BI Web Scraping Git


📚 Currently Exploring

  • ELK Stack (Elasticsearch · Logstash · Kibana): setting up locally via Docker for log analytics and data observability
  • Machine Learning: applying statistical modeling and scikit-learn to recommendation systems
  • Power BI: building analytics dashboards connected to live data pipelines
  • Docker: containerizing data pipelines and local development environments
  • Ubuntu: Linux-based development workflows and CLI tooling

💬 Ask Me About

Python · SQL · Data Engineering · ETL Pipelines · Dimensional Modeling · Data Analytics · Data Handling · Statistics · Data Visualization · Natural-Language Processing (NLP) · Git · GitHub


⚡ Fun Facts

  • Quiet environments and cold rooms are my ideal work setup
  • I'm an NBA fan, especially when playoff season hits.
  • I love watching League of Legends Esports, especially during World Championship.
  • Working toward a home office with fast internet and a mechanical keyboard worth the investment

📫 Reach Me

Pinned Loading

  1. PH-Flood-Control-Pulse-An-EDA-of-Public-Tweets PH-Flood-Control-Pulse-An-EDA-of-Public-Tweets Public

    Jupyter Notebook

  2. Data-Bank-System Data-Bank-System Public

    Entity Relationship Diagram (ERD) to build analyticsl data repository for the Provincial Government of Bataan that centralizes performance measurements across 37+ departments and/or offices.

  3. steam-charts-tracker steam-charts-tracker Public

    Python