SLIIT · Year 4 Data Science · Malabe, Sri Lanka
Data Engineer in the making.
Building production-grade data pipelines, streaming platforms, and lakehouses. I write Python that moves data, SQL that finds truth in it, and systems that keep it honest.
I'm a final-year Data Science undergraduate at SLIIT, Malabe, with hands-on experience spanning the full data engineering stack from ingestion and transformation to streaming pipelines, cloud deployments, and BI/OLAP reporting.
My background covers Python, Spring Boot, the MERN stack, R, PostgreSQL, MongoDB, Docker, SSIS/SSAS, AWS, and Apache Airflow, plus applied statistics and machine learning (scikit-learn, XGBoost, hypothesis testing). I bridge the gap between data science theory and production engineering practice.
I'm currently building out my data engineering stack - AWS, Airflow, dbt, and Redshift - through hands-on projects, while targeting internship roles at Sri Lanka's leading IT companies and competitive remote positions abroad.
Sri Lanka Institute of Information Technology · SLIIT · Malabe
Cumulative GPA: 3.79 / 4.00. Specialising in data engineering, machine learning, and applied statistics. Coursework spans database design, Python, R, statistical modelling, cloud computing, and distributed systems. Consistently recognised on the Dean's List for academic excellence.
Dehiattakandiya National School · Biological Science Stream
Physics: B · Chemistry: B · Biology: S — a foundation in mathematics, physics, and analytical thinking that underpins my engineering approach.
AWS Skill Builder
Data Engineering on AWS - Foundations
Completed
HackerRank
SQL (Basic) Certificate
Completed
HackerRank
SQL (Intermediate) Certificate
Completed
HackerRank
SQL (Advanced) Certificate
Completed
Harvard · CS50
CS50's Introduction to Programming with Python
Completed
Harvard · CS50
CS50's Introduction to Databases with SQL
Completed
SLIIT
Dean's List Recognition - Five Consecutive Semesters
Awarded
Languages
Data Engineering
Databases & Cloud
End-to-end AWS warehouse processing 9 source tables (1.5M+ rows) through Bronze/Silver/Gold medallion architecture. Kimball star schema in dbt (4 facts, 4 dims), two-layer data quality strategy (Great Expectations + 25 dbt tests), least-privilege IAM across three identities, and Dockerized Airflow (CeleryExecutor) with a data-quality gate.
Star-schema warehouse (1 fact, 5 dims) from a 180,000+ row multi-source dataset. 3-package SSIS pipeline with SCD Type 2, SSAS OLAP cube with 4 hierarchies and 2 KPIs, and 4 interactive Power BI reports with drill-through navigation.
Rebuilt a university project from scratch: engineered 38 features from an 85-column, 46,321-row NZ Airbnb dataset, dropped target-leakage columns, and bundled preprocessing + XGBoost into one scikit-learn Pipeline. R² = 0.773 (MAE NZD 70) - a 35% relative improvement over the original R² = 0.57. Deployed live on Streamlit Community Cloud.
Hypothesis-driven statistical analysis of 32,593 students (OULAD dataset). Led data preparation and feature engineering, identifying MNAR missingness in assessment scores. Findings: high-engagement students were 8× more likely to succeed (Cohen's h = 1.18); ANOVA showed engagement explained 52% of outcome variance; logistic regression reached 75-78% accuracy, AUC > 0.80.
Full-stack university operations platform (5 modules, 5-person team). Built the maintenance/incident ticketing module: submission with Cloudinary image attachments, full OPEN → IN_PROGRESS → RESOLVED → CLOSED workflow, technician assignment, comment threads, and Lucene-powered similar-ticket detection.
Led a 5-person team building a full-stack vehicle service booking platform (shared backend, user/admin/employee frontends). Owned booking management: smart time-slot logic, holiday/late-booking handling, lifecycle tracking, PDF report exports, and email/conflict-prevention logic.
Looking for roles in Sri Lanka and competitive remote positions abroad. Available from mid-2026.