Oct 2025 — Present current
Data Engineer / Analytics Engineer
@ EPAM Systems — Tashkent
Data Analytics Engineering and Data Quality delivery track — end-to-end warehouse, pipeline and reporting work on internal client projects.
- Designed a layered data warehouse architecture — staging/ODS, 3NF integration core, and Kimball star-schema data marts — across 11 source tables, choosing the layer boundaries so that every mart figure is traceable back to its source record.
- Built the full DWH load in PL/pgSQL stored procedures, implementing incremental loading with NOT EXISTS deduplication and history handling so re-runs stay idempotent; reduced daily refresh from around 28 to 13 minutes.
- Analysed source systems and produced source-to-target mappings and data-flow documentation covering 14 entities (over 50 attributes), used as the reference for development and validation.
- Built cloud ingestion on AWS: raw CSV/XLSX staged in S3, loaded to Redshift via COPY, with RDS PostgreSQL for operational data and IAM-scoped role-based access.
- Wrote automated data-quality suites in PyTest covering row-count completeness, transformation accuracy and referential integrity; investigated and resolved 250+ discrepancies between source and mart layers.
- Optimised slow-running SQL through indexing, join rewrites and partitioning fundamentals; ran load testing in JMeter via Jenkins CI.
- Delivered 5+ Power BI dashboards on top of the marts, connecting directly to Redshift and writing DAX measures for revenue, margin and payment KPIs.
- Gathered and clarified requirements with stakeholders, translating business questions into modelling and pipeline decisions inside an Agile/SDLC process.
PostgreSQLPL/pgSQLRedshiftS3RDSPythonPyTestPower BIJenkins CI