Paul M.-Data Engineer
Check rate
Experience
Data Engineer
Luxoft
- Built and deployed an end-to-end enterprise data integration platform using CloverDX ETL pipelines, Python, PostgreSQL, and AWS services to ingest, validate, and structure raw analytical datasets supporting AI-powered automation for financial operations and digital banking workflows.
- Designed data extraction connectors collecting fragmented structured and semi-structured input sources and aligning them to unified schema definitions required for downstream analytics workflows.
- Built automated data loading and distribution jobs targeting multi-region storage across S3, RDS, and Redshift, ensuring secure data availability for risk analysis, fraud detection models, credit scoring, and scalable financial reporting.
- Worked closely with business product leaders to evaluate new integration paths and prototype rapid connectors for high-priority data partners.
- Provided on-call operational support to troubleshoot ingestion failures, data latency bottlenecks, and corrupted financial files, performing root-cause analysis through transaction-level replay and controlled environment replication.
- Maintained automated data quality profiling, validation rulesets, and error-handling flows, ensuring consistency and reducing manual reconciliation across systems.
- Created internal technical documentation including data lineage, field definitions, reconciliation rules, financial lifecycle diagrams, and mapping specifications used across engineering, compliance, and support teams.
Data Engineer
Unicage
- Built cloud-based ETL ingestion framework using Airflow, Python, Aurora PostgreSQL, and AWS Lambda to integrate multiple partner data providers into financial-grade web applications.
- Developed custom SQL transformation scripts with field-level validation logic to handle malformed input and edge-case behavior from third-party interfaces.
- Integrated data warehousing concepts including dimensional modeling and incremental loading patterns to support scalable insights tooling.
- Collaborated with security teams to align data access flows with regulatory controls and auditing documentation.
- Introduced automated regression data tests, enabling detection of mapping drift before deployment to production systems.
Data Engineer
Biobot Analytics
- Built large-scale COVID-19 public health data processing pipelines using Databricks, Apache Spark, Snowflake, and AWS to ingest real-time case reporting feeds from hospitals, diagnostic labs, and national open-data programs supporting public health intelligence platforms.
- Integrated disparate raw datasets including vaccination progress tracking, ICU bed utilization, mortality curves, and population density metrics into curated warehouse models designed for advanced epidemiological and operational analysis.
- Designed automated data validation rules and quality scoring frameworks utilizing anomaly detection logic and threshold-based alerting tied to pipeline health metrics.
- Built operational observability dashboards in Grafana and Cloud Monitoring, visualizing pipeline latency, throughput, and schema-change impact to assist proactive issue detection.
- Provided rapid response support during emergency reporting intervals, verifying the correctness of published datasets prior to high-visibility distribution.
Data Developer Intern
Amazon
- Modernized legacy ETL workflows by migrating to modular, service-based pipelines, reducing operational maintenance and improving reliability across data systems.
- Built automated ingestion frameworks for partner data feeds with cleansing and normalization, reducing processing time and improving data accuracy.
- Partnered with Security & Compliance teams to integrate regulated access controls and audit mechanisms, ensuring alignment with enterprise governance and regulatory standards.
Industry experience
See where this freelancer has spent most of their professional time.
Experienced in Information Technology, Banking and Finance, and Healthcare.
Business area experience
See which departments and functions this freelancer has contributed to most.
Experienced in Business Intelligence, Information Technology, Quality Assurance, and Research and Development.
Summary
Cloud-focused Senior Data Engineer with over eight years of hands-on experience designing and delivering high-reliability data processing systems, enterprise ETL pipelines, and distributed integration platforms across financial and AI-driven environments. Deep background integrating complex data sources, optimizing large-scale pipelines, and ensuring data integrity for mission-critical applications. Strong collaboration with cross-functional teams including analysts, architects, and business stakeholders within fast-paced environments.
Skills
- Cloud & Infra: Aws (Lamda, S3, Ec2, Rds, Cloudwatch, Emr), Azure, Docker, Kubernetes
- Etl & Data Pipelines: Cloverdx, Apache Nifi, Airflow, Informatica, Talend, Ssis, Glue, Kafka
- Database & Warehousing: Postgresql, Mysql, Oracle, Redshift, Bigquery, Snowflake
- Language & Scripting: Python, Typescript/Javascript, Java, Sql, Bash
- Architecture: Data Modeling, Batch & Streaming Pipelines, Microservices, Ddd, Event-Driven Integration
- Tools: Git, Ci/Cd, Terraform, Jira, Confluence
Languages
Education
The University of Tokyo
B.S., Computer Science · Computer Science · Japan
Statistics
Experience
Expertise
Qualifications
Profile
Frequently asked questions
Have questions? Find more information here.
Paul is based in Warsaw, Poland.
Paul speaks the following languages: English (Advanced), Japanese (Intermediate).
Paul has at least 8 years of experience. During this time, Paul has worked in at least 2 different roles and for 4 different companies. The average length of individual experience is 2 years and 11 months. Note that Paul may not have shared all experience and actually has more experience.
Based on recent experience, Paul would be well-suited for roles such as: Data Engineer, Data Developer Intern.
Paul's most recent position is Data Engineer at Luxoft.
In recent years, Paul has worked for Luxoft, Unicage, and Biobot Analytics.
Paul is most experienced in industries like Information Technology, Banking and Finance, and Healthcare.
Paul is most experienced in business areas like Business Intelligence, Information Technology, and Quality Assurance. Paul also has some experience in Research and Development.
Paul has recently worked in industries like Banking and Finance, Information Technology, and Healthcare.
Paul has recently worked in business areas like Business Intelligence, Information Technology, and Quality Assurance.
Paul holds a Bachelor in Computer Science from The University of Tokyo.
The availability of Paul needs to be confirmed.
Daily rate distribution
The rates shown represent the typical market range for freelancers in this position based on recent contracts on our platform.
Average rates for similar positions
Rates are based on recent contracts and do not include FRATCH margin.
The average daily rate is the mean of all daily rates from recent contracts of comparable freelancers on our platform.
The median daily rate is the middle value of all daily rates — half of comparable freelancers charge less, half charge more. Unlike the average, it is barely affected by outliers.
Calculated based on our freelancers’ daily rates as of 7 Sep 2026. Actual rates may vary depending on seniority level, experience, skill specialization, project complexity, and engagement length.
Similar freelancers
Discover other experts with similar qualifications and experience
Experts recently working on similar projects
Freelancers with hands-on experience in comparable project as a Data Engineer
Nearby freelancers
Professionals working in or nearby Warsaw, Poland
