Simera Professional Key (SPK)

Roger M

Peru

Data Engineer

$ 4,200/month

5 yrs exp

Data Engineer with 4+ years of experience designing and operating scalable data platforms on AWS, GCP, and Azure. Deep expertise in end-to-end ETL/ELT pipelines, data lake architectures (S3/Iceberg), and analytics-ready datasets using Python, SQL, PySpark, and cloud-native services. Proven track record of cutting costs, reducing incidents by up to 90%, and improving pipeline performance in production environments. Hands-on experience with ML model maintenance, data quality frameworks, CI/CD …

Skills

  • Azure
  • CI/CD
  • data modeling
  • Docker
  • Python
  • Redis
  • Spark
  • SQL
  • Data Engineer
  • Data Pipeline
  • DataBricks
  • Pandas
  • Phyton
  • Power BI
  • Pyspark
  • Python/R
  • CICD
  • Data engineering
  • PowerBI
  • Python3
  • Looker Studio
  • RAG
  • LookerStudio
  • GitHub Actions
  • ETL/ELT pipelines
  • MSQL
  • Looker Data Studio
  • Apache Spark
  • ETL/ELT Pipeline
  • AWS S3
  • Data Quality & Validation

Roger M

Peru

Data Engineer

$ 4,200 /month

Part Time: $ 2,450/month

5 yrs exp

Data Engineer with 4+ years of experience designing and operating scalable data platforms on AWS, GCP, and Azure. Deep expertise in end-to-end ETL/ELT pipelines, data lake architectures (S3/Iceberg), and analytics-ready datasets using Python, SQL, PySpark, and cloud-native services. Proven track record of cutting costs, reducing incidents by up to 90%, and improving pipeline performance in production environments. Hands-on experience with ML model maintenance, data quality frameworks, CI/CD …

Skills

  • Azure
  • CI/CD
  • data modeling
  • Docker
  • Python
  • Redis
  • Spark
  • SQL
  • Data Engineer
  • Data Pipeline
  • DataBricks
  • Pandas
  • Phyton
  • Power BI
  • Pyspark
  • Python/R
  • CICD
  • Data engineering
  • PowerBI
  • Python3
  • Looker Studio
  • RAG
  • LookerStudio
  • GitHub Actions
  • ETL/ELT pipelines
  • MSQL
  • Looker Data Studio
  • Apache Spark
  • ETL/ELT Pipeline
  • AWS S3
  • Data Quality & Validation

Data Engineer

JobLeap
June 2025 - present

Architected and deployed end-to-end data pipelines on AWS (S3, Glue, Athena, Iceberg) with CI/CD via GitHub Actions, reducing time-to-deploy and standardizing deployment practices. Built and maintained billing data pipelines and reporting systems enabling accurate real-time revenue tracking for business stakeholders. Designed a centralized API for business entities using domain-driven data modeling, contributing to a scalable control plane architecture. Significantly reduced Athena query costs by redesigning data models and precomputing datasets, eliminating redundant scans. Improved URL redirection service scalability via Redis-based queueing, increasing throughput and reducing latency under high load. Introduced distributed processing with Apache Spark, enabling the team to handle growing data volumes without pipeline degradation. Contributed to agentic development practices, automating repetitive engineering workflows and reducing manual intervention.

Data Engineer

Canvia
October 2024 - June 2025

Maintained and optimized a production fraud detection system using Docker and system-level tuning, ensuring high availability and reliability. Optimized Kafka and Debezium CDC pipelines, reducing data traffic by 80% through schema filtering and partition tuning. Enhanced PostgreSQL and Redis performance via query optimization and caching strategies, improving overall system stability. Led migration of legacy ETL pipelines to AWS (S3, Glue, Redshift), modernizing data infrastructure and improving scalability. Operated and supported ML systems in production: model monitoring, performance maintenance, and incident response.

Data Engineer

Periferia IT Group
October 2023 - October 2024

Implemented monitoring and data validation systems that reduced production incidents by 90%, improving pipeline reliability across clients. Contributed to Data Warehouse Data Lake migration, improving data accessibility and reducing storage costs. Automated financial reporting across multiple countries, eliminating manual processes and ensuring consistent, timely delivery. Established CI/CD pipelines and data quality frameworks, standardizing deployment and testing practices across the team.

Data Engineer

Predictiva S.A.C
November 2021 - September 2023

Supported migration of data infrastructure to GCP (BigQuery), ensuring data integrity, quality validation, and zero data loss. Developed and maintained ETL workflows using Apache Airflow, automating data ingestion from multiple source systems. Built reporting and analytics automation solutions, reducing manual effort and enabling faster business insights.

Smart Scores

Communication
80
Role Fit
90
Adaptability
90
Problem-solving
85
Professional Presence
80
Drive/Initiative
85

Smart Skills

beta