I'm a Data Engineering, Business Intelligence & Data Science specialist with 16+ years of experience across multinational companies and major economic groups in Peru. I've led data architecture teams and projects on Azure, GCP, and AWS — from pipeline design and data governance to analytics solutions and Power Platform applications.
My focus: turning raw, scattered data into reliable, governed, production-ready pipelines that drive real business decisions.
- 🔭 Currently: Senior Data Engineer at Seidor Analytics, building batch/streaming pipelines on GCP (Dataflow, Composer, Pub/Sub) and BigQuery data models
- 🌎 Based in Lima, Peru — available for remote / freelance work (UTC-5, full overlap with US time zones)
- 🎯 Industries: Banking, Fintech, Insurance, Telecom, Mining, Retail, Consumer Goods, Energy & Fuel, Agribusiness
- Pipeline Engineering — Batch & streaming pipelines (Cloud Dataflow, Cloud Composer, Pub/Sub, Azure Data Factory, AWS Glue)
- Data Modeling — Star schema, Snowflake schema, Data Vault, dimensional modeling for analytical & transactional workloads
- Data Governance & Security — Dataplex, Data Catalog, Unity Catalog, Microsoft Purview, IAM, row/column-level security
- Cloud Data Warehousing — BigQuery, Snowflake, Redshift, Synapse Analytics
- DataOps / CI-CD — Versioned schema deployment across DEV/UAT/PROD, Terraform, Azure DevOps, GitHub Actions
- Power Platform Development — Model-driven apps on Dataverse, Power Apps + Power Automate integrated with Azure Functions, Logic Apps & Azure SQL
- Analytics & ML — Predictive modeling with Amazon SageMaker, Azure ML, BigQuery ML
Selected consulting & staff-augmentation engagements across multi-cloud data platforms (client names withheld under confidentiality):
| Sector | Role | Stack | Highlight |
|---|---|---|---|
| Fintech / Payments (top-10 LatAm processor) | Cloud Data Engineer | GCP · Dataflow · BigQuery · Pub/Sub | ETL/streaming pipelines for a processor handling 4.4B+ transactions/year |
| Banking (leading Peruvian bank) | Lead Data Engineer | Azure/GCP/AWS · Databricks · Synapse · AWS Textract | Multi-cloud data architecture with ML/AI (fraud detection, text recognition) across on-prem & cloud |
| Retail (LatAm conglomerate) | Software Engineer | Backend/Frontend, VTEX integration | Pricing/stock system across 140+ points of sale; 25% faster load times; 99.9% uptime at 500K+ requests/month; 30% logistics efficiency gain |
| Insurance (major Peruvian insurer) | Senior Data Engineer | Snowflake | Scalable pipelines integrating new APIs to feed BI tools org-wide |
| Telecom (multinational, multi-country) | Lead Data Engineer | Azure Data Factory · MongoDB · Python | Cross-country ETL integration with custom data quality flows |
| Mining (large-scale gold producer) | Big Data Engineer | SAP HANA 2.0 · BODS · SAC | Modeled and secured large-scale data systems for financial & order-to-cash processes |
| Consumer Goods (LatAm market leader) | Data Scientist | GCP · Cloud Composer · BigQuery | Regression models & ELT pipelines |
| Beauty / Direct Sales (multinational) | Cloud Data Engineer | Lakehouse (batch & real-time) | Standardized data solutions under a unified Data COE architecture |
| Energy & Fuel Distribution (regional leader) | Senior Data Engineer (Tech Lead) | Multi-cloud | Technical leadership validating vendor solutions & production deployments |
Plus 10+ additional engagements across telecom, banking, agribusiness, construction, and professional services.
| Project | Stack | Description |
|---|---|---|
| gcp-weather-elt-pipeline | Cloud Workflows · Cloud Run · BigQuery · Cloud Storage | Workflow-orchestrated ELT with raw → std → trf layers, private OIDC-invoked service, least-privilege IAM, idempotent MERGE loads and 20 tests, all on the GCP free tier |
| gcp-customer-segmentation-pipeline | Cloud Workflows · Pub/Sub · Cloud Functions · BigQuery · Firestore | Event-driven customer segmentation: 3 Pub/Sub topics, 4 Cloud Functions, rule-based routing with idempotent MERGE, workflow that waits for the async leg, least-privilege IAM and 56 tests, all on the GCP free tier |
| gcp-mdm-country-sync | Cloud Workflows · Cloud Run · BigQuery · Firestore · Pub/Sub | Master data management: change detection in BigQuery, a single idempotent CRUD API as the only write path to Firestore, audit trail and one Pub/Sub event per real change, least-privilege IAM and 118 tests, all on the GCP free tier |
- B.Sc. in Economic Engineering — Universidad Nacional de Ingeniería (UNI)
- Data Mining Specialization — CEPS-UNI
- Financial Management Diploma — ASBANC
- Data Science Specialization — Escuela Data Science, Platzi
Open to freelance / contract data engineering work — pipelines, cloud migrations, data governance, and analytics platforms.
📧 alvaroyalle@yahoo.es · 💼 LinkedIn · 🌐 Upwork