Resume
Ariel Guillermo Sánchez Paipilla
Data Architect · DataOps Engineer · Data Engineer · Full Stack
Colombia · Arielsan99@gmail.com · arielsanchez.tech
Key Achievements
Measurable Impact
98%
Cost Reduction
Database Restructuring
Database Restructuring
30-40%
Infra Savings
EKS Modernization
EKS Modernization
90%+
Security Score
AWS Security Hub
AWS Security Hub
10+
Software Products
Nationally Registered
Nationally Registered
Experience
What I Do
DataOps / DevOps Engineer
- Built reusable Terraform & Terraform Cloud modules for data teams, enabling self-service provisioning and reducing deployment time across multiple squads.
- Optimized pipeline execution times and delivery SLAs through ETL refactoring, query tuning, and data consumption pattern improvements.
- Maintained 90%+ Security Hub compliance; enforced least-privilege IAM policies aligned with zero-trust standards.
- Led EKS cluster modernization: migrated legacy production-only cluster to current versions, achieving 30-40% cost reduction through right-sizing and eliminating extended support fees.
- Implemented data governance with AWS LakeFormation and KMS key rotation for encryption compliance.
- Designed observability for data services (CloudWatch, custom metrics, alerting) and led FinOps strategies for cloud cost optimization.
- Integrated Generative AI (Bedrock, Sagemaker, OpenAI) into production data workflows for automated processing.
AWSTerraformTerraform CloudEKSLambdaLakeFormationKMSSecurity HubCloudWatchBedrockSagemakerJenkinsArgoDocker
Technical Lead — Data Operations & Architecture
- Led data engineering team: defined architecture standards, governance policies, and quality frameworks for the analytics platform.
- Restructured database architecture reducing monthly costs by 98% while improving delivery from daily reports to real-time dashboards every 30 minutes.
- Architected batch/streaming pipelines with PySpark, Glue, Airflow, Lambda, and Kafka across layered zone models (raw → staging → curated → consumption).
- Implemented data governance and security for client audits: masking, encryption, lineage tracking, and metadata management.
- Developed unit, acceptance, and performance tests for data pipelines ensuring integrity across environments.
- Designed least-privilege IAM roles across AWS and GCP; deployed all infrastructure with Terraform (GitHub versioned).
- Built custom dashboards in PowerBI & Looker Studio tailored to each client's KPIs. Optimized Snowflake warehouse performance.
AWSGCPTerraformPySparkGlueAirflowKafkaSnowflakePowerBILooker StudioOpenAIPython
Data Engineer — Architecture & Operations
- Designed and operated data architectures on AWS (Glue, Athena, S3, EMR, Redshift, MWAA) for enterprise BI workloads.
- Led migration to serverless architecture (EventBridge, Lambda, ECS) achieving zero-downtime transition and reduced operational costs.
- Implemented data governance with Collibra: enterprise data dictionary and cataloging standards.
- Applied AWS resource tagging strategy for cost allocation, identifying and optimizing high-consumption database queries.
- Managed security operations: Security Hub, EKS clusters, VPC design, IAM governance.
AWSTerraformPySparkGlueAthenaRedshiftMWAACollibraEKSKafkaSnowflake
Development Engineer
- Designed serverless data solutions with Lambda, Glue, EventBridge, QuickSight, and S3 for automated reporting.
- Built backend/frontend applications for data visualization and BI consumption.
- Automated report generation workflows with Python, reducing manual effort and improving delivery times.
AWSLambdaGlueEventBridgeQuickSightPython
Research Lead / Data Scientist / Full Stack Developer
- Led university research group: coordinated projects, mentored team members, defined technical direction for data-driven studies.
- Published research articles and book chapters in indexed journals; registered 10+ software products nationally (Minciencias).
- Designed social media data tracking and extraction techniques for large-scale research data collection.
- Developed cross-platform web & mobile applications; designed ETL pipelines for research data processing.
PythonPandasETLWeb ScrapingFull StackMobile
Expertise
Core Skills
DataOps & Operations
CI/CD for Data Pipeline Orchestration Observability FinOps Incident Management SLA GovernanceData Architecture
Data Modeling Lake/Warehouse Design Batch & Streaming Data Lineage Zone Models Data MeshCloud & Infrastructure
AWS (20+ services) Terraform Terraform Cloud Docker EKS / ECS GitOpsSecurity & Governance
LakeFormation Collibra Security Hub IAM Least Privilege KMS Rotation Data MaskingEducation
Academic Background
B.S. in Systems and Computer Engineering
Universidad Pedagógica y Tecnológica de Colombia (UPTC)
Technical Degree in Systems
Servicio Nacional de Aprendizaje (SENA)
Spanish: Native
English: Intermediate
Portuguese: Intermediate