Ariel Guillermo Sánchez Paipilla

Data Architect · DataOps Engineer · Data Engineer · Full Stack

Colombia · Arielsan99@gmail.com · arielsanchez.tech

Measurable Impact

98%
Cost Reduction
Database Restructuring
30-40%
Infra Savings
EKS Modernization
90%+
Security Score
AWS Security Hub
10+
Software Products
Nationally Registered

What I Do

DataOps / DevOps Engineer
  • Built reusable Terraform & Terraform Cloud modules for data teams, enabling self-service provisioning and reducing deployment time across multiple squads.
  • Optimized pipeline execution times and delivery SLAs through ETL refactoring, query tuning, and data consumption pattern improvements.
  • Maintained 90%+ Security Hub compliance; enforced least-privilege IAM policies aligned with zero-trust standards.
  • Led EKS cluster modernization: migrated legacy production-only cluster to current versions, achieving 30-40% cost reduction through right-sizing and eliminating extended support fees.
  • Implemented data governance with AWS LakeFormation and KMS key rotation for encryption compliance.
  • Designed observability for data services (CloudWatch, custom metrics, alerting) and led FinOps strategies for cloud cost optimization.
  • Integrated Generative AI (Bedrock, Sagemaker, OpenAI) into production data workflows for automated processing.
AWSTerraformTerraform CloudEKSLambdaLakeFormationKMSSecurity HubCloudWatchBedrockSagemakerJenkinsArgoDocker
Technical Lead — Data Operations & Architecture
  • Led data engineering team: defined architecture standards, governance policies, and quality frameworks for the analytics platform.
  • Restructured database architecture reducing monthly costs by 98% while improving delivery from daily reports to real-time dashboards every 30 minutes.
  • Architected batch/streaming pipelines with PySpark, Glue, Airflow, Lambda, and Kafka across layered zone models (raw → staging → curated → consumption).
  • Implemented data governance and security for client audits: masking, encryption, lineage tracking, and metadata management.
  • Developed unit, acceptance, and performance tests for data pipelines ensuring integrity across environments.
  • Designed least-privilege IAM roles across AWS and GCP; deployed all infrastructure with Terraform (GitHub versioned).
  • Built custom dashboards in PowerBI & Looker Studio tailored to each client's KPIs. Optimized Snowflake warehouse performance.
AWSGCPTerraformPySparkGlueAirflowKafkaSnowflakePowerBILooker StudioOpenAIPython
Data Engineer — Architecture & Operations
  • Designed and operated data architectures on AWS (Glue, Athena, S3, EMR, Redshift, MWAA) for enterprise BI workloads.
  • Led migration to serverless architecture (EventBridge, Lambda, ECS) achieving zero-downtime transition and reduced operational costs.
  • Implemented data governance with Collibra: enterprise data dictionary and cataloging standards.
  • Applied AWS resource tagging strategy for cost allocation, identifying and optimizing high-consumption database queries.
  • Managed security operations: Security Hub, EKS clusters, VPC design, IAM governance.
AWSTerraformPySparkGlueAthenaRedshiftMWAACollibraEKSKafkaSnowflake
Development Engineer
  • Designed serverless data solutions with Lambda, Glue, EventBridge, QuickSight, and S3 for automated reporting.
  • Built backend/frontend applications for data visualization and BI consumption.
  • Automated report generation workflows with Python, reducing manual effort and improving delivery times.
AWSLambdaGlueEventBridgeQuickSightPython
Research Lead / Data Scientist / Full Stack Developer
  • Led university research group: coordinated projects, mentored team members, defined technical direction for data-driven studies.
  • Published research articles and book chapters in indexed journals; registered 10+ software products nationally (Minciencias).
  • Designed social media data tracking and extraction techniques for large-scale research data collection.
  • Developed cross-platform web & mobile applications; designed ETL pipelines for research data processing.
PythonPandasETLWeb ScrapingFull StackMobile

Core Skills

DataOps & Operations
CI/CD for Data Pipeline Orchestration Observability FinOps Incident Management SLA Governance
Data Architecture
Data Modeling Lake/Warehouse Design Batch & Streaming Data Lineage Zone Models Data Mesh
Cloud & Infrastructure
AWS (20+ services) Terraform Terraform Cloud Docker EKS / ECS GitOps
Security & Governance
LakeFormation Collibra Security Hub IAM Least Privilege KMS Rotation Data Masking

Academic Background

B.S. in Systems and Computer Engineering

Universidad Pedagógica y Tecnológica de Colombia (UPTC)

Technical Degree in Systems

Servicio Nacional de Aprendizaje (SENA)

Spanish: Native
English: Intermediate
Portuguese: Intermediate