f
frocode

Frocode

@frocode

Senior Data Enigneer

Algeria
Inglese, Francese
Alcune informazioni sono riportate in lingua inglese.
Chi sono
Hello! FroCode, an experienced Senior Data Engineer and Cloud Specialist certified by IBM and AWS. I specialize in designing, automating, and maintaining robust ETL/ELT architectures that keep your data clean, structured, and instantly ready for analytics. Whether you need a simple automated Python script or an enterprise-grade real-time streaming pipeline... Continua a leggere

Competenze

f
frocode
Frocode
offline • 
Tempo di risposta medio: 1 ora

Consulta i miei servizi

Programmazione e tecnologia
I will build scalable etl data pipelines using python, sql, and apache airflow, dagster

Esperienza lavorativa

RedBull_Futur/IO

Senior Data Engineer (Product Owner)

RedBull Futur/IO • Lavoratore autonomo

Jun 2026 - Present • 4 mos

Managed the end-to-end data platform roadmap and product backlog while collaborating with global stakeholders to translate complex requirements into actionable engineering epics. Engineered a full-scale Medallion architecture on Snowflake using dbt Core and custom macros to manage incremental materializations and maintain strict SCD Type 2 historical logs. Built scalable data ingestion pipelines using PySpark and dlt to extract datasets from high-frequency REST APIs, complex JSON payloads, and enterprise systems like SAP BW via SAP SNP Glue. Optimized Snowflake virtual warehouses and PySpark job configurations by managing partitions, reducing shuffles, and adjusting clustering keys to dramatically slash execution times and compute costs. Developed complex analytical data models, including accumulating snapshot fact tables, to track real-time container routing, capacity planning, and lifecycles across the LATAM region. Enriched internal supply chain tables with external datasets like live weather patterns and port congestion APIs to surface gold-layer data for global interactive visualization dashboards. Orchestrated multi-stage pipelines using Dagster and Apache Airflow while implementing automated data-quality checks and threshold-deviation alerts to notify teams of shipment delays. Established strict data contracts and SLAs between upstream producers and downstream consumers to enforce schema stability and guarantee long-term pipeline reliability. Implemented Liquibase for database schema versioning, allowing the team to automate stateful migrations and execute safe, version-controlled changes directly inside Snowflake. Automated continuous integration and deployment workflows using GitHub Actions to run unit tests, execute SQLFluff linting, and deploy dbt models and pipeline infrastructure safely.

Capgemini

Capgemini

Lavoratore autonomo • 2 yrs 10 mos

Senior Data Engineer

Feb 2025 - Present • 1 yr 8 mos

Authored modular Terraform templates to programmatically provision core AWS services including AWS Glue Data Catalog, EMR clusters, S3 buckets, Athena, and Lake Formation. Designed and published versioned Python transformation packages to standardize internal data-wrangling code patterns and reduce duplication across distributed engineering teams. Administered Kubernetes workloads on EKS, configuring autoscaling rules, persistent volumes, and secure encryption handling via AWS Secrets Manager. Configured granular row and column-level access controls and secure cross-account data shares using AWS Lake Formation governance frameworks. Designed batch extraction workflows to systematically pull raw transactional data into structured analytics layers. Built automated data transformation pipelines and maintained intermediate staging tables to directly fuel operational business intelligence reporting.

QA Test Data Engineer

Jan 2024 - Mar 2025 • 1 yr 2 mos

Developed comprehensive automated quality testing frameworks and test matrices to rigorously evaluate application backend logic, high-volume database structures, and boundary edge cases prior to deployment. Monitored continuous integration testing suites to identify processing regressions early, ensuring clean data payloads and high operational reliability for all downstream software products.