Data Engineer @ Leonardo
Italy · Remote/Hybrid
Lead and develop enterprise data pipelines and analytics infrastructure, owning high-level design from data modelling through to production deployment.
- Lead and develop ETL/ELT pipelines in Python/PySpark on Azure Data Factory, Databricks and Microsoft Fabric, owning high-level design from data modelling to deployment.
- Migrated large datasets to Google Cloud BigQuery, cutting infrastructure costs and improving query performance by ~40%.
- Automated data validation and reconciliation, reducing manual QA effort and lifting data accuracy to near 100%.
- Built Power BI and Apache Superset dashboards surfacing KPIs to leadership, enabling faster decision-making.
- Act as technical lead for the team while developing core components end-to-end.