Mauro Loprete
  • Inicio
  • Experiencia
  • Docencia
  • Proyectos
  • Certificaciones
  • Blog
  • Slides

Experience

Professional Experience

F1RST - Santander Group

Data Engineer | Montevideo, Uruguay | Apr 2025 – Present

  • Built reusable Python packages for PySpark pipelines, implementing robust unit tests and promoting test-driven development within the team.
  • Designed and implemented a Data Contracts framework to strengthen data governance, lineage, and automated quality checks.
  • Developed an internal deployment tool for dbt projects using Jinja templates and Databricks Jobs to standardize and automate production deployments.
  • Created a metadata-driven framework to configure and generate transformations across multiple tables, supporting column-level and data-type transformations expressed in PySpark.
  • Drove the adoption of unit testing and best practices, improving the reliability and maintainability of data pipelines.
  • Managed Databricks cluster dependencies via Python wheels for fast, repeatable environment provisioning.

COGNUS LATAM

Data Engineer | Montevideo, Uruguay (Remote) | Aug 2021 – Apr 2025

  • Developed and maintained data pipelines in Microsoft Fabric, optimizing data integration in scalable environments.
  • Migrated legacy systems to data lakes, following the medallion architecture in Microsoft Fabric.
  • Designed and optimized ETL processes with Pentaho Data Integration, improving performance and data quality.
  • Implemented CI/CD with GitLab CI/CD or Jenkins to automate deployments and monitoring of data pipelines.
  • Used Docker to build reproducible, scalable environments for data solutions.
  • Led a team of 4+ people on a strategic project, managing the development and implementation of analytics solutions.
  • Created and led the analytics reporting process, removing the dependency on IT for extracting information from transactional systems.
  • Implemented a Lakehouse on Databricks, from storage through data modeling.
  • Developed data ingestion and transformation processes with Spark and PySpark.
  • Trained and deployed machine learning models for trend prediction and business process optimization.
  • Delivered internal courses on BI tools and data architectures, promoting the adoption of best practices in information management.

Technologies

  • General: Python, R, SQL (PostgreSQL, SQL Server, MySQL), Pentaho, Docker, and Git.
  • Cloud: Experience working with Databricks and Microsoft Fabric, using these platforms for scalable data integration, transformation, and analytics.
  • Statistical Learning & Machine Learning: Solid knowledge of statistical modeling, supervised and unsupervised learning, with hands-on experience using R (caret, tidymodels) and Python (scikit-learn, TensorFlow, PyTorch) for predictive modeling and data-driven decision making.
  • MLOps & Model Deployment: Experience with ML model lifecycle management using MLflow, TensorFlow Serving, and ONNX. Deployed models to cloud environments with Docker, automated workflows with CI/CD pipelines, and monitored models in production.

Sitio hecho con Quarto, por Mauro Loprete. Licencia: CC BY-SA 2.0.