Experience
Professional Experience
F1RST - Santander Group
Data Engineer | Montevideo, Uruguay | Apr 2025 – Present
- Built reusable Python packages for PySpark pipelines, implementing robust unit tests and promoting test-driven development within the team.
- Designed and implemented a Data Contracts framework to strengthen data governance, lineage, and automated quality checks.
- Developed an internal deployment tool for dbt projects using Jinja templates and Databricks Jobs to standardize and automate production deployments.
- Created a metadata-driven framework to configure and generate transformations across multiple tables, supporting column-level and data-type transformations expressed in PySpark.
- Drove the adoption of unit testing and best practices, improving the reliability and maintainability of data pipelines.
- Managed Databricks cluster dependencies via Python wheels for fast, repeatable environment provisioning.
COGNUS LATAM
Data Engineer | Montevideo, Uruguay (Remote) | Aug 2021 – Apr 2025
- Developed and maintained data pipelines in Microsoft Fabric, optimizing data integration in scalable environments.
- Migrated legacy systems to data lakes, following the medallion architecture in Microsoft Fabric.
- Designed and optimized ETL processes with Pentaho Data Integration, improving performance and data quality.
- Implemented CI/CD with GitLab CI/CD or Jenkins to automate deployments and monitoring of data pipelines.
- Used Docker to build reproducible, scalable environments for data solutions.
- Led a team of 4+ people on a strategic project, managing the development and implementation of analytics solutions.
- Created and led the analytics reporting process, removing the dependency on IT for extracting information from transactional systems.
- Implemented a Lakehouse on Databricks, from storage through data modeling.
- Developed data ingestion and transformation processes with Spark and PySpark.
- Trained and deployed machine learning models for trend prediction and business process optimization.
- Delivered internal courses on BI tools and data architectures, promoting the adoption of best practices in information management.
Technologies
- General: Python, R, SQL (PostgreSQL, SQL Server, MySQL), Pentaho, Docker, and Git.
- Cloud: Experience working with Databricks and Microsoft Fabric, using these platforms for scalable data integration, transformation, and analytics.
- Statistical Learning & Machine Learning: Solid knowledge of statistical modeling, supervised and unsupervised learning, with hands-on experience using R (caret, tidymodels) and Python (scikit-learn, TensorFlow, PyTorch) for predictive modeling and data-driven decision making.
- MLOps & Model Deployment: Experience with ML model lifecycle management using MLflow, TensorFlow Serving, and ONNX. Deployed models to cloud environments with Docker, automated workflows with CI/CD pipelines, and monitored models in production.