NapoliData · US Contractor
Data Engineer & AI Engineer. Contracting through Napoli Data LLC: distributed data pipelines and cloud data platforms across AWS and Azure, ETL/ELT orchestration with Glue, Step Functions, Airflow and Databricks, and LLM-based assistants for pipeline observability with LangChain, Anthropic Claude, Amazon Bedrock and Ollama for local inference.
AWSAzureDatabricksSnowflakeAirflowKafkaClaude · BedrockOllama
CLARA Analytics · US Contractor
Senior Data Engineer. Designed and operated a multi-tenant serverless data standardization platform on AWS. Carrier files land into raw, cleansed and standardized layers on Apache Iceberg, orchestrated by Step Functions, Glue and Lambda. Onboarded new carriers end to end, from file contract to validated production load. Incremental loads with SCD Type 2 and Iceberg MERGE, automated post-load validation, and production incident response. Wrote a local LLM analyzer (LangChain, Ollama, DeepSeek) that ranked root-cause hypotheses for on-call.
Proactiviti · US Contractor
Data Engineer. Distributed pipelines on AWS (Glue, Lambda, Step Functions, Athena) and Airflow on Kubernetes with KubernetesPodOperator. Real-time ingestion from REST APIs and Kafka; automated ETL from third-party APIs into Azure SQL; hybrid flows across Azure Data Factory, Synapse and Databricks. Tuned Redshift queries and PostgreSQL RDS for scalability and cost. Automated monitoring with LLM agents built on LangChain and Claude.
−20%unplanned incidents · LLM monitoring agents
Fivvy · US Contractor
AWS Data Engineer. Data pipelines on Lambda, Glue and RDS Aurora, with Airflow, Step Functions and EC2 for large volumes; Pandas and PySpark for analysis at scale; S3 and Athena for storage and querying; CloudWatch for monitoring. Re-engineered an ETL process running every 15 minutes on EC2, RDS and S3.
−46%utilization cost · faster data refreshes
Aprende Institute · US Contractor
AWS BI Data Engineer. Designed, implemented and maintained the cloud data architecture on AWS: EC2, S3, Lambda to automate ETL and maintenance tasks, Glue crawlers to catalog data in S3, and Redshift and Athena for analytics, with query tuning and schema-design practices for speed and accuracy.
Johnson & Johnson · US Contractor
Senior Data Engineer. Large-scale ETL pipelines on AWS (Python, Lambda, Glue, Athena, S3) turning raw MySQL and S3 data into features for data scientists and reports for business analysts. Put machine-learning solutions into production (environment setup, Bitbucket version control, unit testing in Flask/Python) and improved monitoring with CloudWatch and EventBridge.
Prisma Medios de Pago
Data Scientist Project Leader. Big Data & Analytics at a multi-brand payment processor: led the data-science team's projects, planned enterprise-wide data solutions, built and monitored predictive models, and defined business metrics and delivery timelines.
Banco Galicia
Data Analyst (Credit Risk) → Senior Data Scientist (Marketing & BI). Propensity models for insurance and loans, a cross-sell recommendation engine, RFM and unsupervised customer segmentation, a real-time recommender on Oracle and an NLP ReMarketing chatbot. In Credit Risk: a lending qualification engine and automated qualification of the credit-card portfolio. Measurable sales lift, material call-center savings.
Case study overview →
Banco Patagonia
Data Analyst. Credit Risk Management (2013-2016): credit scoring and evaluation models for individuals and companies. Built, updated and optimized scores and mass ratings. Commercial Operations Support (2009-2013): statistical studies on credit-card claims and report automation.