Data & Analytics Engineering

I build data pipelines — ingestion, orchestration, delivery.

Python · SQL · scheduled workflows

Statistician (UNICAMP). Three years building the ingestion and orchestration layer — scrapers, RPA bots, API integrations and scheduled workflows that collect, validate and prepare data for analysis.

PythonSQLPlaywrightn8nDockerPower BIAzure
PROFILE
Rodrigo Carvalho
Statistician · UNICAMP
Campinas, Brazil · UTC−3

Stack, by pipeline layer

The tools I work with at each stage, from extracting data at the source through to delivering it somewhere it can be queried.

ingestion
PythonrequestsSeleniumPlaywrightAutomation 360REST APIsSAP extracts

Extraction from sources with no usable export path: headless browsers, RPA bots and third-party APIs.

storage
SQL ServerPostgreSQLMySQLOracleSQLite

Primarily relational. Schema design and data modeling, plus query optimization for performance on large tables.

transformation
pandasnumpySQL: window fns · CTEsPower QueryDAXdimensional modeling

Star schemas, incremental load logic, and the statistical layer on top: regression, ANOVA and time series.

orchestration
n8nMakeNode-REDGitHub ActionsDockerAzure

Scheduling, retry handling and the integration between steps, so workflows run unattended.

serving
Power BIsemantic modelsDAXFastAPICSV / JSON exports

Delivery to the consumer, whether that is a dashboard, an API, a scheduled export or an inbox.

Selected work

code and data on GitHub

Experience

jul 2024 — now
Marketing Data Analyst · Pefaz

Built the lead-capture pipeline end to end: n8n and Make workflows that collect, validate and enrich prospect data from multiple channels, then create and assign CRM records in real time through the HubSpot and RD Station REST APIs. Predictive lead scoring and A/B analysis (regression, ANOVA) run on top of it; Power BI reports conversion, pipeline velocity, cost per lead and campaign ROI off the same data.

sep 2023 — jul 2024
Data Analyst, Pricing · GrupoSC

Automated the price-collection layer with Python and Selenium, cutting collection time by 80%, and merged SAP transactional extracts into the elasticity model senior leadership priced from. Delivered the real-time price tracking and competitive intelligence dashboards on top.

may 2022 — apr 2023
BI Data Analyst · Catho

Stood up the market-intelligence scrapers feeding weekly competitive benchmarks, and rewrote the SQL underneath the job-marketplace KPIs for performance and scalability on large datasets.

jul 2021 — jul 2022
Associate Software Engineer · Accenture Brazil

Automated critical business processes in Automation Anywhere (Automation 360) and built the pipelines carrying RPA output into SQL databases. Deployed PostgreSQL and Node-RED in Docker containers on Azure, improving pipeline orchestration and reliability.

oct 2020 — jun 2021
Data Science Intern · PlanD Data Intelligence

Time-series forecasting for sales planning and inventory optimization, and data mining for consumption patterns — on the relational databases I built and maintained in SQL.

Contact

Get in touch.

Based in Campinas, São Paulo, Brazil.

Education & certifications
degree

B.Sc. Statistics

University of Campinas (UNICAMP)

certs
Azure Fundamentals · AZ-900Automation Anywhere · Advanced RPA