NEW
Proxify is bringing transparency to tech team performance based on research conducted at Stanford. An industry first, built for engineering leaders.
Learn more
Caio M.
Datatekniker
Caio er en alsidig dataekspert med over fem års erfaring inden for software- og datateknik, datavidenskab og analyse.
Derudover havde Caio en central rolle hos Nubank, hvor han ledte et team på 200 medarbejdere og styrede initiativet Data Governance and Cost Accountability. Dette initiativ forbedrede gennemsigtigheden af projekt- og opgaverelaterede omkostninger betydeligt til gavn for alle involverede teammedlemmer.
Under sin ansættelse hos Accenture udviste Caio arbejdede han flittigt med at gennemføre GCP Data-løsninger i kundernes infrastruktur. Denne indsats havde stor betydning, da den strømlinede dataindsamlingsprocessen fra forskellige kilder og reducerede den tid, der tidligere blev brugt på manuel dataindsamling.
Hovedekspertise
- Python 6 år

- SQL 6 år

- ETL 5 år

Andre færdigheder
- OAuth2 4 år

- GraphQL 4 år

- PowerShell 3 år

Udvalgt oplevelse
Beskæftigelse
Data Engineer
Proxify AB - 5 måneder
-
Working In Multiple Clients, providing Data Engineering & Analytics consultancy and development
-
Automating Web Page Navigation and Scraping using Playwright and BeautifulSoup
-
Providing Data Analytics and Data Modelling solutions with multiple frameworks
-
Implementing ETL jobs to integrate data from different sources into Data Warehouses or Data Lakes
Teknologier:
- Teknologier:
HTML
Azure Blob storage
TensorFlow
NumPy
OpenCV
XGBoost
Keras
Pandas
Open source
LaTeX
PyCharm
BigQuery
- CSV
OAuth2
- Command-line interface
Unix
VSCode
SciPy
Scikit-learn
- ELT
Matplotlib
- Data Analytics
Azure Synapse
- Recurrent neural network
PL/SQL
XML
- NLP
Machine Learning
BeautifulSoup
SQLAlchemy
Tableau
Plotly
- Dimensional modeling
- Fact Data Modeling
Redshift
dbt
- Prompt Engineering
LangChain
Looker
-
Datatekniker
Nubank - 3 flere år 5 måneder
- Reformerede dataindsamlingsarkitekturen for sociale medier og reducerede beregningstid og omkostninger med over 90 procent.
- Stod i spidsen for et initiativ til datastyring og omkostningsansvarlighed i et team på 200 medarbejdere, hvilket forbedrede omkostningssynligheden i forbindelse med projekter og opgaver.
Analytics Engineer
Nubank - 3 flere år 5 måneder
-
Spearheaded the reformulation of social media data collection architecture, achieving a remarkable reduction of computing time and costs by over 90%;
-
Integrated a new pipeline with the company’s data lake, enabling universal access to Social Media datasets and fostering collaboration;
-
Led a Data Governance and Cost Accountability initiative within a team of 200 members, enhancing transparency and providing visibility into costs associated with projects and tasks;
-
Delivering meaningful data and insights empowered the team to concentrate on content analytics and performance, facilitating informed decision-making and optimizing workflow efficiency.
Teknologier:
- Teknologier:
HTML
Scala
Azure Blob storage
- Data Science
Azure Data Factory
TensorFlow
NumPy
OpenCV
XGBoost
Keras
Pandas
ClojureScript
R (programming language)
Open source
LaTeX
PyTorch
PyCharm
BigQuery
- CSV
OAuth2
- Command-line interface
Unix
VSCode
SciPy
Scikit-learn
- ELT
Matplotlib
- Data Analytics
Azure Synapse
Random Forest
- PCA
Convolutional neural network
- Recurrent neural network
PL/SQL
XML
- NLP
Machine Learning
Cuda
BeautifulSoup
SQLAlchemy
Tableau
Clojure
Plotly
- Dimensional modeling
- Fact Data Modeling
Redshift
dbt
- Prompt Engineering
LangChain
Looker
Dataflow
-
Datatekniker
ClearSale - 1 år 2 måneder
- Hjalp med at opskalere et nyt biometriprodukt ved at definere præstationsmålinger og overvåge programmets adfærd.
- Identificerede huller i dataindsamlingen, så teamet blev opmærksom på problemer i realtid.
- Kundeadoptionen tidoblede, og de månedlige indgående forespørgsler steg fra et par tusinde til millioner.
Product Intelligence
ClearSale - 1 år 2 måneder
-
Played a key role in scaling up a new Biometry product by defining metrics for performance evaluation and monitoring the application’s behavior;
-
Identified gaps in data collection processes, enabling the team to address issues in near real-time and enhance overall data quality;
-
Collaborated with Product and Sales teams to reformulate the sales pitch, emphasizing improvements driven by data insights;
-
Successfully contributed to a tenfold increase in client adoption and a significant surge in monthly incoming requests, from a few thousands to millions, showcasing the product's enhanced value proposition.
Teknologier:
- Teknologier:
HTML
Oracle
Scala
Azure Blob storage
- Data Science
Azure Data Factory
TensorFlow
NumPy
OpenCV
XGBoost
Keras
Pandas
R (programming language)
Open source
LaTeX
PyCharm
BigQuery
- CSV
OAuth2
- Command-line interface
Unix
VSCode
SciPy
Scikit-learn
- ELT
Matplotlib
- Data Analytics
Azure Synapse
Random Forest
- PCA
Convolutional neural network
- Recurrent neural network
PL/SQL
XML
- NLP
Machine Learning
- Computer Vision
Cuda
BeautifulSoup
SQLAlchemy
Tableau
Plotly
- Dimensional modeling
- Fact Data Modeling
Redshift
Looker
Dataflow
-
Datatekniker
Accenture Brazil - 5 måneder
- Arbejdede med implementering af GCP Data-løsninger på kundernes infrastruktur;
- Udviklede en løsning til dataindsamling fra mere end ti datakilder for at skabe et datalager, hvilket sparer tid på manuel indsamling af data fra dem;
- Hjalp med at reducere omkostningerne fra tredjepartskilder med 50 procent ved at bruge de rå data til at udvikle vores egen Data Viz og annullere overflødige Analytics-kontrakter.
Data & AI
Accenture Brazil - 5 måneder
-
Contributed to the implementation of GCP Data solutions on clients’ infrastructure, enhancing data management capabilities and optimizing workflow efficiency;
-
Designed and implemented a solution to aggregate data from over 10 sources to establish a centralized Data Warehouse, significantly reducing manual data collection efforts and streamlining data processing workflows;
-
Played a key role in cost reduction initiatives by leveraging raw data to develop in-house Data Visualization tools, resulting in a 50% reduction in costs associated with third-party sources and the cancellation of redundant Analytics contracts.
Teknologier:
- Teknologier:
Oracle
Azure Blob storage
- Data Science
Azure Data Factory
TensorFlow
NumPy
OpenCV
Keras
Pandas
Open source
LaTeX
PyTorch
PyCharm
BigQuery
- CSV
OAuth2
- Command-line interface
Unix
VSCode
SciPy
Scikit-learn
- ELT
Matplotlib
- Data Analytics
Azure Synapse
Random Forest
- PCA
Convolutional neural network
PL/SQL
- NLP
Machine Learning
- Computer Vision
Cuda
BeautifulSoup
SQLAlchemy
Tableau
Plotly
- Dimensional modeling
- Fact Data Modeling
Redshift
Talend
Looker
Dataflow
-
Data Scientist
Netshoes Brazil - 11 måneder
-
Collaborated closely with the Marketing department to optimize the targeting of advertisements and mail campaigns to customers, enhancing their effectiveness;
-
Conceptualized and implemented a source of truth dataset for Customers’ data, leading to an increase in the frequency of model training and improving overall data quality;
-
Leveraged more up-to-date analysis to drive a daily increase of R$50k in gross income by refining the targeting of mailings and advertisements, thereby maximizing revenue generation efforts.
Teknologier:
- Teknologier:
Oracle
Azure Blob storage
- Data Science
Azure Data Factory
TensorFlow
NumPy
OpenCV
Keras
Pandas
R (programming language)
Open source
LaTeX
PyTorch
PyCharm
BigQuery
- CSV
OAuth2
- Command-line interface
Unix
SciPy
Scikit-learn
- ELT
Matplotlib
- Data Analytics
Azure Synapse
Random Forest
- PCA
Convolutional neural network
- Recurrent neural network
PL/SQL
Machine Learning
- Computer Vision
SQLAlchemy
Plotly
- Dimensional modeling
- Fact Data Modeling
Talend
-
Presales Architect
T-Systems Brazil - 2 måneder
-
Assisted in managing the team by providing insights to understand the team's performance, facilitating informed decision-making and strategic planning;
-
Developed visualizations to analyze and identify clients requiring more attention, enabling proactive engagement and relationship management;
-
Utilized gathered insights to optimize resource allocation and prioritize efforts towards proposals with higher success probabilities, resulting in improved efficiency and effectiveness in client interactions.
Teknologier:
- Teknologier:
NumPy
Pandas
R (programming language)
Open source
LaTeX
PyCharm
BigQuery
- CSV
Unix
SciPy
Scikit-learn
Matplotlib
- Data Analytics
PL/SQL
-
Uddannelse
BSc.Informationssystemer
USP - University of São Paulo · 2017 - 2020
Find din næste udvikler inden for få dage, ikke måneder
Book en 25-minutters samtale, hvor vi:
- udfører behovsafdækning med fokus på udviklingsopgaver
- Forklar vores proces, hvor vi matcher dig med kvalificerede, godkendte udviklere fra vores netværk
- beskriver de næste trin for at finde det perfekte match på få dage
