https://github.com/higorcazuza81/higorcazuza81
A little about me
https://github.com/higorcazuza81/higorcazuza81
big-data data-architecture data-engineer data-engineering data-engineering-pipeline database python sql
Last synced: over 1 year ago
JSON representation
A little about me
- Host: GitHub
- URL: https://github.com/higorcazuza81/higorcazuza81
- Owner: higorcazuza81
- Created: 2024-11-17T20:57:19.000Z (almost 2 years ago)
- Default Branch: main
- Last Pushed: 2025-03-29T11:37:23.000Z (over 1 year ago)
- Last Synced: 2025-03-29T12:27:28.413Z (over 1 year ago)
- Topics: big-data, data-architecture, data-engineer, data-engineering, data-engineering-pipeline, database, python, sql
- Homepage:
- Size: 3.55 MB
- Stars: 0
- Watchers: 1
- Forks: 0
- Open Issues: 0
-
Metadata Files:
- Readme: README.md
Awesome Lists containing this project
README
### 👋 Hi there! I'm Cazuza
I'm an engineer, but not the kind that builds buildings. Ever heard the saying "data is the new oil"?
Well, you can think of me as the engineer who designs and builds the drilling rigs. The main difference? Mine aren't in the sea, but in the cloud (AWS, to be specific). And instead of crude oil, these rigs run on a whole lot of code and coffee.
Jokes aside, my job is to build robust, automated data pipelines with Python, Airflow, and dbt so that companies can make smarter decisions. Welcome to my little corner of GitHub!
---
### 🛠️ My Toolbox
---
### ✨ Featured Projects
💡 Asymptora - Our Data Engineering Lab
This is the project portfolio I develop with my wife (who is also a Data Engineer). Asymptora is our space to explore, test, and implement end-to-end data solutions. Our focus is on applying the core pillars of modern data engineering:
-
Resilient Architecture: Designing modular and scalable Data Lakehouses on AWS. -
Reliable Orchestration: Automating and monitoring data workflows with Airflow and dbt. -
High-Performance Processing: Transforming large data volumes with Spark/PySpark.
[View Asymptora's Main Repository]
🚗 FIPE Table Project - Automotive Market Analysis
A data pipeline that extracts, processes, and stores updated data from the FIPE Table (a Brazilian standard for vehicle prices), delivering metrics for price trend analysis and supporting buy/sell decisions in the auto sector.
Tech Stack: Python, AWS Lambda, S3, SQL, Airflow.
🔬 Research Pipeline - Ânima Educação
Responsible for architecting the data pipeline for a research project on Female Entrepreneurship, ensuring data collection, cleaning, and availability for strategic statistical analysis of the results.
Tech Stack: Python, Pandas, dbt, PostgreSQL.
---
### 📊 My GitHub Stats