Skip to content
View PArthur006's full-sized avatar
🧑‍💻
Focado no sucesso!
🧑‍💻
Focado no sucesso!

Highlights

  • Pro

Block or report PArthur006

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
PArthur006/README.md

✨ Hi, I'm Pedro Arthur 👋

💻 Analysis and Systems Development (ADS) Student @ UCB
🚀 Data, AI & Cloud Intern @ First Decision
📊 Focused on Data Engineering & Modern Data Stack

Portfolio

Typing SVG


💡 About Me

I am an Analysis and Systems Development (ADS) student at Universidade Católica de Brasília (UCB) and currently working as an Intern at First Decision, focusing on Data Migration, AI, and Cloud environments.

My goal is to build scalable, efficient, and reliable data architectures. I work daily with the challenges of moving and transforming corporate data to ensure it is ready for analysis and business intelligence.

My daily battlefield involves:
Data Migration & ETL: Extracting and moving critical data from legacy environments (SAP ECC) to modern databases (SAP HANA).
Data Pipelines: Cleaning, transforming, and structuring raw data to make it accessible for analytics.
Continuous Learning: Expanding my foundation in Cloud Computing (AWS/OCI) and big data processing frameworks (PySpark/Databricks).

🎯 Current Goal: Transitioning into a Junior Data Engineer role, solving complex architectural problems, and building end-to-end data pipelines.


🚀 Tech Stack & Tools


📌 Featured Projects

🛡️ Secure Lakehouse Pipeline (Data Engineering, LGPD & Security)

An orchestrated end-to-end data pipeline built with PySpark and Prefect. It extracts data from corporate silos, applies Dynamic Data Masking and Salted Hashing (SHA-256) on PII to comply with privacy laws, and prepares features for Machine Learning.

🛩️ ANAC Data Pipeline (Medallion Architecture & Resilience)

Analytical pipeline processing over 20 years of Brazilian aviation data. Built a Star Schema Data Warehouse (PostgreSQL) using PySpark for Silver layer transformations, with a Graceful Degradation mechanism via SQLAlchemy for high availability.

✈️ Decola-Brasil (Flight Booking Portal)

Backend application focused on software architecture (MVC) and database integration. Built to handle reservations, demonstrating solid backend logic and data structuring.
Demo Decola-Brasil


📈 GitHub Stats



📫 Let's Connect

LinkedIn Gmail

Pinned Loading

  1. My-Personal-Portfolio My-Personal-Portfolio Public

    Portfólio pessoal interativo (SPA) construído com React, com tema dinâmico (Light/Dark), animações e fundo interativo com Particles.js. Vitrine dos meus projetos em Front-end, Back-end e Game Devel…

    JavaScript 1

  2. epf-decola epf-decola Public

    Aplicação web de um sistema de reservas aéreas com mapa de assentos interativo, desenvolvida em Python com o microframework Bottle para a disciplina de Orientação a Objetos.

    Python 1

  3. secure_lakehouse_pipeline secure_lakehouse_pipeline Public

    Arquitetura Data Lakehouse (Medallion) construída com PySpark e Delta Lake. Implementa orquestração com Prefect, adequação criptográfica à LGPD e engenharia de features para Machine Learning.

    Python 1

  4. anac-data-pipeline anac-data-pipeline Public

    Pipeline de ETL e modelagem dimensional (Star Schema) com dados da ANAC. Foco na estruturação de banco de dados relacional para consumo analítico e visualização de métricas do setor aéreo.

    Python 1

  5. FIT-PA FIT-PA Public

    C# 1

  6. SENAI-Curso-FullStack SENAI-Curso-FullStack Public

    Repositório central de códigos, arquiteturas e documentação de testes da Qualificação Profissional Full-Stack (SENAI).

    JavaScript 1