This is your work, valued

Florianópolis, SC - Brazil

Arthur Raulino Kretzer

Advanced
@ArthurKretzer

I specialize in data engineering and artificial intelligence, building scalable data pipelines and enabling advanced analytics for smart manufacturing systems.

data_engineering_exercises. This is a repository for data engineering exercises. It has exercises from data scrapping to airflow.

9

tutorial-ming-stack. This repo was created for a tutorial on XIV Symposition on Computing Systems Engineering (https://sbesc.lisha.ufsc.br/sbesc2024/Home) held on November 26-29, 2024 in Recife - PE - Brazil, called "How to generate value in industry with IoT data with a simple service stack?".

5

lpbf-layer-segmentation. This project was made in Alkimat during a computer vision master's course on UFSC (Universidade Federal de Santa Catarina). The scope of the development was to segment a printed layer on an image shoot from a Laser Powder Bed Fusion (LPBF) machine developed in Alkimat.

5

streaming-pipeline. This project contains experiments for Data Lakehouse streaming pipelines in Kubernetes for my master thesis comparing edge and cloud deployments.

4

dbt-with-spark-thrift-server. This project provides an open-source framework for building modern data lakehouses by integrating Hive Metastore for schema management, Spark Thrift Server for distributed query execution, and DBT for SQL-based transformations. Designed for development and experimentation, it also includes MinIO as a local S3-compatible object store.

4

airflow-and-minio-example. Jupyter Notebook

3

debezium-cdc-example. A Change Data Capture (CDC) example using Debezium for data streaming.

2

digitalocean-k8s-iac. This repository contains Terraform code to provision a Kubernetes cluster with 4 nodes on DigitalOcean.

2