DevOps Engineer @ ClearML | Observability & Kubernetes (CKAD) | Cloud & MLOps infrastructure

Public link to this page at Filippo Brintazzoli - Resume

🌓 Toggle dark mode with cmd/ctrl + shift + L


<aside> 👋

Product-oriented DevOps engineer specializing in Kubernetes, observability, and GPU scheduling infrastructure for MLOps teams - in the cloud and especially on-premises. I build and scale systems that stay legible and reliable under pressure, focused on keeping ML workloads reliable at scale and cutting teams' time from experiment to production. CKAD and Prometheus certified.

</aside>

Experience 👨🏻‍💻

DevOps Engineer (Freelance)

*ClearMl, Remote, Italy – (Apr 2024 - Present)*

All-things infrastructure and platform - Kubernetes focus.

Observability Tech Lead

Infovista (NLA Cloud Platform), Remote, Italy – (Jan 2023 - Apr 2024)

Over the last couple of years, as a Tech Lead and part of the Platform team, I delivered Observability-as-a-Service for the NLA Cloud Platform.

The centralized observability platform is collecting data (metrics, logs, traces) from distributed multi-cluster environments, observing thousands of microservices in production.

A single instance of our observability platform has been keeping up with 5 million active series and ingesting up to 4k logs per second, collected from up to 70 Kubernetes clusters, hundreds of Nodes, and 5k running Pods in production.

Skills in Grafana, Prometheus, Thanos, Loki, Fluent Bit, Fluentd, *Tempo, Argo Workflows,* and *MinIO.*