Skip to content

ana@sre: ~

whoami

Profile picture of Ana Luiza Primo

Hi! I'm Ana Luiza

Site Reliability Engineer/DevOps Engineer | International Career

cat sobre.txt

# about

About me

I'm a Site Reliability Engineer focused on the reliability, observability and performance of production systems. I've worked on high-traffic customer-facing platforms, defining SLOs/SLIs, bringing MTTR down and building observability end to end.

I also create content about SRE/DevOps careers and mentor people who want to grow in the field, including those aiming at the international market.

# experience

Professional experience

March 2025 – present

Site Reliability Engineer

International contracts — United States · Remote

  • Production deploys and canary deployment automation with GitHub Actions.
  • Rollout of observability tooling.
  • Building AI-assisted observability solutions.
  • On-call rotation.

2024 – March 2025

Site Reliability Engineer — M level

Itaú Unibanco · São Paulo, Brazil

  • Observability team supporting the main channels of the bank's app — login, home and authentication.
  • Mapping the applications and their architectures in order to act on problems and incidents.
  • Building alerts and dashboards with the team to monitor the app and keep it available to customers.

2021 – 2024

Site Reliability Engineer — M level

iti — Itaú's digital bank · São Paulo, Brazil

  • Incident orchestration, channel support and post-mortem ceremonies.
  • SLO culture and proactive observability; instrumentation of .NET Core and Kotlin microservices.
  • Supported hundreds of microservices on Kubernetes/AWS EKS with Splunk, Grafana, AppDynamics, Jaeger, Loki and Elasticsearch.
  • On-call engineer, eliminating toil through automation and running performance tests with JMeter.

2020 – 2021

Site Reliability Engineer — J level

Stone Payments · São Paulo, Brazil

  • Operations and infrastructure team owning the full service lifecycle: deployment, availability, performance, change management and emergencies.
  • Capacity planning and infrastructure-as-code automation on Google Cloud.
  • Worked closely with development teams and with governance, ensuring adherence to standards and compliance.

education

  • Production Engineering — INATEL · 2021–2025
  • Project Management — Ohio University · 2025
  • Business English — Ohio University · 2025

stack & tools

  • Grafana
  • Prometheus
  • Loki
  • OpenTelemetry
  • Datadog
  • Zabbix
  • AWS EKS
  • Azure
  • Kubernetes
  • Terraform

# projects

My projects

  • 100DiasDeKubernetes

    A public log of the #100DiasDeKubernetes challenge written entirely in Portuguese, translating and expanding on Anais Urlichs’ material with Brazilian references. A collaborative repository, open to forks and pull requests.

    • Kubernetes
    • Docs
    • Community
  • datadog-automation

    A Flask API that generates Datadog dashboards and monitors from application and AWS account data, covering 9 dashboard types and 12 monitor types. Includes an interactive wizard and ships containerized with Docker and Nginx.

    • Python
    • Datadog
    • Observability
  • arquitetura-celular

    A Terraform project that provisions a cell-based architecture on AWS: independent cells in separate AZs, each with its own VPC, NAT gateway, EKS cluster and ALB — so failures stay contained within a cell.

    • Terraform
    • AWS
    • High availability

# newsletter