Skip to content
View nevesm's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report nevesm

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
nevesm/README.md

Marcus Neves

Site Reliability & Platform Engineer. I keep transactional systems online and cloud bills honest.

Eight years in infrastructure, six of them in DevOps/SRE across payments and financial services — environments where a minute of downtime is revenue that doesn't come back. That teaches you two things quickly: automate everything that can fail at three in the morning, and don't spend a cent more than reliability requires.

I also write backend and full-stack code. Not as a hobby — because infrastructure decisions you can't defend at the application layer tend to be the wrong ones.

What I work on

FinOps — finding and removing waste in cloud accounts. Rightsizing, commitment coverage, spot instances in production with workloads prepared for interruption, storage and network. A 30%+ reduction is the result I've repeated most.

Observability — Prometheus, Grafana, Loki, Tempo, OpenTelemetry. I've led migrations from commercial SaaS tooling to an open stack, with alerts versioned as code. Half that work isn't migrating — it's cutting the noise nobody acts on.

Reliability — SLIs and SLOs with error budgets, end-to-end incident management, blameless post-mortems, and steadily removing toil.

Where I've operated — payments and financial services, including instant-payment infrastructure on Kubernetes and checkout security layers (WAF, DDoS mitigation, fraud prevention). Fleets of hundreds of Linux and Windows servers across hybrid estates.

Currently

  • SRE and cloud infrastructure for a cloud-native platform in the energy sector (US, remote — end client under NDA)
  • Building platform-kit — the Terraform and Helm stack I deploy on every observability migration
  • Writing about cloud cost, observability and what actually breaks in production

Tools I reach for

Kubernetes Terraform Ansible Helm · AWS Azure GCP OpenStack OCI · Prometheus Grafana Loki Tempo OpenTelemetry · Linux Nginx RabbitMQ PostgreSQL SQL Server · Python Go TypeScript Java · GitLab CI Jenkins Azure DevOps

Elsewhere

I work under NDA and don't discuss client environments — including former ones. What I can show you is how I think: it's in the repositories and in what I publish.

🇧🇷 Português

Engenheiro de confiabilidade e plataforma. Mantenho sistemas transacionais no ar e contas de nuvem honestas.

Oito anos de infraestrutura, seis deles em DevOps/SRE nos setores de pagamentos e serviços financeiros — ambientes onde um minuto parado é receita que não volta. Isso ensina duas coisas rápido: automatize tudo que pode falhar às três da manhã, e não gaste um centavo a mais do que a confiabilidade exige.

Também escrevo backend e full-stack. Não como hobby — porque decisão de infraestrutura que você não consegue defender na camada de aplicação costuma ser a decisão errada.

No que trabalho

  • FinOps — encontrar e eliminar desperdício em contas de nuvem. Rightsizing, cobertura de compromisso, spot em produção com aplicação preparada para interrupção, storage e rede. Redução de 30%+ é o resultado que mais repeti.
  • Observabilidade — Prometheus, Grafana, Loki, Tempo, OpenTelemetry. Já liderei migrações de ferramentas SaaS proprietárias para stack aberta, com alertas versionados em código. Metade desse trabalho não é migrar: é cortar o ruído que ninguém atende.
  • Confiabilidade — SLI/SLO com orçamento de erro, gestão de incidente ponta a ponta, post-mortem sem culpado, e redução contínua de toil.
  • Onde já operei — pagamentos e serviços financeiros, incluindo infraestrutura de pagamentos instantâneos em Kubernetes e camada de segurança de checkout (WAF, mitigação de DDoS, prevenção a fraude). Parques com centenas de servidores Linux e Windows em ambiente híbrido.

Atualmente

  • SRE e infraestrutura de cloud para uma plataforma cloud-native do setor de energia (EUA, remoto — cliente final sob NDA)
  • Construindo o platform-kit — a stack Terraform e Helm que subo em toda migração de observabilidade
  • Escrevendo sobre custo de nuvem, observabilidade e o que realmente quebra em produção

Trabalho sob NDA e não falo de ambiente de cliente — inclusive dos anteriores. O que dá para mostrar é como eu penso: está nos repositórios e no que eu publico.

Pinned Loading

  1. ghpr ghpr Public

    🖥️✅🟢 CLI para aprovar pull requests através do terminal

    Shell 3

  2. myengineer-io/aws-notificator myengineer-io/aws-notificator Public

    ☁️🔔💬 Um notificador de alertas da AWS para WhatsApp e Discord, plug and play!

    HCL 27 8

  3. donmarx/showdownhelper donmarx/showdownhelper Public

    JavaScript 1 1