Skip to content
View dwaradwara's full-sized avatar
  • tbilisi
  • 08:13 (UTC -12:00)

Block or report dwaradwara

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
dwaradwara/README.md

B2B SaaS · Incident Response · APIs · Databases · Observability

I turn customer-impacting technical issues into evidence-backed resolutions and engineering escalations.


👋 30-Second Introduction

I’m a Technical Support Engineer with 3+ years of customer-facing B2B experience, owning technical issues from initial report through resolution or well-documented engineering escalation.

My hands-on work spans Browser DevTools, server logs, REST/HTTP, SQL/PostgreSQL, Linux, Docker, AWS, and observability, with portfolio projects covering B2B SaaS, production support, fintech/payments, databases, and cloud operations.


🧭 Incident Investigation Workflow

flowchart LR
    A[Customer Issue / Alert] --> B[Reproduce]
    B --> C[DevTools / Logs / Metrics / SQL]
    C --> D[Isolate Failing Layer]
    D --> E{Resolve in Support?}
    E -->|Yes| F[Smallest Safe Fix]
    E -->|No| G[L2 / L3 Escalation]
    F --> H[Validate Recovery]
    G --> H
    H --> I[Client Update / RCA / Runbook]
Loading

Reproduce → Collect evidence → Isolate → Resolve or escalate → Validate → Document


📌 Evidence Highlights

8 / 8 → HTTP 503

Controlled upstream outage reproduced consistently in the B2B incident lab.

10 / 10 → HTTP 200

Recovery validated repeatedly after restoring the failed dependency.

114 → 0

RabbitMQ backlog drained after isolating a worker outage in NIGHTWATCH.

~500k rows

PostgreSQL query-plan investigation and targeted indexing in the managed-services lab.

AWS ECS staging

OpsPilot deployed with ECS Fargate, RDS, Redis, Terraform, OIDC, and observability.

ClickHouse + Kafka

Replicated analytics support environment with Kafka ingestion, Keeper, and diagnostics.


⚡ Featured Support Engineering Work

🎯 B2B Incident Response Support Lab

B2B Technical Support · Application Support

DevTools FastAPI PostgreSQL Prometheus

  • Browser DevTools + request-ID log correlation
  • real upstream HTTP 503 dependency outage
  • SEV-1 workflow and L1 → L2/L3 escalation
  • back-office troubleshooting
  • automated outage and recovery validation

🌙 NIGHTWATCH Production Support Lab

Production Support · Application Support

Linux RabbitMQ Redis PostgreSQL Grafana

  • queue backlog and worker outage diagnosis
  • PostgreSQL locks and connection exhaustion
  • Nginx / DNS / HTTP failure isolation
  • Redis dependency troubleshooting
  • metrics, logs, traces, and operational records

🚦 OpsPilot

Production Support · Cloud Support

AWS ECS Terraform PostgreSQL Redis

  • AWS staging environment on ECS Fargate
  • SLOs, error budgets, metrics, logs, and traces
  • GitHub Actions + AWS OIDC
  • transactional outbox and dependency recovery
  • controlled incident drills and runbooks

💳 Fintech Incident Operations Lab

Fintech · Payments Support

Datadog Prometheus PostgreSQL Redis

  • payment incident triage and reconciliation
  • DLQ / queue failure investigation
  • Datadog custom metrics and monitor lifecycle
  • duplicate / ambiguous payment handling
  • stakeholder updates and post-incident analysis

🎯 Hiring Manager Fast Path

Hiring for Start here
Technical / B2B Support B2B Incident Response → Supabase Support
Application / Production Support NIGHTWATCH → L2 Production Support
Cloud / DevOps Support OpsPilot → Linux DevOps Operations
Database / Data Platform Support ClickHouse Support → Supabase Support
Fintech / Payments Fintech Incident Operations → Usage-Based Billing
Linux / Infrastructure Support Ubuntu / KVM → Linux Infrastructure Toolkit
AI SaaS Support AI Support Specialist Lab

🧰 Core Technology Stack

Area Technologies
Support & troubleshooting Browser DevTools, REST/HTTP, JSON, logs, SQL, curl, Bash
Systems & containers Linux, Ubuntu, Docker, Nginx, Kubernetes, KVM/libvirt
Cloud & automation AWS, Terraform/OpenTofu, Ansible, GitHub Actions, ArgoCD
Data & messaging PostgreSQL, Redis, ClickHouse, Kafka, RabbitMQ, Supabase
Observability Prometheus, Grafana, Datadog, CloudWatch, OpenTelemetry, Zabbix
Applications Python, FastAPI, Flask, Spring Boot / Java, LangSmith


📚 Full Technical Portfolio

View all 15 hands-on project repositories

Technical / Application Support

Production / Application Operations

Cloud / DevOps / Infrastructure

Specialist Support


📌 Current Direction

Technical Support Engineering · Application Support · Production Support

Cloud, infrastructure, database, and observability skills support my primary focus: diagnosing and resolving customer-impacting technical issues.


🤝 Connect

Open to Technical Support, Application Support, and Production Support opportunities in Tbilisi / remote EMEA.


Pinned Loading

  1. opspilot opspilot Public

    Production-support and reliability engineering platform on AWS with ECS, PostgreSQL, Redis, Terraform, observability, CI/CD, and controlled incident drills.

    Python

  2. l2-production-support-lab l2-production-support-lab Public

    Hands-on L2 production support lab covering Linux, SQL, Docker, monitoring, alerting, incident management and centralized logging.

    Python

  3. supabase-support-lab supabase-support-lab Public

    Hands-on Supabase support engineering lab covering PostgreSQL, RLS, Auth, Storage, Data API, logs and incident troubleshooting.

    HTML

  4. ai-support-specialist-lab ai-support-specialist-lab Public

    AI SaaS support engineering lab using Python, FastAPI, OpenAI, LangSmith, AWS CloudWatch, incident triage, root-cause analysis, and CI.

    Python

  5. linux-devops-operations-lab linux-devops-operations-lab Public

    Linux DevOps operations lab with Ansible, Docker, Nginx, Zabbix, Grafana, KVM/libvirt, monitoring, incident response, and rollback workflows.

    Jinja

  6. gitops-argocd-devops-lab gitops-argocd-devops-lab Public

    GitOps DevOps lab with Kubernetes, ArgoCD, Docker, GitHub Actions, self-healing, rollback, and incident troubleshooting.

    Python