, and the discipline to design, automate, and operate complex services so that reliability becomes a first-class engineering... objectives (SLOs), service-level indicators (SLIs), and error budgets for critical services, and use those measures to drive...
using Kubernetes (EKS), AWS services, Helm charts, and containerized deployment strategies. Collaborate with business.... Drive CI/CD automation, platform monitoring, and operational excellence using Jenkins, GitHub Actions, Prometheus, Grafana...
, and connector configurations using GitOps patterns. Implement comprehensive observability using Prometheus, Grafana, Datadog.... Experience operating Kafka on Kubernetes (Strimzi, Confluent Operator). Exposure to managed Kafka services (AWS MSK, Azure Event...
, and connector configurations using GitOps patterns. Implement comprehensive observability using Prometheus, Grafana, Datadog.... Experience operating Kafka on Kubernetes (Strimzi, Confluent Operator). Exposure to managed Kafka services (AWS MSK, Azure Event...
, cloud migration strategies, and data services architecture on a global scale. Drive platform automation (IaC) utilizing..., etc.) and implementing modern automation and monitoring tools (Ansible, Terraform, Prometheus, Grafana). Demonstrated experience managing...
, observability, and incident management solutions using tools such as Prometheus, Grafana, ELK Stack, CloudWatch, Datadog, Splunk... management, monitoring, and security services used in enterprise cloud environments. Hands-on production experience designing...
, and connector configurations using GitOps patterns. Implement comprehensive observability using Prometheus, Grafana, Datadog.... Experience operating Kafka on Kubernetes (Strimzi, Confluent Operator). Exposure to managed Kafka services (AWS MSK, Azure Event...
or professional-services background. Federal or defense-customer delivery experience. Infrastructure-as-Code expertise (Terraform...Senior Software Engineer (Cloud / DevOps) – AWS Professional Services Location: Chantilly, VA (Onsite Preferred...
available cloud platforms on Amazon Web Services. This is a deeply hands-on engineering role spanning architecture, infrastructure... environments across compute, networking, storage, identity, and managed data services, with strong attention to scalability...
, and the discipline to design, automate, and operate complex services so that reliability becomes a first-class engineering... objectives (SLOs), service-level indicators (SLIs), and error budgets for critical services, and use those measures to drive...