Your browser does not support javascript! Please enable it, otherwise web will not work for you.

Site Reliability Engineer II

Home > Python programming jobs

Site Reliability Engineer II in USA

  • PagerDuty
  • Full time
  • Email
  • Atlanta, GA

Responsibilities

  • Support and improve foundational networking, compute platforms, Kubernetes clusters, and ingress and traffic management systems.
  • Harden existing systems and help roll out new infrastructure capabilities to improve platform reliability and scalability.
  • Monitor system health using metrics, logs, and alerts.
  • Participate in 24/7 on-call rotations to detect, respond to, and resolve incidents.
  • Participate in standups, planning, retrospectives, and technical discussions while communicating progress and risks early.
  • Stay current on technical trends and suggest innovative tools and approaches.

Requirements

  • At least 3 years of experience in Site Reliability Engineering, DevOps, or Platform Engineering roles.
  • Hands-on experience operating Linux-based systems in production.
  • Working knowledge of networking fundamentals including load balancing, DNS, TLS, and ingress traffic flow.
  • Experience with container orchestration such as EKS or Kubernetes.
  • Experience with cloud-native infrastructure such as AWS, GCP, or Azure, including networking and compute concepts.
  • Proficiency in at least one programming language such as Python, Ruby, or Go.
  • Experience with Infrastructure as Code such as Terraform or CloudFormation.
  • Preferred: experience with AWS cloud networking concepts including VPCs, subnets, routing, security groups, and load balancers.
  • Preferred: experience operating or contributing to production Kubernetes platforms, including cluster upgrades, networking, or ingress configuration.
  • Preferred: experience with monitoring, observability, and logging platforms such as DataDog, New Relic, SumoLogic, Splunk, Prometheus, or Grafana.
  • Preferred: familiarity with service meshes, ingress controllers, or API gateways such as Envoy, Istio, or NGINX.

Benefits

  • Hybrid work model based in established office locations, including Atlanta, with flexible work arrangements.
  • Competitive salary and comprehensive benefits package.
  • Company equity and Employee Stock Purchase Program, subject to eligibility.
  • Retirement or pension plan, generous paid vacation, paid holidays, and sick leave.
  • Dutonian Wellness Days and HibernationDuty companywide paid days off.
  • Paid parental leave, including up to 22 weeks for a pregnant parent and 12 weeks for a non-pregnant parent, subject to local standards.
  • 20 hours of paid volunteer time off per year.
  • Companywide hack weeks and mental wellness programs.

Salary: $113k - $172k/yr

PagerDuty

PagerDuty builds an AI-powered platform for digital operations management that enhances business resilience and drives operational efficiency. It's designed for enterprises looking to streamline their operations and effectively respond to incidents, ensuring that they can maintain performance in ...

Similar positions

Software Engineer in Test

  • Roadie
  • Full time
  • USA
  • 09/13/2026
  • Salary: Competitive
  • Remote

Principal Software Engineer - Builder Experience

  • Elastic
  • Full time
  • Remote
  • 09/13/2026
  • Salary: €73k-€116k
  • Greece

Principal Software Engineer - Builder Experience

  • Elastic
  • Full time
  • Remote
  • 09/13/2026
  • Salary: zł 369k-zł 584k
  • Poland

Senior Devops Engineer

  • Woliba
  • Full time
  • USA
  • 09/13/2026
  • Remote

Senior Core Infrastructure Engineer

  • Oracle
  • Full time
  • USA
  • 09/13/2026
  • Salary: Competitive
  • Nashville, TN