ServiceNow
ServiceNow

10001+ employees

WebsiteLinkedIn
Information Technology
Software
Cloud Computing
Enterprise Software
IT Service Management
About ServiceNow

ServiceNow is a leading cloud computing company that specializes in digital workflows to help enterprises automate and manage their IT operations, employee workflows, and customer service processes. Founded in 2004, the company provides a comprehensive platform that enables organizations to improve operational efficiencies, reduce costs, and enhance user experiences through automation and AI-driven insights. ServiceNow's products span IT service management, IT operations management, HR service delivery, and customer service management, positioning it as a critical partner for digital transformation across industries. With a strong market presence and a commitment to innovation, ServiceNow serves thousands of customers worldwide, driving enterprise productivity and agility.

5 days ago

Sr. Staff Software Engineer – SRE, Release & Test Platforms

Full-time
Senior, Lead
Software Engineer

Open LazyApply and auto-apply to jobs like this in one click

📋

Description
  • Join us to build the next generation of cloud-native reliability, release, and test platforms that enable engineering excellence, developer productivity, and high-confidence ServiceNow releases through automation, observability, and AI-driven operations.
  • Responsibilities include building and operating cloud-native engineering platforms for software validation, release qualification, and operational readiness; designing production-like release and test environments; developing automated quality gates to assess release health and production readiness; integrating automated testing, observability, reliability signals, and deployment intelligence into CI/CD pipelines; building reusable test frameworks, self-service environments, test data, mock services, and developer productivity tooling; advancing shift-left engineering through automated validation and continuous verification; automating failure detection, policy validation, deployment verification, security checks, and reliability assessments; leading Kubernetes-based platform evolution for scalable test infrastructure and release automation; resolving recurring infrastructure issues through sustainable software, systems, and networking solutions; and partnering with engineering teams on design reviews, architecture standards, and automation-first reliability practices.

🎯

Requirements
  • 12+ years of experience in software, systems, platform, or reliability engineering
  • Deep Kubernetes expertise across architecture, operations, networking, storage, security, autoscaling, and multi-cluster environments
  • Experience building and operating large-scale Kubernetes platforms for cloud-native, mission-critical services
  • Experience integrating Kubernetes with CI/CD, GitOps, automated testing, and deployment validation
  • Experience designing cloud-native platforms for ephemeral environments, release qualifications, and automated validation
  • Proven ability to lead engineering excellence across developer productivity, platform engineering, release confidence, and modernization
  • Experience with progressive delivery (canary releases, feature flags, automated rollback, deployment verification)
  • Experience with chaos engineering, resilience validation, disaster recovery, and reliability assessments
  • Proficiency designing, authoring, testing, and debugging code in a team setting using languages such as Python, Go, Java, or Ruby
  • Experience using AI-assisted engineering for intelligent testing, release risk analysis, incident diagnostics, and operational automation
  • Strong skills in coding, observability, SLOs, and cross-team collaboration to improve reliability and performance
  • Good to have: expertise in observability and monitoring at scale
  • Good to have: experience with DevOps automation, CI/CD pipelines, and agile methodologies (e.g., GitLab CI/CD)
  • Good to have: experience with enterprise-scale test automation frameworks (Playwright, Selenium, Cypress, REST Assured, PyTest, JUnit/TestNG, etc.)
  • Good to have: experience with test orchestration, test impact analysis, flaky test detection, parallel execution, and intelligent regression testing
  • Good to have: experience with service virtualization, contract testing, synthetic testing, and test data management
  • Good to have: experience with infrastructure configuration management tools such as Ansible
  • Good to have: expertise with Kubernetes ecosystem technologies (Helm, Argo CD, Argo Workflows, Kustomize, Istio/Linkerd, Gateway API/Ingress, Prometheus, OpenTelemetry, container runtimes)
  • Good to have: experience implementing GitOps using Argo CD, Flux, or similar
  • Good to have: experience operating Kubernetes across AWS (EKS), Azure (AKS), and Google Cloud (GKE)

🏖️

Benefits
  • Competitive salary
  • Supportive teams and career progression opportunities
  • Benefits plans and programs
  • Mental health resources with coaching and 24/7 support
  • Family support resources and parental leave programs
  • Company-wide designated global well-being days (holidays where everyone is off)
  • Good working culture supporting work-life balance
  • Parental leave programs
  • Childcare and caregiving benefits
  • Learning experience platform and tuition reimbursement program
  • Global cross-functional mentoring program
  • Team building activities, employee belonging groups, volunteering, and community outreach programs