Talent Navigator
Back to jobs

Staff Site Reliability Engineer (Government)

SentinelOne

Partner networkOnsiteFull Time

$156k to $200k per year

Posted 7 days ago

Apply now

About this role

  • As a Staff Site Reliability Engineer, you will join our Government SRE team and own both the technical reliability of government environments and the coordination of compliant, efficient deployments into them
  • This team is at the critical intersection of the unique compliance requirements of a regulated environment and the need to establish a consistent software experience for users and developers in commercial environments
  • You will work closely with cross-functional teams, including security, compliance, operations, validation, and engineering, to lead best practices for cloud infrastructure, continuous delivery, and government release processes
  • Drive continuous software delivery, resolve incidents, run post mortems, and create automation strategies for deployment, self-testing, and alerting
  • Lead and execute incident management for production issues, ensuring rapid recovery, root cause analysis, and preventative follow-up actions
  • Improve and optimize the observability strategy by collaborating with application engineering teams to design monitoring solutions that enhance alerting capabilities and reduce noise
  • Define, implement, and monitor SLOs, SLIs, and SLAs in collaboration with product and engineering teams to align with business objectives
  • Design, develop, and maintain software solutions that address operational, compliance, and pipeline challenges
  • Own and coordinate all government environment releases, driving process improvements to enhance the release pipeline’s efficiency, reliability, and visibility
  • Understand product architecture and service dependencies to manage risk and implement effective testing strategies
  • Partner cross-functionally with engineering, product, SecOps, compliance, and leadership teams to align priorities, define testing strategies, and resolve challenges
  • Ensure all infrastructure and deployments meet FedRAMP, government regulations, and industry standards, while maintaining required release documentation and risk assessments

Benefits

  • Medical, dental, and vision coverage
  • Employee assistance program
  • Gym reimbursement
  • Incentive-based challenges
  • Mental health and mindfulness
  • Unlimited Time Off
  • Grandparent Leave
  • Volunteer Time Off
  • Paid Sick Time
  • Paid Holidays
  • 16 weeks Gender-Neutral Parental Leave
  • Restricted Stock Unit Program
  • Flexible Spending Accounts
  • Life Insurance
  • Short and Long Term Disability Insurance
  • 401K
  • Team building activities
  • Celebrations and social gatherings
  • Community volunteering events
  • Global all hands and local town hall events- We’re looking for people who are relentlessly curious and committed to continuous learning. AI is reshaping every function across our business, and we enable every team member, regardless of role or level, to build fluency in AI tools and concepts
  • Those who thrive here actively seek out new solutions, experiment thoughtfully, and apply what they learn to drive better, faster, smarter outcomes
  • Due to Federal Government contract requirements, U.S. Citizenship and a work location in the United States is required.FedRAMP staff may be subject to customer or third-party background checks up to and including Secret Clearance if required by their role at SentinelOne
  • Strong understanding of compliance frameworks relevant to government deployments (e.g., FedRAMP, DoD, NIST 800 53, NIST 800 137)
  • Experience working directly with government agencies or in highly regulated industries
  • 2+ years of production experience with a container orchestration system (Kubernetes preferred) and Continuous Delivery
  • Demonstrated experience with at least one main programming language (Python, Go, Ruby, etc.) and proficiency in bash scripting to improve operational workflows
  • Familiarity with testing strategies and automation in large-scale environments
  • Multi cloud experience in AWS/GCP (expertise within AWS preferred)
  • Experience with industry standard observability stacks (Prometheus, Grafana, ELK, OpenTelemetry, etc.) and incident management processes
  • Proven background implementing and supporting FedRAMP, security, risk management, and compliance processes for software releases
  • 8+ years of experience in SRE, DevOps, or Infrastructure Engineering for SaaS products, with 4+ years running operations at a large scale
  • Familiarity with GitOps frameworks, IaC tooling (Terraform or Pulumi), and deployment strategies (blue green, rolling deploys, canary deploys)

Required skills

rapid-recoverykubernetespythongolangrubyamazon-web-servicesgoogle-cloud-platformprometheusgrafanaelkopentelemetryterraformpulumigreenhouse

Similar roles

United StatesFull TimeStaff

$153k to $259k per year

gitlabpostgresqlkubernetesgreenhouse
Posted yesterdayView role
United KingdomFull TimeSenior
rubygitlabpostgresqlrails
Posted yesterdayView role
NetherlandsFull TimeSenior
pythonc-sharpreactguesswork
Posted yesterdayView role