Government Careers
  • Senior Site Reliability Engineer (Government)

  • SentinelOne
  • New York, New York 10261 United States View Map

At SentinelOne, we are driven by a clear purpose: to give the advantage to those who secure our futureAs AI reshapes how organizations build, operate, and innovate, the responsibility to protect them becomes more critical than everWhen you join SentinelOne, your work helps protect global enterprises, critical infrastructure, and the technologies shaping tomorrowIf you are motivated by meaningful challenges and want your impact to be real, measurable, and global, you will find purpose hereWe're looking for people who are relentlessly curious and committed to continuous learningAI is reshaping every function across our business, and we enable every team member, regardless of role or level, to build fluency in AI tools and conceptsAs a Senior Site Reliability Engineer, you will join our Government SRE team and own both the technical reliability of government environments and the coordination of compliant, efficient deployments into themThis team is at the critical intersection of the unique compliance requirements of a regulated environment and the need to establish a consistent software experience for users and developers in commercial environmentsYou will work closely with cross-functional teams, including security, compliance, operations, validation, and engineering, to lead best practices for cloud infrastructure, continuous delivery, and government release processesDrive continuous software delivery, resolve incidents, run post mortems, and create automation strategies for deployment, self‑testing, and alertingLead and execute incident management for production issues, ensuring rapid recovery, root cause analysis, and preventative follow‑up actionsImprove and optimize the observability strategy by collaborating with application engineering teams to design monitoring solutions that enhance alerting capabilities and reduce noiseDefine, implement, and monitor SLOs, SLIs, and SLAs in collaboration with product and engineering teams to align with business objectivesDesign, develop, and maintain software solutions that address operational, compliance, and pipeline challengesOwn and coordinate all government environment releases, driving process improvements to enhance the release pipeline's efficiency, reliability, and visibilityUnderstand product architecture and service dependencies to manage risk and implement effective testing strategiesPartner cross‑functionally with engineering, product, SecOps, compliance, and leadership teams to align priorities, define testing strategies, and resolve challengesEnsure all infrastructure and deployments meet FedRAMP, government regulations, and industry standards, while maintaining required release documentation and risk assessmentsBenefitsIncentive‑based Wellness ChallengesMedical, Dental & VisionGym ReimbursementCareer Wellness PerksMental Health & MindfulnessPaid Parental Leave401kFlexible Spending AccountsShort & Long Term Disability InsuranceEmployee Assistance ProgramLife InsuranceUnlimited Time OffPaid Sick TimePaid HolidaysHappy HoursParties & CelebrationsTeam Building ActivitiesAll‑Hands & Town Hall GatheringsQualificationsThose who thrive here actively seek out new solutions, experiment thoughtfully, and apply what they learn to drive better, faster, smarter outcomes.Multi cloud experience in AWS/GCP (expertise within AWS preferred).Demonstrated experience with at least one main programming language (Python, Go, Ruby, etc.) and proficiency in bash scripting to improve operational workflows.5+ years of experience in SRE, DevOps, or Infrastructure Engineering for SaaS products, with 4+ years running operations at a large scale.Familiarity with testing strategies and automation in large scale environments.Experience with industry standard observability stacks (Prometheus, Grafana, ELK, OpenTelemetry, etc.) and incident management processes.Strong understanding of compliance frameworks relevant to government deployments (e.g., FedRAMP, DoD, NIST 800 53, NIST 800 137).Familiarity with GitOps frameworks, IaC tooling (Terraform or Pulumi), and deployment strategies (blue green, rolling deploys, canary deploys).Experience working directly with government agencies or in highly regulated industries.2+ years of production experience with a container orchestration system (Kubernetes preferred) and Continuous Delivery.Proven background implementing and supporting FedRAMP, security, risk management, and compliance processes for software releases.#J-18808-Ljbffr

At SentinelOne, we are driven by a clear purpose: to give the advantage to those who secure our futureAs AI reshapes how organizations build, operate, and innovate, the responsibility to protect them becomes more critical than everWhen you join SentinelOne, your work helps protect global enterprises, critical infrastructure, and the technologies shaping tomorrowIf you are motivated by meaningful challenges and want your impact to be real, measurable, and global, you will find purpose hereWe're looking for people who are relentlessly curious and committed to continuous learningAI is reshaping every function across our business, and we enable every team member, regardless of role or level, to build fluency in AI tools and conceptsAs a Senior Site Reliability Engineer, you will join our Government SRE team and own both the technical reliability of government environments and the coordination of compliant, efficient deployments into themThis team is at the critical intersection of the unique compliance requirements of a regulated environment and the need to establish a consistent software experience for users and developers in commercial environmentsYou will work closely with cross-functional teams, including security, compliance, operations, validation, and engineering, to lead best practices for cloud infrastructure, continuous delivery, and government release processesDrive continuous software delivery, resolve incidents, run post mortems, and create automation strategies for deployment, self‑testing, and alertingLead and execute incident management for production issues, ensuring rapid recovery, root cause analysis, and preventative follow‑up actionsImprove and optimize the observability strategy by collaborating with application engineering teams to design monitoring solutions that enhance alerting capabilities and reduce noiseDefine, implement, and monitor SLOs, SLIs, and SLAs in collaboration with product and engineering teams to align with business objectivesDesign, develop, and maintain software solutions that address operational, compliance, and pipeline challengesOwn and coordinate all government environment releases, driving process improvements to enhance the release pipeline's efficiency, reliability, and visibilityUnderstand product architecture and service dependencies to manage risk and implement effective testing strategiesPartner cross‑functionally with engineering, product, SecOps, compliance, and leadership teams to align priorities, define testing strategies, and resolve challengesEnsure all infrastructure and deployments meet FedRAMP, government regulations, and industry standards, while maintaining required release documentation and risk assessmentsBenefitsIncentive‑based Wellness ChallengesMedical, Dental & VisionGym ReimbursementCareer Wellness PerksMental Health & MindfulnessPaid Parental Leave401kFlexible Spending AccountsShort & Long Term Disability InsuranceEmployee Assistance ProgramLife InsuranceUnlimited Time OffPaid Sick TimePaid HolidaysHappy HoursParties & CelebrationsTeam Building ActivitiesAll‑Hands & Town Hall GatheringsQualificationsThose who thrive here actively seek out new solutions, experiment thoughtfully, and apply what they learn to drive better, faster, smarter outcomes.Multi cloud experience in AWS/GCP (expertise within AWS preferred).Demonstrated experience with at least one main programming language (Python, Go, Ruby, etc.) and proficiency in bash scripting to improve operational workflows.5+ years of experience in SRE, DevOps, or Infrastructure Engineering for SaaS products, with 4+ years running operations at a large scale.Familiarity with testing strategies and automation in large scale environments.Experience with industry standard observability stacks (Prometheus, Grafana, ELK, OpenTelemetry, etc.) and incident management processes.Strong understanding of compliance frameworks relevant to government deployments (e.g., FedRAMP, DoD, NIST 800 53, NIST 800 137).Familiarity with GitOps frameworks, IaC tooling (Terraform or Pulumi), and deployment strategies (blue green, rolling deploys, canary deploys).Experience working directly with government agencies or in highly regulated industries.2+ years of production experience with a container orchestration system (Kubernetes preferred) and Continuous Delivery.Proven background implementing and supporting FedRAMP, security, risk management, and compliance processes for software releases.#J-18808-Ljbffr

Government Careers

Government Careers

Government jobs offer stability, competitive benefits, and the chance to make a meaningful impact on your community and country.

Whether you’re starting your career or seeking new opportunities, these roles provide pathways for growth, security, and service.

Explore positions across a wide range of fields and take the first step toward a rewarding future in public service.

Show more

MORE JOBS