Site Reliability Engineer (SRE)
REMOTE, Remote
Job Id:
0000175164
Job Category:
Information Technology
Job Location:
REMOTE, Remote
Security Clearance:
No Clearance
Business Unit:
Piper Companies
Division:
Not Defined
Position Owner:
Anne Green
Piper Companies is seeking a Site Reliability Engineer (SRE) to join the Ops Stack / Platform Engineering team with a leading enterprise organization in a remote capacity. The Site Reliability Engineer (SRE) will focus on observability, system health, infrastructure patching, vulnerability remediation, and automation while supporting large-scale platform and infrastructure environments. The Site Reliability Engineer (SRE) is a long term contract opportunity that allows you to work remote in the United States.
Responsibilities of the Site Reliability Engineer (SRE):
• Build, maintain, and enhance observability and monitoring capabilities across enterprise infrastructure and platform environments.
• Develop dashboards, alerts, metrics, and logging strategies to improve visibility into system health, availability, and performance.
• Monitor platform availability, reliability, capacity, and performance and proactively identify potential issues.
• Own and improve OS and infrastructure patching processes, including scheduling, deployment, validation, and remediation.
• Support vulnerability remediation initiatives and collaborate with security teams to resolve infrastructure findings.
• Automate operational processes involving patching, monitoring, remediation, and routine system maintenance.
• Troubleshoot complex production and infrastructure issues and conduct root cause analysis.
• Partner with Platform Engineering, Infrastructure, Security, and application teams to improve platform reliability.
Requirements of the Site Reliability Engineer (SRE):
• 4+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, Cloud Engineering, or Infrastructure Engineering.
• Strong experience with observability, monitoring, logging, alerting, and system-health strategies.
• Hands-on experience with infrastructure patching, vulnerability remediation, or infrastructure lifecycle management.
• Experience supporting Linux and/or Windows environments.
• Strong scripting and automation experience using Python, PowerShell, Bash, or similar technologies.
• Familiarity with infrastructure automation and configuration management technologies such as Ansible, Terraform, Puppet, or Chef.
• Experience with observability platforms such as Splunk, Datadog, Dynatrace, Grafana, Prometheus, Elastic, or comparable technologies preferred.
• Familiarity with containers and Kubernetes is a plus.
Compensation for the Site Reliability Engineer (SRE):
• $120,000 - $145,000
• Full Comprehensive Benefits: Health, Vision, Dental, PTO, Paid Holiday and Sick Leave if Required by Law.
Keywords:
Site Reliability Engineer, SRE, Site Reliability Engineering, Platform Engineering, DevOps, Cloud Engineering, Infrastructure Engineering, observability, monitoring, logging, alerting, system health, infrastructure patching, OS patching, vulnerability remediation, vulnerability management, automation, Python, PowerShell, Bash, Linux, Windows, AWS, Azure, GCP, Ansible, Terraform, Puppet, Chef, Splunk, Datadog, Dynatrace, Grafana, Prometheus, Elastic, Kubernetes, containers, incident response, root cause analysis, RCA, infrastructure automation, capacity planning, high availability, production support, enterprise infrastructure
This job opens for applications on 9/25/2026. Applications for this job will be accepted for at least 30 days from the posting date.
#REMOTE
#LI-AG1