SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 165,000 - 0 / annual
First Due is seeking a Platform Site Reliability Engineer (SRE) / DevOps Engineer to build, scale, and maintain the infrastructure, deployment pipelines, and operational systems powering their fire and EMS software platform. This is an individual contributor role reporting to the Director of Platform Engineering.
Key responsibilities span three areas:
**Infrastructure & Platform Engineering:** Design, implement, and maintain scalable, secure, and highly available cloud infrastructure. Build and manage infrastructure-as-code solutions to support repeatable and reliable deployments. Continuously improve platform reliability, resiliency, scalability, and performance. Partner with engineering teams to ensure services are designed and operated with reliability and observability in mind. Support disaster recovery, backup, and business continuity initiatives.
**DevOps & Automation:** Design, implement, and maintain CI/CD pipelines to enable efficient and reliable software delivery. Automate manual operational tasks and improve deployment processes across environments. Partner with engineering teams to streamline development workflows and reduce operational overhead. Improve release management practices and deployment strategies. Champion DevOps best practices across the organization.
**Reliability & Operations:** Manage on-call rotations and incident response. Develop and maintain runbooks and operational documentation. Implement monitoring, alerting, and observability solutions. Conduct post-incident reviews and drive continuous improvement. Support distributed and remote engineering organizations.
The ideal candidate is passionate about automation, operational excellence, cloud infrastructure, and continuous improvement. They thrive in fast-paced environments, enjoy solving complex technical challenges, and are comfortable driving consensus on standards and influencing engineering teams to adopt best practices.
**Requirements:**
- 5+ years of experience in SRE, DevOps, or platform engineering roles
- Strong proficiency with cloud platforms (AWS, GCP, or Azure)
- Hands-on experience with infrastructure-as-code tools (Terraform, CloudFormation, or similar)
- Expertise in CI/CD pipeline design and implementation (Jenkins, GitLab CI, GitHub Actions, or similar)
- Strong scripting and automation skills (Python, Bash, Go, or similar languages)
- Experience with containerization and orchestration (Docker, Kubernetes)
- Solid understanding of networking, security, and compliance principles
- Experience with monitoring and observability tools (Prometheus, Grafana, ELK, Datadog, or similar)
- Excellent communication and collaboration skills
- Experience supporting distributed and remote engineering organizations
- Must be authorized to work for any US employer; visa sponsorship not available
- Must pass criminal background check and E-Verify