SlipstreamJobsFresh Startup & VC-Backed Jobs

Senior Software Engineer, Site Reliability Engineering

Thumbtack - Remote - Remote - posted 2026-10-02

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Salary: CAD 180,200 - 233,200 / annual

Thumbtack is a home services marketplace that helps millions of homeowners manage repairs, maintenance, and improvements by connecting them with local service professionals. The Site Reliability Engineering team is responsible for building and maintaining a reliable, secure, and scalable platform that serves the entire Thumbtack ecosystem. In this role, you will be a key contributor to SRE, designing and supporting resilient systems with a focus on high performance, availability, and throughput. You'll minimize service disruptions, downtime, and latency across the stack—from Linux systems to customer-facing applications. Your work will have high leverage, impacting how Engineering, Applied Science, and other teams deliver, run, and observe systems. You will design, create, and maintain software and systems to improve availability, scalability, and efficiency of Thumbtack's services. You'll set the architectural direction of infrastructure and platform services while supporting the broader engineering organization. You'll design and implement tools and processes for deployment, change management, and infrastructure management. You'll troubleshoot and debug critical systems throughout the software development lifecycle, contribute to the evolution of platform capabilities, perform capacity planning and demand forecasting, and participate in rotating on-call duties. This is a collaborative role where you'll work across product development, developer experience, and backend infrastructure teams to build Thumbtack's ecosystem of platform services. REQUIREMENTS: - 5+ years of experience managing infrastructure and systems - Extensive fluency in AWS and Linux - Ability to effectively read, write, and debug code in production-facing systems - Expertise in designing, analyzing, and troubleshooting large-scale distributed systems across web technologies (DNS, TLS, HTTP/S, TCP/IP) - Demonstrable knowledge of instrumenting, operating, and observing distributed microservices in a production cloud environment - Ability to decompose complex problems while understanding necessary tradeoffs - Clear and effective communication with cross-functional partners at various technical levels

Similar roles