SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Gather AI is building a vision-powered platform that uses autonomous drones and existing equipment to digitize warehouse operations and supply chain workflows. The company is pioneering robotics-driven warehouse intelligence to improve efficiency, safety, and on-time delivery.
As Distributed Systems Architect, you will serve as a technical anchor for the engineering organization, responsible for evolving mature, complex SaaS systems and setting technical direction across the entire stack. Your scope spans backend, frontend, platform, and infrastructure decisions.
Key responsibilities include:
• Own enterprise-grade architecture: Align backend, frontend, platform, and infrastructure decisions into a coherent system that scales reliably.
• Replace firefighting with predictable execution: Lead efforts to retire scalability and reliability debt, safely upgrade critical infrastructure, and eliminate single-threaded dependencies.
• Set engineering standards: Drive high-level decisions on testing strategies, validation guardrails, and system design that impact delivery speed and maintainability.
• Modernize production systems: Apply database-level strategies (schema evolution, replication tradeoffs) and platform primitives (orchestration, runtime constraints) as architectural tools.
• Mentor senior talent: Establish shared practices, align terminology, and elevate the engineering organization through mentorship and knowledge transfer.
• Apply first-principles thinking: Develop deep understanding of data systems, latency-sensitive services, and failure modes.
Required expertise:
• Distributed systems: Mastery of performance-sensitive REST APIs, caching strategies, and core distributed systems concepts.
• Relational databases: Expert-level PostgreSQL (upgrades, query tuning, replication, schema evolution) and advanced SQL.
• Cloud & containers: Deep production experience with AWS and/or Azure, Kubernetes, and Docker.
• Backend development: Strong production experience in Node.js and/or Python.
• Reliability & security: Hands-on observability (logs, metrics, traces), CI/CD, SRE fundamentals (SLIs/SLOs), IAM, and secrets management.
Bonus experience includes logistics/warehouse/robotics domain knowledge, reasoning about multi-regional architectures, driving org-wide standards, and familiarity with applied ML and LLM/VLM tooling.