SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Beamery is a jobs, skills and tasks data platform helping organizations navigate AI and automation challenges across talent lifecycle management. The company powers recruitment, mobility, upskilling, diversity, work architecture and workforce planning for Fortune 500 customers.
As a Principal Platform Engineer, you will solve the toughest reliability, scalability and infrastructure problems with company-wide impact. You'll design architectures and RFCs for platform and infrastructure initiatives, maintain hands-on credibility by helping teams ship scalable services, and partner with Product and Engineering leadership on large multi-team platform initiatives.
Key responsibilities include: setting architectural vision with other Principal Engineers; creating and advocating company-wide standards for operational excellence, observability and incident response; owning and evolving the platform tech radar with data-backed technology decisions; coaching and mentoring engineers across the organization; collaborating with Sales, Customer Success and Fortune 500 customers on solutions; and staying connected to production through on-call rotation.
You'll need deep SRE and cloud infrastructure expertise with a proven track record designing and delivering scalable, reliable cloud-based services. Required experience includes: production Kubernetes cluster management at scale (lifecycle, upgrades, resource optimization, security, high availability); Infrastructure as Code with Terraform and GitOps practices; strong software engineering foundations with Go preferred (NodeJS a plus); observability, SLOs, alerting and cost management for large systems; operational experience with Kafka, MongoDB, PostgreSQL, Elasticsearch and Istio; and a FinOps mindset driving cost-awareness across teams. Familiarity with LLMOps and experience with model gateways, routing (LiteLLM) and model tracking (MLflow) is a bonus.
You should have previous experience as an individual contributor in an Engineering leadership position (Staff+, Principal, Architect) with on-call leadership and incident command at scale.