SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Antimetal is building the future of infrastructure management through AI-powered agents that investigate, resolve, and prevent infrastructure issues. As a Platform Engineer, you'll own the foundational systems that power the entire engineering organization's velocity and product delivery.
You'll work across four tightly coupled areas, with flexibility to match scope to your strengths:
**Platform Acceleration**: Architect and optimize development infrastructure for agent and product work, including internal admin tooling (Anvil), CI/CD, deploy pipelines, dev/demo environments, and Claude-Code-native skills. Your work multiplies productivity across the entire engineering org.
**Service Infrastructure**: Build and maintain core infrastructure powering Antimetal's services on Kubernetes, including deployment pipelines, observability (OTEL, Datadog), shared libraries, secret management, and platform abstractions that let product teams ship reliable services without managing the substrate.
**Data Infrastructure**: Own the schema, storage, and access patterns for critical systems including investigation traces, agent trajectories, resource graphs, customer telemetry, and typed data structures. You'll ensure data integrity is non-negotiable and know when to normalize versus defer structure.
**Connectivity**: Own the MCP gateway routing every tool call, OAuth and credential management for customer integrations, self-hosted MCP server deployment, and the unified integration catalog. Reliability and enterprise trust are paramount—token refresh at scale, proactive health checks, and isolation that holds under stress.
You bring at least 3 years of platform, infrastructure, or backend engineering experience, preferably at high-growth or AI-native companies. You have strong fundamentals in service-oriented architectures, networking, and systems design. You're proficient in TypeScript with hands-on Kubernetes production experience. You can take ambiguous goals from prototype to production using customer feedback and operational data. You have strong code and data modeling skills and take a product-focused approach to platform work, measuring success by engineer adoption. You use AI aggressively in your workflow and are a heavy Claude Code user.
Bonus experience includes MCP, OAuth, API gateways, multi-tenant integration platforms, modern observability, SRE best practices, internal admin/developer tooling, and agent/LLM systems work.