SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Workstream is a Series B, mission-driven HR, payroll, and hiring platform purpose-built for the hourly workforce. The company serves 2.7 billion hourly workers (80% of the global workforce) and counts Burger King, Carl's Jr./Hardee's, IHOP, KFC, and Culvers among its customers. Backed by Founders Fund, BOND, and Coatue, Workstream is expanding its product portfolio to deliver AI-native services that complete mature business outcomes with AI agents plus a small human review layer, at better margins than traditional alternatives.
This is a founding engineer opportunity on Workstream's AI-Native Services (AINS) team. As the first engineer on this new initiative, you will spearhead the team from zero to one, shaping its architecture, setting the engineering bar, and building the team that follows. This is a Staff-level role with direct influence over how AINS grows and a direct line to leadership. You will own the technical components—whether they work, how they fail, and how fast bespoke work becomes reusable. Pricing, deal acceptance, and margin models sit with the Head of Delivery and Head of GTM.
Key responsibilities include:
- Build and operate production agents and agentic workflows: orchestration, tool use, data connections, deterministic logic, state, and permissions.
- Work forward-deployed with design customers—sit in their workflow, map the messy reality, and build against live data.
- Build the evaluation harness: eval datasets from real customer work, pre-launch thresholds, regression suites, and post-launch monitoring.
- Build the human-review layer: exception queues, review tooling, escalation paths, and tracing that makes every agent decision auditable.
- Run shadow-mode parallel runs against the customer's existing process or provider and drive discrepancies to zero before go-live.
- Turn customer-specific work into reusable agents, integrations, and playbooks—productization is the default, customization the exception.
- Handle production incidents end to end, and feed every failure back into evals and controls so it cannot recur.
- Help set the technical bar as the team grows.
This is a full-time, remote position with occasional in-office visits as needed to support close cross-functional collaboration and key team initiatives.
REQUIREMENTS:
- You have personally built and operated at least one AI agent or agentic workflow in production. You can clearly explain how it worked, where it failed, how you evaluated it, and what you kept deterministic or human-led.
- Strong engineering judgment across models, orchestration, APIs, data and integrations, permissions, observability, and evaluations.
- You have worked directly with customers or end users—discovery, implementation, or incident resolution—and communicate clearly with non-technical people.
- You can break a messy workflow into deterministic steps, agentic decisions, human judgment, and exception paths—and define how each part will be tested.
- You care about keeping plausible but wrong AI output away from customers: grounded context, constrained actions, verification, evals, review, monitoring.
- You thrive in zero-to-one ambiguity: shipping weekly, owning outcomes, and doing whatever the service needs—including unglamorous review tooling and data plumbing.
- Exposure to a regulated domain (payroll, payments, tax, benefits, labor compliance) is a plus, not a requirement.
- Demonstrated production experience matters more than title or years.