SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Foodics is a leading restaurant management ecosystem and payment tech provider founded in 2014, headquartered in Riyadh with offices across 5 countries. The company serves customers in over 35 countries and has processed over 6 billion orders. Foodics recently raised $170 million in the largest SaaS funding round in MENA.
You will build internal AI systems that engineers use daily: agents that generate and maintain tests, pipelines that triage failures before humans see them, tooling that speeds up code review and debugging, and evaluation infrastructure that makes AI features testable. This is a builder's role where you write production code, own systems in CI, and are measured on whether engineers actually use what you ship.
Key responsibilities include:
**Test Automation Across the Stack:**
- Backend: API and contract testing, service-level and integration coverage, data setup, and test design that survives schema changes
- Frontend: web E2E and component-level coverage with Playwright, visual and RTL regression, and fast suites that gate merges
- Mobile: native and cross-platform coverage with Appium or Maestro, device-farm strategy, offline and sync behavior, and payment-peripheral paths
- Shared fixtures, environment and test-data management, parallelization, and CI pipelines with meaningful signals
- Performance and load testing
**AI Layer:**
- Test generation from specs, code, and production traffic with maintenance solutions
- Failure triage that classifies red builds: real bug, flake, environment issue, or test rot
- Self-healing locators and suite health tooling—flake detection, quarantine, coverage-gap analysis
- Evaluation infrastructure for AI features: datasets, scoring, regression detection
- Region-specific evaluation for Arabic/English behavior, RTL interfaces, and POS/tax/payment rules
**Agentic AI and Orchestration:**
- Agentic AI that performs real work: reads diffs, runs relevant suites, reproduces failures, proposes fixes, opens PRs
- Agents that own quality workflows end-to-end—exploratory testing, coverage-gap hunting, release-risk assessment
- Orchestration under load: multi-step planning, tool use, retries, state/memory, sandboxed execution, multi-agent handoffs
- Integration with existing stack (CI, Jira, observability, MCP-style interfaces)
- Judgment to know when a plain pipeline beats an agent
**Technical Foundation:**
You should be current on test automation frameworks (Playwright, Appium, Maestro), agentic AI orchestration and tool use, context engineering (retrieval, chunking, reranking, caching), evaluation methodologies (offline/online evals, LLM-as-judge), reliability patterns (structured output, guardrails, fallback design), and LLM operations (tracing, prompt management, latency/cost budgeting, model routing).
**Requirements:**
- Production software engineer with recent hands-on work on LLM-backed systems that real users depend on
- Strong Python; comfortable in at least one of .NET, Java, or TypeScript
- Tested, maintained code—not notebooks
- Real automation depth across multiple surfaces (backend and UI—web or mobile); experience keeping suites green without deleting hard tests
- Real experience building evaluation systems with measurable improvements
- Practical depth with modern LLM toolkit: prompting, structured output, tool use, retrieval, agentic AI orchestration, with clear understanding of trade-offs
- Credible testing fundamentals; test design, automation frameworks, and CI/CD should not be new
- Bias toward adoption; measure work by what other engineers use, not by demos
Foodics offers inclusive and diverse culture with flexibility and hybrid work, competitive compensation with bonuses and share potential, personal development with training and annual learning stipend, a talented team of 30+ nationalities across 14 countries, and autonomy with mentoring and challenging goals.