SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
PhysicsX is a deep-tech company building an AI-driven simulation software stack for engineering and manufacturing across advanced industries including Aerospace & Defense, Materials, Energy, Semiconductors, and Automotive. The company enables high-fidelity, multi-physics simulation through AI inference across the entire engineering lifecycle, unlocking optimization and automation in design, manufacturing, and operations.
As IT Lead Engineer, you will own the infrastructure enabling engineers to run CPU/GPU-heavy physics simulations at scale. This is a high-ownership role in a scale-up where you'll build processes as much as run them. You'll be the escalation point when performance issues need real root-cause analysis rather than quick fixes.
Key responsibilities include:
- Own and evolve the HPC environment: cluster administration, job scheduling (Slurm/PBS/LSF), performance tuning, and capacity planning for compute-heavy workloads
- Manage Entra ID governance including Access Reviews, Identity Protection, and Privileged Identity Management at scale
- Provide hands-on IT support across the company with an automation-first mindset
- Mentor and develop junior IT/support engineers, building a culture of ownership and technical rigor
- Act as escalation point for complex or high-priority issues across HPC, dev tooling, and general IT support
- Configure, support and improve developer tooling (GitHub Actions/Workflows, JFrog, Chainguard)
The role operates within a flat organizational structure where good ideas win regardless of hierarchy. You'll work hybrid from the London office with flexibility for remote work, balancing focused ambitious work with sustainable pace.
REQUIREMENTS:
- 7–10 years of experience in IT infrastructure, systems administration, or HPC support, with at least 2–3 years in a lead or senior technical role
- Demonstrable hands-on experience managing HPC infrastructure including cluster administration, job schedulers, and performance tuning for CPU/GPU workloads
- Experience with compliance frameworks (SOC 2, ISO 27001) and ability to translate control requirements into technical implementation (access reviews, logging, encryption, change management)
- Strong Linux administration skills, plus comfort supporting mixed Linux/macOS/Windows environments
- Experience supporting developer tooling and workflows including CI/CD systems, version control, license management, and internal engineering platforms
- Experience in cloud-native environments across multiple cloud vendors
- Strong experience designing and setting up zero trust environments
- Strong networking fundamentals and hands-on operational experience: routing and switching, segmentation, firewalls, VPN, enterprise wireless, DNS/DHCP/IPAM, TLS/PKI, with confident packet-level troubleshooting
- Excellent structured troubleshooting methodology: ability to isolate layers, prove causes, prevent recurrence, and teach the method to others
- Strong automation and scripting ability (Python, Bash, PowerShell) with infrastructure as code and configuration management as default working practice
- Comfort working in a scale-up environment with fewer guardrails and more ownership
NICE TO HAVE:
- Strong AV experience: designing and building AV from scratch
- Compliance automation platforms (Vanta, Drata)
- Relevant certifications (CCNA, CCNP, SC-300, MS-102, RHCE, Jamf)
- Experience configuring security tooling and working with security teams