SlipstreamJobsFresh Startup & VC-Backed Jobs

AIML - Apple Foundation Models: ML Research Engineer, RL & Agentic Reasoning

Lex - Zurich, Switzerland - In-office

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Join Apple's Generative AI team in Zurich as a Machine Learning Research Engineer focused on reinforcement learning (RL) and agentic reasoning for foundation models. You will develop and scale RL methods to improve reasoning, instruction following, multi-turn dialogue, and reduce hallucinations in large language models. Your work will directly impact Apple Intelligence features like Siri, reaching billions of users while contributing to state-of-the-art research. Key responsibilities include designing and training agents with tool use, planning, and API integration capabilities; building and refining reward models, evaluators, datasets, and simulation environments for RLHF, RLAIF, and RLVF; running large-scale experiments and translating findings into both research contributions and practical improvements. You'll collaborate within a Europe-based team of approximately 35 RL/ML experts while coordinating closely with Apple's foundation model groups in Cupertino and New York. Required qualifications: MSc, PhD, or equivalent research/industry experience in Computer Science, Machine Learning, Electrical Engineering, or related field. Strong background in reinforcement learning and deep learning with hands-on experience training large-scale models, particularly LLMs. Proficiency in Python and modern ML frameworks (PyTorch, JAX) with demonstrated experience in distributed training. Ability to collaborate in interdisciplinary teams and communicate complex concepts clearly. Preferred qualifications include publications in top ML/AI venues or equivalent open-source/industry contributions; hands-on experience with tool use, planning, retrieval, and agentic integrations for LLMs; experience with data curation, evaluation frameworks, and safety/guardrail methods; ability to design and implement experiments at scale with innovative approaches to challenging problems.

Similar roles