SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Rapyuta Robotics is seeking a Senior Application Support Engineer (L3) to support and maintain deployed robotics and warehouse automation systems in production environments. You will be the primary point of contact for clients facing issues with deployed solutions, ensuring timely and effective resolution.
Key Responsibilities:
Operations Support: Resolve issues faced by customers from ongoing operations including UI, localization, sensor calibration, and integration problems. Investigate and resolve complex system-level issues, collaborating with cross-functional teams as needed. Document issues, root causes, and solutions for internal knowledge sharing and continuous improvement.
CI/CD Deployment: Regularly deploy CI/CD updates to client sites, ensuring minimal downtime and seamless integration. Manage and troubleshoot deployment workflows. Maintain and improve deployment scripts and automation tools.
Site Deployment and Support: Help in setting up telemetry for debugging issues. Setup monitoring to prevent issues from occurring.
Continuous Improvement: Lead deep root cause analysis (RCA) and drive permanent fixes in collaboration with development teams. Create tools to help people do deployments efficiently or debug issues faster. Provide feedback to development teams to improve product reliability and performance.
Requirements:
5+ years of experience in Application Support, Production Support, DevOps, or Robotics System Support. Experience supporting complex distributed systems in production environments. Strong programming skills in Python, C++, or similar languages. Strong experience with Linux environments and command-line tools. Ability to perform deep system debugging using logs, telemetry, and system metrics. Experience with Docker/containerized environments, computer networking fundamentals, distributed systems troubleshooting, CI/CD pipelines and deployment automation, and monitoring and observability tools. Experience performing root cause analysis (RCA) for production incidents. Experience with one or more of: ROS/ROS2 based robotic systems, Autonomous Mobile Robots (AMR), ASRS/Fleet management systems, localization/navigation/mapping systems, or sensor integration (LiDAR, cameras, IMU, etc.). Strong ownership mindset for production systems. Ability to lead incident investigations and guide junior engineers. Strong cross-team collaboration with engineering, QA, and operations. Proactive approach to system reliability and continuous improvement. Ability to perform effectively during high-impact production incidents.