SlipstreamJobsFresh Startup & VC-Backed Jobs

Data Operations Engineer

Origin Health - Bengaluru, Karnataka, India - In-office

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Origin Medical Research Lab, the research arm of Origin Medical, is seeking a Data Operations Engineer to support AI-driven healthcare solutions focused on maternal health equity. Based in Bengaluru, you'll work at the intersection of clinical data engineering and AI research, developing foundational data services and capabilities that enable advanced healthcare delivery. Key responsibilities include designing and maintaining clinical data pipelines for seamless data ingestion, transformation, and delivery to support AI research projects. You'll collaborate with cross-functional teams to gather data requirements, understand diverse data sources, and implement comprehensive data strategies. You'll develop preprocessing techniques to clean, normalize, and preprocess raw clinical data, making it suitable for training and analysis. The role involves applying feature engineering methods to extract relevant features that enhance AI algorithm performance, as well as evaluating and integrating appropriate tools and technologies for data storage, processing, and monitoring. You'll also design, develop, and implement dataset analysis and visualization tools for clinical research. Performance optimization is critical—you'll optimize data pipelines for scalability and performance, considering data volume, processing speed, and resource utilization. The position emphasizes continuous improvement, staying current with advancements in data engineering, AI research methodologies, and best practices to enhance data operations and streamline processes. Required qualifications include 1–2 years of experience in a DataOps or data engineering role, or comparable exposure through internship or project work. You'll need a bachelor's degree in computer science or related field (information science, computer applications, data analytics). Essential technical skills include proficiency in Python, libraries such as Pandas, Matplotlib, Seaborn, and JSON, plus experience with data science and computer vision concepts. Modular programming practices and version control (Git) are expected. Desirable qualifications include basic PyTorch knowledge, networking fundamentals, database experience (preferably PostgreSQL), data visualization and cleaning tools, and medical image processing familiarity. You should demonstrate strong analytical and documentation skills, comfort working with data engineering at scale, excellent teamwork and communication abilities in English, and an eager-to-learn mindset.

Similar roles