SlipstreamJobsFresh Startup & VC-Backed Jobs

AI Vision Engineer

Nexxa.ai - Toronto, ON, Canada - In-office - posted 2026-08-27

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Nexxa.ai is building AI systems for heavy industries, enabling machines and operations to think, decide, and act autonomously across manufacturing, infrastructure, and logistics. We are seeking an AI Vision Engineer to design, build, and deploy next-generation computer vision systems for real-world industrial applications. You will work across the full vision stack—classical and deep-learning-based computer vision, vision-language models (VLMs), multimodal reasoning, and real-time inference on edge and cloud platforms. You'll collaborate closely with AI and engineering teams to develop production-ready vision solutions, improve model performance, and shape the next generation of intelligent visual systems. Key responsibilities include: - Design, train, evaluate, and deploy computer vision models for industrial applications - Build and optimize CV pipelines for object detection, segmentation, classification, OCR, tracking, and visual understanding - Develop and fine-tune vision-language models for multimodal reasoning, visual question answering, and document understanding - Design and optimize real-time inference pipelines for edge devices and cloud deployment - Build scalable data pipelines for image and video collection, annotation, augmentation, training, and evaluation - Fine-tune and evaluate open-source vision and multimodal foundation models - Develop robust evaluation frameworks and benchmarks measuring accuracy, robustness, latency, and business impact - Optimize models for production constraints including quantization, pruning, and hardware-accelerated inference - Collaborate with product, engineering, and research teams to translate business requirements into technical solutions - Contribute to architecture decisions, technical design reviews, and AI/CV best practices - Stay current with advancements in computer vision, multimodal AI, and autonomous systems We're looking for a builder comfortable working across classical computer vision, deep learning, and multimodal generative AI. You should be curious, adaptable, eager to learn emerging technologies, and excited about solving challenging real-world problems. You take ownership, move quickly, and thrive with autonomy. REQUIREMENTS: - Bachelor's degree in Computer Science, Computer Engineering, Electrical Engineering, Mathematics, Statistics, Artificial Intelligence, or related technical field (or equivalent practical experience) - 3+ years of industry experience in Computer Vision, Machine Learning, Applied AI, or related fields - Demonstrated experience independently owning and delivering computer vision projects from concept to production - Strong Python programming skills - Hands-on experience with PyTorch and modern deep learning workflows - Experience developing and deploying computer vision models (detection, segmentation, classification, OCR) in production - Experience with image and video processing libraries such as OpenCV - Experience with common CV/detection frameworks (YOLO, Detectron2, MMDetection, or similar) - Experience working with VLMs, multimodal models, or Generative AI applications - Strong understanding of machine learning fundamentals, model evaluation, experimentation, and deployment - Experience with Hugging Face Transformers and open-source AI ecosystems - Familiarity with data annotation workflows, dataset curation, hyperparameter optimization, and inference optimization - Experience building production-grade software and AI systems - Strong analytical and problem-solving skills - Excellent communication and collaboration skills PREFERRED QUALIFICATIONS: - Master's degree in Computer Science, AI, Machine Learning, Computer Vision, or related field - Experience with OCR, document understanding, or visual reasoning systems - Experience with Vision-Language Models and multimodal AI applications - Familiarity with 3D vision, SLAM, or sensor fusion (camera + LiDAR) for industrial/robotics applications - Experience with real-time inference optimization (TensorRT, ONNX Runtime, quantization, pruning) - Experience deploying models on edge hardware (NVIDIA Jetson, embedded systems) - Experience with LangChain, LangGraph, or agentic AI frameworks combining vision and language models - Experience with vector databases and visual/semantic retrieval systems - Experience deploying AI systems on AWS, GCP, or other cloud platforms - Experience with Docker, Kubernetes, and MLOps workflows - Experience with PostgreSQL and large-scale data systems - Experience with model serving, distributed training, and inference optimization at scale - Contributions to open-source projects, technical blogs, research publications, Kaggle competitions, or demonstrable CV/AI work - Experience in startup or high-growth environments

Similar roles