SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
FourKites is an AI-driven supply chain visibility platform that processes over 3.2 million supply chain events daily for 1,600+ global brands. As a Senior Data Scientist, you will design, build, and productionize machine learning models that power core prediction problems across the platform, including ETA/ATA forecasting and message-based status extraction.
You will own models end-to-end: from data pipeline design through training, deployment, monitoring, and retraining. Working with real-world logistics data (GPS pings, check calls, carrier data), you'll develop regression, classification, and time-series forecasting models, as well as NLP/LLM-based extraction pipelines for text-based ETA and status updates. You'll diagnose gaps between offline evaluation and live production performance, build automated training pipelines using Airflow, and maintain model monitoring and observability using tools like Grafana.
Key responsibilities include replacing manual or rule-based processes with ML-driven automation, translating model improvements into measurable business impact (operational savings, efficiency gains), and mentoring other data scientists on technical approach and best practices. You'll make independent build-vs-buy and architecture tradeoff decisions and collaborate closely with product, engineering, and operations teams.
Required qualifications: strong ML fundamentals (regression, classification, time-series forecasting), NLP experience (text extraction, entity recognition, LLM-based extraction), proven production ML experience shipping models at scale, strong Python and SQL skills (pandas, scikit-learn), cloud and data infrastructure experience (AWS, Airflow), model monitoring tooling experience (Grafana), comfort with noisy real-world data, ability to diagnose offline-to-production performance gaps, track record of ML-driven automation, and excellent cross-functional communication skills.
Nice-to-have: logistics/supply chain domain experience, real-time/streaming data experience (Kafka), and exposure to LLM/GenAI applications in production.