SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Archer Technologies is an aerospace company building all-electric vertical takeoff and landing aircraft for sustainable air mobility. The AI Products Org develops software for the general aviation industry.
As a Staff Data Engineer on the AI Platform team, you will design, build, and operate the data infrastructure powering large-scale model training and inference. You own the pipelines, storage systems, and data quality mechanisms upstream of the ML platform, ensuring models train on clean, high-throughput, well-governed data.
Key responsibilities:
- Design and maintain high-throughput, fault-tolerant ingestion and transformation pipelines feeding training workloads at scale, focusing on latency, throughput, and correctness
- Build and operate the data lakehouse, defining table formats (Iceberg, Paimon, Parquet), partitioning strategies, and compaction policies optimized for ML consumption patterns
- Instrument pipelines with data quality checks, lineage tracking, and anomaly detection so model failures trace back to data problems quickly
- Partner with ML engineers to define feature stores, dataset versioning, and experiment-to-production data contracts; integrate with tools like MLflow for dataset and artifact tracking
- Work closely with AI researchers, platform engineers, and software engineers to understand data access patterns, optimize query performance, and unblock training runs
Required qualifications:
- 5+ years of professional data engineering experience (excluding internships)
- BS/MS/PhD in Computer Science, Data Engineering, Software Engineering, or related field