SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 180,000 - 220,000 / annual
You.com is building the AI Search Infrastructure that powers modern AI systems, creating a trusted knowledge layer for agents, applications, and enterprises to retrieve real-time, accurate, and citation-backed information. The platform combines proprietary vertical indexes with LLM-optimized retrieval systems.
You will join as a hands-on Data Engineer to build and scale the modern data platform. Working closely with Finance, Engineering, Product, and Analytics teams, you'll develop reliable, high-performance data pipelines and systems for both batch and real-time data processing.
Key responsibilities include:
- Build and maintain scalable data pipelines (batch and streaming) using Databricks, Spark, Kafka, and AWS services
- Develop ETL/ELT workflows using DBT, PySpark, SQL, and Fivetran
- Build pipelines from source systems (Salesforce, billing, product events, API logs) into clean analytics layers
- Work with Finance on data solutions, metrics, forecasting models, revenue accounting, COGS, and margin reporting
- Partner with marketing and growth teams on segmentation, campaign targeting, and lifecycle analytics
- Develop and maintain reverse ETL pipelines to sync data to Salesforce, HubSpot, Braze, and other downstream systems
- Create curated datasets and dashboards to support analytics, reporting, and go-to-market initiatives
- Support AI/ML and agent-based applications by preparing datasets for MCP integrations and AI-driven applications
- Monitor pipeline performance, implement data quality checks, validations, and alerting mechanisms
- Collaborate across teams to define data contracts and ensure system consistency
Requirements:
- 6+ years of experience in data engineering or related field
- Strong hands-on experience with Databricks, AWS (S3, Glue, Athena, EMR, etc.), and Kafka
- Proficiency in Python (PySpark) and SQL for large-scale data processing
- Experience building and maintaining ETL/ELT pipelines (DBT/Airflow or similar)
- Experience with data ingestion tools such as Fivetran or similar
- Familiarity with reverse ETL/data activation workflows and syncing to Salesforce, HubSpot, Braze
- Exposure to or experience with AI/ML data pipelines, including RAG architectures, vector databases, or embeddings workflows
- Familiarity with agent-based systems, MCP integrations, or LLM-powered applications (strong plus)
- Experience working with Finance and building finance-specific metrics and pipelines (strong plus)
- Understanding of data modeling and working with large-scale datasets (batch and streaming)
- Experience creating dashboards and supporting reporting workflows with BI tools
- Strong problem-solving skills and ability to debug production data issues
- Strong communication and cross-team collaboration skills