SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 230,000 - 340,000 / annual
Baselayer is rebuilding the identity and verification layer for US businesses. The company fuses public records, IRS data, sanctions lists, web signals, and fraud telemetry from 2,200+ financial institutions into a single business graph that resolves entities in milliseconds with 98% match rates. Trusted by over 20% of US financial institutions and expanding into gig platforms, marketplaces, and AI companies.
You will own product features across the identity graph surface end-to-end: schema design, data ingestion, entity resolution, API endpoints, scope and multi-tenancy, and performance optimization. The role spans the full stack from ingestion pipelines through customer-facing APIs.
Key responsibilities:
- Ingest and normalize heterogeneous public records into the graph, owning the full pipeline from ingestion to normalizer to repository to API surface
- Build and maintain entity resolution and linking systems, matching the same business or person across sources with mismatched keys, misspelled names, and differently formatted addresses
- Design and ship customer-facing APIs for search, lookup, monitoring, and webhook delivery
- Own schema design and zero-downtime migrations on a database running 24/7
- Optimize performance and scale for 10x volume, ensuring ingestion pipelines complete before the next batch and API latencies remain stable as the graph deepens
- Build multi-tenancy and access control with every query scoped by org/user/permission at the data layer
The team is small, focused on real-time entity resolution at scale, and operates with minimal process between idea and shipping. The graph is built and match rates are proven; the hardest problems ahead include graph embeddings, fraud propagation models, sub-100ms latency traversal, and expanding beyond finance.
Minimum Requirements:
- Shipped product features touching ingestion, storage, API, and customer in production end-to-end
- Owned a Postgres schema that real money or real decisions depended on, managing its evolution
- Strong async Python with experience running async services at scale
- Deep Postgres expertise: schema design under live load, EXPLAIN ANALYZE proficiency, zero-downtime migration tooling
- FastAPI (or equivalent): shipped real APIs with auth dependencies, OpenAPI contracts, and structured response models
- Built multi-tenant APIs where access control is load-bearing and scoped correctly at the data layer
What Sets You Apart:
- Entity resolution, record linkage, or fuzzy matching at scale
- Search infrastructure (full-text, fuzzy, or vector) kept in sync with a Postgres system of record
- KYC/KYB/fraud/underwriting data-pipeline experience
- Address parsing/normalization at country scale
- Webhook delivery infrastructure: retries, signing, idempotency, and ordering guarantees
- GCP experience with Cloud Run, Cloud Tasks, Cloud SQL, and batch pipelines