SlipstreamJobsFresh Startup & VC-Backed Jobs

Senior Software Engineer, Identity Graph

Baselayer - San Francisco, CA, United States - Hybrid - posted 2026-10-02

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Salary: USD 230,000 - 340,000 / annual

Baselayer is rebuilding the identity and verification layer for US businesses. The company fuses public records, IRS data, sanctions lists, web signals, and fraud telemetry from 2,200+ financial institutions into a single business graph that resolves entities in milliseconds with 98% match rates. Trusted by over 20% of US financial institutions and expanding into gig platforms, marketplaces, and AI companies. You will own product features across the identity graph surface end-to-end: schema design, data ingestion, entity resolution, API endpoints, scope and multi-tenancy, and performance optimization. The role spans the full stack from ingestion pipelines through customer-facing APIs. Key responsibilities: - Ingest and normalize heterogeneous public records into the graph, owning the full pipeline from ingestion to normalizer to repository to API surface - Build and maintain entity resolution and linking systems, matching the same business or person across sources with mismatched keys, misspelled names, and differently formatted addresses - Design and ship customer-facing APIs for search, lookup, monitoring, and webhook delivery - Own schema design and zero-downtime migrations on a database running 24/7 - Optimize performance and scale for 10x volume, ensuring ingestion pipelines complete before the next batch and API latencies remain stable as the graph deepens - Build multi-tenancy and access control with every query scoped by org/user/permission at the data layer The team is small, focused on real-time entity resolution at scale, and operates with minimal process between idea and shipping. The graph is built and match rates are proven; the hardest problems ahead include graph embeddings, fraud propagation models, sub-100ms latency traversal, and expanding beyond finance. Minimum Requirements: - Shipped product features touching ingestion, storage, API, and customer in production end-to-end - Owned a Postgres schema that real money or real decisions depended on, managing its evolution - Strong async Python with experience running async services at scale - Deep Postgres expertise: schema design under live load, EXPLAIN ANALYZE proficiency, zero-downtime migration tooling - FastAPI (or equivalent): shipped real APIs with auth dependencies, OpenAPI contracts, and structured response models - Built multi-tenant APIs where access control is load-bearing and scoped correctly at the data layer What Sets You Apart: - Entity resolution, record linkage, or fuzzy matching at scale - Search infrastructure (full-text, fuzzy, or vector) kept in sync with a Postgres system of record - KYC/KYB/fraud/underwriting data-pipeline experience - Address parsing/normalization at country scale - Webhook delivery infrastructure: retries, signing, idempotency, and ordering guarantees - GCP experience with Cloud Run, Cloud Tasks, Cloud SQL, and batch pipelines

Similar roles