Sr. Data Engineer
dentsuAbout the role
Job Description:
We’re looking for an Identity Data Engineer who is passionate about data quality, intellectually curious about how real-world identities get resolved, and ready to get deep into the details.
You'll work directly with PII-class data at a low level — examining records, interrogating match logic, and developing a genuine understanding of why our matching engines make the decisions they do. Our matching engines link consumer and household identity signals across diverse data sources, combining deterministic logic with increasingly AI-assisted probabilistic resolution. You'll help enhance these engines — improving match rates, reducing false positives, and extending asset coverage. As our AI-augmented matching capabilities grow, so will this role. There is a real long-term track here for an engineer who wants to go deep on identity.
What You'll Do
Identity Data Engineering
Design, build, and maintain Snowflake-based pipelines that produce and refresh our core consumer and household identity assets on a regular cadence.
Write complex SQL and Python to transform, deduplicate, and enrich identity data at scale — including direct work with PII fields such as names, addresses, emails, and phone numbers.
Investigate data anomalies and quality issues at a record level, tracing match decisions back to source signals and surfacing root causes.
Build and maintain data models that represent consumer and household identity linkage across multiple input sources.
Matching Engine Enhancement
Partner with senior engineers and data scientists to enhance our AI-assisted matching engine — contributing to feature design, scoring logic, model evaluation, and threshold tuning.
Implement and test matching algorithm improvements — both AI-driven and rule-based — and measure their real impact on precision, recall, and overall asset quality.
Build evaluation tooling: ground-truth comparisons, match quality dashboards, and regression detection across engine versions.
Help drive the evolution of our matching pipeline toward more intelligent, AI-augmented identity resolution, actively using AI tools as part of your day-to-day engineering workflow.
Collaboration & Delivery
Work cross-functionally with Data Science, Product, and downstream engineering teams to translate identity requirements into reliable, scalable solutions.
Participate in code reviews and architectural discussions; apply engineering best practices across the full delivery lifecycle — design, implement, test, and deploy via CI/CD.
Document data models, pipeline logic, and algorithm decisions clearly for both technical and non-technical audiences.
Support QA processes and on-call responsibilities for production identity asset pipelines.
Build automated validation frameworks and quality tracking pipelines that continuously monitor asset health — including data completeness, match consistency, and anomaly detection — and surface results through clear, actionable reporting.
What You Bring
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s