DA

Staff, Backend Engineer - Catalog

DataHub
Palo Alto, USAfull_timePosted 15 Jul 2026

About the role

<div class="content-intro"><p>DataHub is an AI & Data Context Platform adopted by over 3,000 enterprises, including Apple, CVS Health, Netflix, and Visa. Innovated jointly with a thriving open-source community of 13,000+ members, DataHub's metadata graph provides in-depth context of AI and data assets with best-in-class scalability and extensibility.</p> <p>The company's enterprise SaaS offering, DataHub Cloud, delivers a fully managed solution with AI-powered discovery, observability, and governance capabilities. Organizations rely on DataHub solutions to accelerate time-to-value from their data investments, ensure AI system reliability, and implement unified governance, enabling AI & data to work together and bring order to data chaos.</p></div><h2><strong>The Challenge</strong></h2> <p>As AI and data products become business-critical, enterprises face a metadata crisis:</p> <ul> <li>No unified way to track the complex data supply chain feeding AI systems</li> <li>Engineering teams struggling with data discovery, lineage, and governance</li> <li>Organizations needing machine-scale metadata management, not just human-browsable catalogs</li> </ul> <h2><strong>Why This Matters</strong></h2> <p><strong>This is where infrastructure meets impact.</strong> The metadata layer you'll build will directly power the next generation of AI systems at massive scale. Your code will determine how safely and effectively thousands of organizations deploy AI, affecting millions of users worldwide.</p> <h2><strong>The Role</strong></h2> <p>We're looking for an exceptional Staff, Backend engineer to lead development of DataHub's Platform framework – the core that connects diverse data systems and powers our metadata collection capabilities.</p> <h2><strong>You'll Build</strong></h2> <ul> <li>Scalable, fault-tolerant ingestion systems for enterprise-scale metadata</li> <li>Clean, intuitive APIs for our connector ecosystem</li> <li>Event-driven architectures for real-time metadata processing</li> <li>Schema mapping between diverse systems and DataHub's unified model</li> <li>Versioning systems for AI assets (training data, model weights, embeddings)</li> </ul> <h2><strong>You Have</strong></h2> <ul> <li class="my-2 [&+p]:mt-4 [&_strong:has(+br)]:inline-block [&_strong:has(+br)]:pb-2">8+ years building production-grade distributed systems</li> <li class="my-2 [&+p]:mt-4 [&_strong:has(+br)]:inline-block [&_strong:has(+br)]:pb-2">Advanced Python and API design expertise</li> <li class="my-2 [&amp

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Generate Application Kit

Free account required — sign up in 30s

Company

DataHub

View all open roles →