Principal Data Lead
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+21 more
Job description
The Principal Data Lead will lead data strategy execution, data supply chain operations, metadata maturity, governed data asset promotion governance, data quality standards, and data-domain coordination for a large-scale federal data and analytics modernization program. The program supports governed data assets, reusable analytics, dashboards, APIs, public-facing reporting, and AI-enabled services.
This role will improve source onboarding, data promotion, metadata maturity, data quality, lineage, stewardship, and data-domain coordination. The Principal Data Lead will ensure data assets are discoverable, governed, documented, quality-controlled, and suitable for self-service analytics and responsible AI expansion.
Responsibilities
- Lead data strategy execution, data supply chain and metadata maturity leadership, governed data asset promotion governance, data quality standards, and data-domain coordination across source systems and data owners.
- Oversee applicable coverage areas, including data platform operations, distributed processing, programming and query languages, jobs and workflows, schema design, optimization, pipeline reliability, governance layers, metadata maturity, machine-readable documentation, lineage, governed data asset promotion, analytics discoverability, and domain stakeholder engagement.
- Operate and improve source onboarding, ingestion, transformation, data quality, platform operations, compute governance, connectors, endpoints, workspace administration, and documentation for reusable data assets.
- Advance the program’s data trust model by defining and applying admission, promotion, ownership, stewardship, lineage, refresh, trust indicator, documentation, and retirement standards for governed data assets.
- Establish and monitor data quality and processing timeliness practices, including completeness, conformance, quality pass rates, defect trends, freshness against service levels, rework drivers, schema drift events, and defect remediation.
- Coordinate with data owners, source-system teams, product teams, public-facing dashboard teams, BI teams, AI teams, and approved consumers to improve data contracts, ingestion readiness, metadata standards, and downstream reuse.
- Ensure metadata maturity advances as a tracked, multi-year effort and that AI use cases do not scale beyond appropriate users until supporting metadata quality, lineage, documentation, and governance are sufficient for reliable outputs.
- Support governed self-service analytics, certified dashboards, public-facing products, secure data sharing, reusable APIs, modernization of legacy analytics workflows, and publication workflows through governed data assets, documentation, and data governance standards., eSimplicity supports a remote work environment operating within the Eastern time zone so we can work with and respond to our government clients. Expected hours are 9:00 AM to 5:00 PM Eastern unless otherwise directed by manager.
Occasional travel for training and project meetings. It is estimated to be less than 5% per year.
Requirements
- Bachelor’s degree in Computer Science, Information Systems, Engineering, Math, or other related scientific or technical discipline. With twelve years of general information technology
- 14+ year’s experience in software engineering, including implementing engineering best practices, iterative/continuous engineering principles
- All candidates must pass public trust clearance through the U.S. Federal Government. This requires candidates to either be U.S. citizens or pass clearance through the Foreign National Government System which will require that candidates have lived within the United States for at least 3 out of the previous 5 years, have a valid and non-expired passport from their country of birth and appropriate VISA/work permit documentation
- Demonstrated ability to lead enterprise data strategy, data engineering, data governance, metadata, data quality, or data platform operations in a complex analytics environment.
- Experience with lakehouse or comparable data platform operations, including ingestion, transformation, schema design, optimization, jobs and workflows, pipeline reliability, data quality, documentation, and production support.
- Working knowledge of distributed processing, programming and query languages, governance layers, metadata management, lineage, data cataloging, and governed data asset promotion.
- Ability to define and manage data quality standards, data promotion criteria, source onboarding patterns, data contracts, stewardship practices, data-domain coordination, and reusable data product documentation.
- Experience collaborating with product, engineering, BI, AI/ML, security, privacy, support, and business stakeholders to support self-service analytics, certified dashboards, public-facing products, APIs, and responsible AI readiness.
- Knowledge of sensitive Government data handling, approved data-use practices, least-privilege access, privacy-aware data publication, public data controls, cell suppression, and Section 508/WCAG considerations for public-facing data products.
- Ability to comply with customer-specific security, privacy, accessibility, quality, training, and data-handling requirements for assigned systems and data.
Preferred qualifications
- Experience supporting federal, public sector, healthcare, or other regulated data, analytics, oversight, reporting, or public transparency programs.
- Experience with modern cloud platforms, lakehouse or comparable data platforms, governance layers, secure data-sharing patterns, query engines, connectors, APIs, BI tools, workflow tools, source control, CI/CD, work management tools, documentation tools, observability tools, and approved monitoring patterns.
- Experience advancing metadata maturity, machine-readable documentation, data stewardship, catalog discoverability, semantic consistency, dashboard certification, public-facing publication packages, and governed data assets.
- Familiarity with AI/ML data readiness, model-serving data dependencies, retrieval-augmented generation patterns, AI service governance, inference logging, AI evaluation artifacts, and metadata prerequisites for responsible AI expansion.
- Hands-on experience building and optimizing data pipelines within the Databricks platform
- Preferred certifications may include AWS, Databricks data platform, cloud data analytics, Certified Data Management Professional, DAMA, data governance, data quality, analytics engineering, SAFe, or related data platform credentials.
Benefits & conditions
United States Remote $188,700 - $200,000 a year - Full-time, Pulled from the full job description
- 401(k)
- Health insurance
- Paid time off
- Vision insurance
- Dental insurance
- Disability insurance
- Paid holidays, eSimplicity offers a comprehensive benefits package, including medical, dental, and vision coverage, 401(k) retirement benefits, paid time off, paid holidays, life and disability insurance, and additional wellness and employee support programs. Eligibility may vary based on employment status and applicable plan terms.
About the company
eSimplicity is a modern digital services company that partners with government agencies to improve the lives and protect the well-being of all Americans, from veterans and service members to children, families, and seniors. Our engineers, designers, and strategists cut through complexity to create intuitive products and services that equip federal agencies with solutions to courageously transform today for a better tomorrow.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.indeed.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Making Data Warehouses Fast: A Developer’s Story
Data Engineer Salary UK
Highest Paying Tech Companies for Developers
Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production