> Markdown version of [/jobs/ext/2172206-principal-data-architect](https://www.wearedevelopers.com/jobs/ext/2172206-principal-data-architect). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Principal Data Architect - **Company:** Allen Institute for Brain Science - **Location:** Seattle, WA, United States (Remote available) - **Experience:** Expert - **Salary:** $167,850.0 - $209,750.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Big Data, Cloud Storage, Data Governance, Data Infrastructure, Extract Transform Load (ETL), Data Retention, Information Lifecycle Management, Information Management, Information Sciences, Meta-Data Management, Metadata Standards, Search Technologies, Data Storage Technologies, Large Language Models, Information Technology, Data Management - **Published:** August 21, 2026 - **Apply:** https://www.dice.com/job-detail/e951fbf7-c5ce-4e99-ad42-b096f8eb8c25 ## About the Role * Bachelor's Degree in Information and Library Science, Data Management, Computer Science, or a related field; or equivalent combination of degree and relevant experience * Minimum of 7 years of relevant experience in data governance, data stewardship, or large-scale data management environments * Deep expertise in metadata standards, data lifecycle management, and governance frameworks * Experience working across complex, multi-stakeholder environments, ideally in scientific or research settings * Strong ability to influence without authority and translate complex concepts across diverse audiences * Familiarity with scientific datasets, ideally in biology or biomedical research environments, including high resolution microscopy, genomics, physiology, and video * Deep knowledge of data governance frameworks and regulatory standards, including NIST and/or HIPAA compliance * Understanding of cloud storage environments, data infrastructure, or research computing ecosystems Preferred Education and Experience * Master's degree or Ph.D. in Information Science, Data Management, or a related field, with over 5 years of experience in large data-intensive environments, ideally in a scientific or academic research setting * Advanced skills in developing and applying data valuation frameworks to prioritize data retention and decommissioning, particularly within budget-constrained environments * A track record of success aligning data management practices to scientific and technical goals in a complex, multi-stakeholder environment * Skilled in navigating and influencing at the intersection of technical and scientific teams, with the ability to communicate complex data management principles clearly to diverse audiences * Familiarity with information management principles, such as metadata standards (e.g., FAIR data principles) and taxonomy development for data cataloging * Experience with archival standards, data preservation techniques, and regulatory frameworks related to scientific data * Ability to work collaboratively across scientific, technical, and operational teams * Strong organizational and communication skills, with the ability to translate technical and policy concepts across diverse audiences * Proven experience using relevant AI technologies to find and work with data (LLMs, vector search, MCPs, agentic search). Physical Demands * Fine motor movements in fingers/hands to operate computers and other office equipment ## Description * Establish and implement Institute-wide data and metadata standards to improve data discoverability, usability, and consistency across scientific and technology teams, influencing technical direction and decision-making across domains * Own and advance the Institute's data governance and stewardship program, including data cataloging, and lifecycle management practices, operating with a high degree of autonomy as a recognized subject matter expert * Develop guidance and best practices for data ETL, lifecycle management, including creation, storage, archiving, and retirement of datasets aligned with scientific mission * Drive adoption of data governance practices across programs, ensuring alignment and sustained use of standards and tools, with impact measured through uptake and effectiveness * Ensure data governance practices align with relevant NIST, HIPAA, and institutional data security standards where applicable. Work with technology, security, and legal teams to help define processes for managing sensitive or restricted datasets and implement appropriate safeguards and compliance practices * Lead the development and evolution of a central metadata standard for large-scale, multimodal biological data. Influence decisions related to data storage, infrastructure, and cost optimization, helping balance scientific priorities with resource constraints * Serve as a technical and strategic advisor to scientific, engineering, and operational leaders on data management practices, tradeoffs, and long-term sustainability * Lead high-impact, cross-functional initiatives to improve data accessibility, reuse , and long-term sustainability across the Institute *Note: Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions. This description reflects management's assignment of essential functions; it does not proscribe or restrict the tasks that may be assigned.* ## Related Videos - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Architecting the Future: Leveraging AI, Cloud, and Data for Business Success](https://www.wearedevelopers.com/videos/1096-architecting-the-future-leveraging-ai-cloud-and-data-for-business-success) - [Startup Presentation: StorX - Future of Cloud Storage](https://www.wearedevelopers.com/videos/1176-startup-presentation-storx-future-of-cloud-storage) - [Data: The Deciding Factor in AI Success](https://www.wearedevelopers.com/videos/100310-data-the-deciding-factor-in-ai-success) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Blueprints for Success: Steering a Global Data & AI Architecture](https://www.wearedevelopers.com/videos/1577-blueprints-for-success-steering-a-global-data-ai-architecture) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers)