> Markdown version of [/jobs/ext/2106181-principal-data-engineer](https://www.wearedevelopers.com/jobs/ext/2106181-principal-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Principal Data Engineer - **Company:** Sovrn, Inc. - **Location:** Boulder, CO, United States (Remote available) - **Experience:** Expert - **Salary:** $200,000.0 - $240,000.0 - **Contract:** Temporary contract - **Skills:** Artificial Intelligence, Amazon Web Services, Big Data, Code Review, Data as a Services, Data Architecture, Information Engineering, Data Governance, Data Infrastructure, Data Security, Software Design Patterns, DevOps, Distributed Computing Environment, Python (Programming Language), Meta-Data Management, Site Reliability Engineering Practices, Software Engineering, Data Streaming, Alwayson, Large Language Models, Snowflake, Prompt Engineering, Backend, Data Lakes, Pyspark, Semi-structured Data, Data Lineage, Apache Kafka, Server Operating Systems & Platforms, Databricks - **Published:** August 18, 2026 - **Apply:** https://job-boards.greenhouse.io/sovrn/jobs/7872983 ## About the Role * 10+ years of software engineering experience, with a strong data engineering and backend track record * 5+ years working specifically in adtech data infrastructure, SSP, DSP, exchange, or ad server environments * Deep fluency in the programmatic ecosystem: OpenRTB, bid request/response flows, auction mechanics, supply path optimization, or similar * Excellent understanding of real-time streaming and batch pipelines, big data, and data lakes; hands-on experience in distributed data processing in the AWS ecosystem * Strong understanding of second-layer big data platforms such as Snowflake and Databricks, applicable use cases, best practices, implementation, and support considerations * Strong experience in structured, unstructured, and semi-structured data techniques; metadata management, data lineage, and data governance * Experience with data security and compliance (PII, CCPA, GDPR, etc.) * Demonstrated experience leading AI or agentic engineering efforts in production environments; not just experimentation, but shipped, operated, and iterated on * Hands-on experience with LLM integration patterns: RAG, vector DBs, tool use, multi-step agentic workflows, prompt engineering, and evaluation frameworks * Ability to clearly document and communicate architectural concepts at multiple levels; ability to lead problem definition, solution designs, and implementation work plans * Understanding of DevOps and SRE practices * Comfort driving technical decisions in ambiguous, fast-moving environments * A point of view on where AI is and isn't the right too, and the credibility to make that case Location: Location: Sovrn offices are located in Boulder, CO and New York, NY. We believe collaboration and shared purpose lead to deeper relationships, more creative problem-solving, and more learning, faster. For employees located near one of our offices, working in-person is our default. We trust people to use personal agency when focused work or life circumstances call for flexibility.. For exceptional candidates we would consider remote locations in these states: AZ, CA, CO, GA, MD, MA, MI, MO, NJ, NY, NC, OK, TX, UT, WA. #LI-Remote, #LI-Hybrid We understand that no candidate is perfectly qualified for any job. Experience comes in different forms; many skills are transferable; and passion goes a long way. Even more important than your resume is a clear demonstration of accountability and the ability to thrive in a fluid and collaborative environment. We expect you to learn new things in this role and encourage you to apply if your experience is close to what we're looking for. ## Description We're looking for a Principal Software Engineer (Data) with deep roots in adtech data infrastructure and a genuine conviction about what AI-native data engineering looks like in practice. This is a specialized principal-level engineering role - one that carries all the architectural ownership and technical leadership expectations of a Principal Software Engineer, focused on Sovrn's Data Collective. From a generative/agentic AI capabilities standpoint, we already use LLMs and agentic tooling across our data stack and, we're looking for is someone who can help us take that from general adoption to intentional practice - who has strong opinions about where AI creates real leverage in a high-throughput adtech environment, and who can bring the rest of the engineering organization along with them. Languages/components/tools in our stack: Python, Pyspark, Kafka, Databricks, AWS What you'll be doing: Data Platform & Architecture * Own the design and evolution of data platform systems that operate at exchange scale; high throughput, real-time streaming, and always-on batch pipelines * Lead architectural decisions across data infrastructure: pipeline design, data modeling, lakehouse architecture, and data services layers * Specify data platform components and configurations required for pipeline implementation; define pipeline observability to understand and improve performance at massive scale * Research, implement, and evolve methods to process and democratize data across the organization * Drive technical standards, design reviews, and engineering best practices across a senior team * Partner with product, data science, and platform teams to ship end-to-end AI & Agentic Engineering Leadership * Establish and champion AI engineering practices across the team, from prompt engineering and RAG patterns to agentic workflow design, LLM evaluation, and progressive implementation of agentic design patterns * Identify high-leverage opportunities to apply AI in our data stack: intelligent pipeline optimization, anomaly detection, automated data quality, forecasting, and LLM-powered data services * Lead the evolution of our existing LLM and agentic tooling from passive use to intentional, well-architected integration within our data platform * Set standards for how we evaluate, trust, and operate AI-powered systems in production, including observability, fallback behavior, and model governance * Help the broader engineering team build fluency and confidence with AI tooling, not just tolerance of it Collaboration & Mentorship * Provide domain expertise across the organization to enable business growth through data services and data models * Provide counsel to all consumers and stakeholders of data to enable efficient and impactful use of our data assets * Mentor and level up engineers through code review, design collaboration, and hands-on guidance; foster a culture of innovation and continuous learning * Operate with high autonomy across ambiguous, high-impact problems, Will you now or in the future require sponsorship for employment visa status (e.g., H-1B visa status)?* Select... Are you currently working in the US on a temporary visa?* Select... Do you have adtech experience (SSP, DSP, exchange, or ad server environments)? * Select... Describe your experience in operating at the forefront of applied AI, especially in exploring, prototyping, evaluating, and deploying cutting-edge generative AI technologies.* ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [The Best Job Search Websites of 2025](https://www.wearedevelopers.com/magazine/368-the-best-job-search-websites-of-2025) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift)