> Markdown version of [/jobs/ext/2878337-sr-engineer-data-infrastructure](https://www.wearedevelopers.com/jobs/ext/2878337-sr-engineer-data-infrastructure). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr. Engineer, Data Infrastructure - **Company:** Amazon.com, Inc. - **Location:** Seattle, WA, United States - **Experience:** Expert - **Salary:** $140,000.0 - $200,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Automated Storage and Retrieval Systems, Law Practice Management Software, Cloud Computing, Data Deduplication, Data Infrastructure, Data Stores, Data Systems, Distributed Data Store, NoSQL, Unstructured Data, Workflow Management Systems, Large Language Models, Indexer, Production Code, Data Pipelines - **Published:** September 13, 2026 - **Apply:** https://www.careerjet.com/job/us3da381b1e8d24c624d0e988c9c0cb45c/eaa ## About the Role Demonstrated experience building systems that ingest, transform, index, store, and retrieve large volumes of unstructured data at scale. Hands-on production coding experience as an active engineering contributor. Experience owning infrastructure end-to-end, including deployment, observability, and on-call or incident response. Familiarity with large asynchronous processing pipelines. Experience with large relational and/or NoSQL data stores. Ability to operate with full ownership of downstream infrastructure responsibilities without relying on separate platform, SRE, or DBA teams. Nice to Have: Background in eDiscovery, legal technology platforms, or document-centric data systems, including deduplication, threading, family relationships, metadata extraction, or document enrichment. Experience with search, indexing, crawling, or retrieval systems, including web-scale crawling. Familiarity with durable execution and workflow orchestration frameworks. Experience with AI or LLM processing pipelines. Experience with data modeling for complex, heterogeneous document collections. ## Description The Senior Engineer, Data Infrastructure will own a major vertical of distributed data infrastructure that processes hundreds of millions of tokens per minute and handles significant variance in workloads across a shared, multi-tenant environment. This is a full-ownership role covering design, implementation, deployment, observability, and incident response. A core challenge of this position is building fair, isolated, and predictable systems under contention, where individual customer workloads can vary by 1,000x or more. The successful candidate will be a hands-on engineer who writes production code and operates the systems they build., Own, operate, and continuously improve a designated vertical of distributed data infrastructure end-to-end. Design and implement systems for the ingestion, transformation, storage, and retrieval of large volumes of structured and unstructured data. Ensure workload isolation and fairness across a multi-tenant environment where customer workloads vary significantly in size and complexity. Maintain infrastructure definitions, deployment configuration, production promotion, and observability tooling as part of the engineering role. Identify and remove bottlenecks across ingestion, storage, retrieval, orchestration, and AI-processing pipelines. Own incident response for systems within the assigned scope, including clear failure mode documentation and reliable recovery paths. Establish technical patterns and provide working implementations that elevate the broader engineering team's architectural decision-making., Do you enjoy solving complex problems and driving influential changes? Are you curious about the systems used to run the largest cloud computing infrastructures in the world? Do yo… + 1 day ago + ## Related Videos - [Fault Tolerance and Consistency at Scale: Harnessing the Power of Distributed SQL Databases](https://www.wearedevelopers.com/videos/1520-fault-tolerance-and-consistency-at-scale-harnessing-the-power-of-distributed-sql-databases) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Optimizing Discovery: PostgreSQL's Role in Transforming GetYourGuide's Search](https://www.wearedevelopers.com/videos/1647-optimizing-discovery-postgresql-s-role-in-transforming-getyourguide-s-search) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [NoSQL Data Modeling for Front-end Developers](https://www.wearedevelopers.com/videos/297-nosql-data-modeling-for-front-end-developers) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer)