> Markdown version of [/jobs/ext/2270419-senior-machine-learning-engineer-applied-science-data-frameworks](https://www.wearedevelopers.com/jobs/ext/2270419-senior-machine-learning-engineer-applied-science-data-frameworks). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Machine Learning Engineer, Applied Science Data Frameworks - **Company:** Adobe Systems - **Location:** San Jose, CA, United States - **Experience:** Expert - **Salary:** $183,300.0 - $265,350.0 - **Contract:** Permanent contract - **Skills:** Training Data, Amazon Web Services, Apache HTTP Server, Automation of Tests, Microsoft Azure, Big Data, Cloud Computing, Computer Clusters, Continuous Integration, Information Engineering, Data Infrastructure, Extract Transform Load (ETL), Data Structures, Distributed Data Store, Distributed Systems, Fault Tolerance, Python (Programming Language), Machine Learning, Operational Databases, Queue Management Systems, Tensorflow, Search Technologies, SQL Databases, Data Processing, Data Ingestion, Pytorch, Apache Spark, Adobe, Containerization, Information Technology, Deployment Automation, Dask, Integration Frameworks, Data Management, Machine Learning Operations, Feature Extraction, Data Pipelines, Docker, Databricks - **Published:** August 27, 2026 - **Apply:** https://adobe.wd5.myworkdayjobs.com/external_experienced/job/San-Jose/Senior-Machine-Learning-Engineer--Applied-Science-Data-Frameworks_R171469 ## About the Role * 5-6 years of professional experience building and operating distributed systems or data infrastructure in production environments. * Solid understanding of distributed computing concepts and experience with framework like Apache Spark, Ray, Dask, or equivalent. * Familiarity with cloud platforms (AWS or Azure) and data platforms such as Databricks or Spark. * Proficiency in Python and strong software engineering fundamentals - system design, data structures, algorithms. * Familiarity with ML frameworks such as PyTorch or TensorFlow; hands-on ML experience is a plus but not required. * Basic familiarity with MLOps practices including CI/CD pipelines, containerization (Docker), and deployment automation. * Familiarity with batch inference architectures and large-scale data processing patterns is a plus. * Bachelor's degree in Computer Science, Engineering, or a related field; MS is a plus. * Strong communication skills and ability to collaborate across engineering and research teams. About Adobe ## Description We are looking for a Senior Machine Learning Engineer to join our Applied Science Data Frameworks team responsible for building the foundational infrastructure that powers large-scale multimodal AI training and inference. This role is ideal for someone with strong distributed systems and data engineering fundamentals who is eager to work in an ML-adjacent environment, contributing to training data loaders, distributed inference frameworks, feature enrichment pipelines, and dataset management systems that enable ML teams to train foundation models at petabyte scale. You'll work on high-impact projects involving distributed data loading for PyTorch training workloads, batch inference pipelines for feature enrichment, semantic search infrastructure for dataset discovery, and production-grade ML data pipelines that support generative AI model development. Your systems will process billions of images, * Contribute to building and maintaining distributed training data loaders that handle multi-source data ingestion, temporal sampling, and real-time transformations for large-scale model training workflows. * Help implement and maintain feature enrichment pipelines and dataset registry systems that support multimodal model training across images, video, documents, and text. * Build and maintain batch inference pipelines for large-scale feature extraction, * processing assets through distributed GPU clusters with queue management and fault tolerance. * Develop data processing systems using frameworks like Apache Ray, Spark, DuckDB, or similar distributed computing tools for SQL-based data ingestion and Apache Arrow-based storage formats. * Support semantic search capabilities and vector database infrastructure (e.g., * OpenSearch, LanceDB) for dataset discovery and embedding-based retrieval. * Contribute to CI/CD infrastructure for ML systems including self-hosted runner * management, Docker image builds, automated testing pipelines, and deployment automation. * Collaborate with ML research teams to translate training requirements into reliable, scalable data loading and preprocessing infrastructure. * Write reusable framework components, SDKs, and documentation to help accelerate platform adoption across modeling teams. * Optimize data pipeline performance across dimensions like startup latency, throughput, memory footprint, and GPU utilization. * Contribute to observability and reliability standards for production data systems * supporting 24/7 training workloads. ## Related Videos - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [3x Performance: A Humbling Journey](https://www.wearedevelopers.com/videos/100165-3x-performance-a-humbling-journey) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) - [NoCode LiveCode: Leveraging AI Tools to Craft Fully Functional Apps!](https://www.wearedevelopers.com/videos/1169-nocode-livecode-leveraging-ai-tools-to-craft-fully-functional-apps) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLops – Deploying, Maintaining And Evolving Machine Learning Models in Production](https://www.wearedevelopers.com/magazine/115-mlops-deploying-maintaining-and-evolving-machine-learning-models-in-production) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)