> Markdown version of [/jobs/ext/2027117-data-science-engineer](https://www.wearedevelopers.com/jobs/ext/2027117-data-science-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Science Engineer - **Company:** Adobe Systems - **Location:** San Jose, CA, United States - **Experience:** Experienced - **Salary:** $138,600.0 - $200,700.0 - **Contract:** Permanent contract - **Skills:** Adobe Analytics, Adobe Acrobat, Adobe Creative Cloud, Adobe Experience Manager, Application Programming Interfaces (APIs), Artificial Intelligence, Airflow, Amazon Web Services, Amazon S3, Data Analysis, Apache HTTP Server, Confluence, JIRA, Microsoft Azure, Cloud Computing, Cloud Storage, Collaborative Software, Continuous Delivery, Continuous Integration, Information Engineering, Data Governance, Data Transformation, Distributed Data Store, Fault Tolerance, Github, Graph Database, Apache Hadoop, Apache Hive, Python (Programming Language), Machine Learning, NumPy, Open Source Technology, Cloud Services, Standard Sql, DataOps, Simple Data Format, SQL Databases, Data Streaming, Parquet, Data Import/Export, Scripting, Data Ingestion, Large Language Models, Prompt Engineering, Apache Spark, Electronic Medical Records, Pandas, Adobe, Pyspark, Collibra, Apache Kafka, Operational Systems, Data Management, Presto, Multiplatform, Data Pipelines, Amazon Elastic Mapreduce (EMR), Jenkins, Databricks, Programming Languages - **Published:** August 11, 2026 - **Apply:** https://www.careerbuilder.com/job-details/data-science-engineer-san-jose-ca--9f52b245-b4b7-4341-bda5-5fcef063a3eb ## About the Role * Master's degree or equivalent experience is preferred. * 8+ years of consistent track record as a data engineer. * 5+ years validated ability in distributed data technologies e.g., Hadoop, Hive, Presto, Spark etc. * 3+ years of experience with Cloud based technologies - Databricks, S3, Azure Blob Storage, Notebooks, AWS EMR, Athena, Glue etc. Familiarity and usage of different file formats in batch/streaming processing i.e., Delta/Parquet/ORC etc. * 2+ years' experience with streaming data ingestion and transformation using Kafka, Kinesis etc. * Outstanding SQL experience. Ability to write optimized SQLs across platforms. * Proven hands - on experience in Python/PySpark/Scala and ability to manipulate data using Pandas, NumPy, Koalas etc. and using APIs to transfer data. * Experience with CI/CD tools i.e., GitHub, Jenkins etc. * Working experience with Open- source orchestration tools i.e., Apache Air Flow/ Azkaban etc. * Teammate with excellent communication/teamwork skills when it comes to closely working with data scientists and machine learning engineers daily. Nice to have * Showcase your work if you are an open - source contributor. Passion to contribute to Open-source community is highly valued. * Experience with Data Governance tools e.g., Collibra and Collaboration tools e.g., JIRA/ Confluence etc. * Familiarity with Adobe solutions such as Adobe Experience Platform, Adobe Analytics, Customer Journey Analytics, and Adobe Journey Optimizer is a plus. * Experience with LLM Models/ Agentic workflows using Copilot, Claude, LLAMA, Databricks Genie etc. is highly preferred. Skills in building context and prompt engineering solutions including classical RAG, Knowledge graph, MCPs, Agentic frameworks like n8n, etc. are highly desirable., Adobe Acrobat, Adobe Product Family, Amazon Simple Storage Service (S3), Amazon Web Services (AWS), Analysis Skills, Apache, Apache Hadoop, Apache Hive, Apache Spark, Application Programming Interface (API), Artificial Intelligence (AI), Atlassian JIRA, Best Practices, Cloud Computing, Communication Skills, Compensation and Benefits, Continuous Deployment/Delivery, Continuous Integration, Cross-Functional, Customer Experience, Customer Relations, Data Analysis, Data Management, Data Science, Electronic Medical Records, GitHub, Global Branding, Jenkins, Machine Learning, Microsoft Windows Azure, Multiplatform/Cross-Platform, Open Source, Python Programming/Scripting Language, SQL (Structured Query Language), Scala Programming Language, Team Player, Writing Skills ## Description As a member of the Data Engineering team, you will have significant responsibility to help build large scale cloud-based data and analytics platform with enterprise-wide consumers. This role is inherently multi-functional, and the ideal candidate will work across teams. The position requires ability to own things, come up with innovative solutions, try new tools and technologies. What you will do * Build fault tolerant, scalable, quality data pipelines using multiple cloud- based tools. * Build analytical, personalization capabilities using modern and brand new technologies employing Adobe tools like AEP, AJO and CJA. * Build LLM agents to optimize and automate data pipelines following best engineering practices. * Deliver End to End Data Pipelines to run Machine Learning Models in a production platform. * Innovative solutions to help broader organization take significant actions fast and efficiently. * Chip in to data engineering and data science frameworks, tools, and processes. * Implement outstanding data operations and implement standard methodologies to use resources in an optimum way. * Architect data ingestion, data transformation, data consumption, data governance frameworks. * Help build production grade ML models and integration with operational systems. * Work in a collaborative environment and contribute to the team as well as organization's success. ## Related Videos - [Improving quality with Agentic AI with Rovo Dev and Xray](https://www.wearedevelopers.com/videos/2005-improving-quality-with-agentic-ai-with-rovo-dev-and-xray) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [3x Performance: A Humbling Journey](https://www.wearedevelopers.com/videos/100165-3x-performance-a-humbling-journey) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Collaboration Quantified: Lessons from Open Source Developer Networks](https://www.wearedevelopers.com/videos/1422-collaboration-quantified-lessons-from-open-source-developer-networks) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)