> Markdown version of [/jobs/ext/1173099-ai-ml-engineer-2](https://www.wearedevelopers.com/jobs/ext/1173099-ai-ml-engineer-2). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # AI/ML Engineer 2 - **Company:** Day & Zimmermann - **Location:** Philadelphia, PA, United States - **Experience:** Experienced - **Salary:** $101,840.0 - $165,490.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Data Analysis, Microsoft Azure, Cloud Computing, Program Optimization, Code Review, Continuous Integration, Information Engineering, Data Mining, Data Visualization, Cursor (Graphical User Interface Elements), Database Queries, Statistical Hypothesis Testing, Python (Programming Language), Knowledge Management, Machine Learning, Open Source Technology, Tensorflow, Software Engineering, Unstructured Data, User-Centered Design, Web Application Frameworks, Data Logging, Google Cloud, Feature Engineering, Chatbots, Pytorch, Large Language Models, Prompt Engineering, Git, Containerization, Git Flow, Scikit Learn, Information Technology, Code Testing, Xgboost, Machine Learning Operations, Hardware Infrastructure, Virtual Agents, Api Design, Data Pipelines - **Published:** July 3, 2026 - **Apply:** https://diversityjobs.com/main/sendform/8/8/28176/1/16994720?backUrl=%2Fcareer%2F16994720%2FAi-Ml-Engineer-2-Pennsylvania-Philadelphia ## About the Role * Deep hands-on experience building production LLM applications with leading providers (Anthropic, OpenAI, Google, open-source models); expertise in prompt engineering, context engineering, retrieval-augmented generation (RAG) with vector databases, AI agent design, tool use, and Model Context Protocol (MCP) server development; experience authoring custom skills and building evaluation harnesses for LLM systems. * Proficiency in ML frameworks (scikit-learn, XGBoost, PyTorch, TensorFlow); strong foundation in statistical analysis, hypothesis testing, feature engineering, model optimization, deployment, and monitoring of predictive models across regression, classification, clustering, and time series; ability to design and validate experiments, define online/offline metrics, and apply quantitative methods to production problems. * Strong proficiency in Python as the primary language across ML and AI application development; comfort writing production-grade, maintainable, well-tested code. * Strong SQL skills with expertise in data extraction, cleaning, transformation, and pipeline development for structured and unstructured data; experience deploying ML/AI workloads across on-premises infrastructure and cloud platforms (AWS, Azure, GCP). * Proficient in Git, branching strategies, code review, and CI/CD pipelines; experience with containerization, API development, and observability (logging, tracing, cost and token monitoring) for LLM applications; proficient use of AI coding tools (Claude Code, Codex, Cursor) including configuring them, authoring custom skills and commands, and integrating them into team workflows; familiarity with guardrails and security considerations for production AI systems. * Ability to lead technical discovery with stakeholders, make architecture decisions, review code, mentor junior engineers, and raise the technical bar across the team. * Skilled at creating dashboards, data visualizations, live demos, and technical presentations to convey complex AI/ML findings to technical and non-technical stakeholders. * Experienced in gathering requirements, working in cross-functional teams, and producing clear documentation, user guides, and technical reports., * Bachelor's Degree in Arts/Sciences (BA/BS) in IT, Computer Science, or related field highly preferred Required * Will consider 7 years relevant experience in lieu of degree * 4+ years of relevant experience * Great attitude and team player. * Successful completion of background screening process. Essential Functions * Visual acuity (e.g., needed to prepare and analyze data, to transcribe documents, to view a computer, to read, to inspect objects, to operate machinery) * Manual Dexterity (e.g., picking, pinching, typing, or other working that uses the fingers) * Grasping (e.g., use of hand to apply pressure) * Hearing * Talking * Capacity to think, concentrate and focus over long periods of time * Ability to write complex documents in the [English] language * Ability to read complex documents in [English] language * Capacity to express thoughts orally (e.g., accurately, quick and loudly convey spoken instructions to workers) * Capacity to reason and make sound decisions * Ability to regularly perform all job functions at Company's office or work site ## Description We believe great work happens when flexibility and connection come together, and when we respect the whole person, not just the professional. That is why we operate under one flexible work model with a hybrid schedule of three in-office days per week, aligned to one of our three Greater Philadelphia area offices. Your Talent Acquisition partner will work closely with you to confirm the office location that best aligns with the role and your home location, ensuring you have the right balance of connection, collaboration, and flexibility to thrive Responsible for the end-to-end design, development, and deployment of advanced AI and machine learning solutions in support of D&Z AI initiatives. This senior engineering role requires equal expertise in generative AI application development and traditional machine learning, spanning the delivery of LLM-based agents, RAG systems, MCP (Model Context Protocol) servers, and predictive models to production. The AI/ML Engineer 2 owns complex projects independently, drives technical architecture decisions, evaluates and integrates emerging AI tools and frameworks, and mentors junior team members through code review and technical guidance. This role collaborates directly with cross-functional stakeholders to scope, design, and integrate AI/ML solutions into business processes, and communicates complex technical concepts clearly to both technical and non-technical audiences. Responsibilities * Designs, builds, and deploys production-grade generative AI applications including AI agents, RAG chatbots, workflow automations, and MCP servers using modern frameworks and vector databases. 30% * Designs, develops, and deploys production-grade predictive models using machine learning techniques across regression, classification, clustering, and time series. 20% * Authors custom skills, prompts, slash commands, and evaluation harnesses to accelerate team delivery and improve AI application quality, using AI coding tools such as Claude Code, Codex, and Cursor. 15% * Defines and owns production monitoring, drift detection, online/offline metrics, and retraining strategies for AI/ML systems; applies statistical methods and evaluation frameworks to validate outputs and drive iterative improvement. 10% * Owns data engineering and pipeline infrastructure underlying AI/ML systems, including retrieval architectures, embeddings, context engineering, and model training/fine-tuning workflows. 10% * Reviews code, mentors junior engineers, and contributes to team-wide engineering standards, documentation, and internal knowledge management patterns. Presents results and delivers live demonstrations of AI/ML solutions to technical and executive audiences. 15% ## Related Videos - [Coffee with Developers - Maria Apazoglou](https://www.wearedevelopers.com/videos/1209-coffee-with-developers-maria-apazoglou) - [TikTok's Privacy Innovation](https://www.wearedevelopers.com/videos/1036-tiktok-s-privacy-innovation) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Explainable machine learning explained](https://www.wearedevelopers.com/videos/589-explainable-machine-learning-explained) - [Serverless deployment of (large) NLP models ](https://www.wearedevelopers.com/videos/158-serverless-deployment-of-large-nlp-models) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [What Industries Outside of AI Are Hiring The Most AI Experts?](https://www.wearedevelopers.com/magazine/98-what-industries-outside-of-ai-are-hiring-the-most-ai-experts) - [13 AI Tools You Have to Try](https://www.wearedevelopers.com/magazine/219-13-ai-tools-you-have-to-try) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)