> Markdown version of [/jobs/ext/2522785-principal-engineer-machine-learning-engineering-quantization-sw](https://www.wearedevelopers.com/jobs/ext/2522785-principal-engineer-machine-learning-engineering-quantization-sw). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Principal Engineer, Machine Learning Engineering (Quantization SW) - **Company:** Qualcomm - **Location:** San Diego, CA, United States (Remote available) - **Experience:** Expert - **Salary:** $200,800.0 - $301,200.0 - **Contract:** Permanent contract - **Skills:** Agile Methodology, Artificial Intelligence, Systems Engineering, Artificial Neural Networks, Code Review, Information Systems, Computer Engineering, Continuous Integration, Software Debugging, Python (Programming Language), Logical Volume Manager, Machine Learning, Tensorflow, Smart Devices, Software Engineering, Pytorch, Large Language Models, Information Technology, ONNX (Open Neural Network Exchange) Format, HuggingFace - **Published:** August 3, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=5946550db164dae9 ## About the Role * Bachelor's degree in Computer Science, Engineering, Information Systems, or related field and 8+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience., Master's degree in Computer Science, Engineering, Information Systems, or related field and 7+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience. OR PhD in Computer Science, Engineering, Information Systems, or related field and 6+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience., * Strong Software Engineering/Development skills combined with a solid foundation in AI and general ML techniques. * Proven hands-on experience evaluating and optimizing Generative AI workflows for accuracy, performance, and other key metrics. * Prior experience with ML model optimization frameworks and a familiarity with applying techniques such as quantization, pruning etc. * Knowledge of neural networks, with hands-on experience using ML frameworks such as PyTorch, ONNX etc. * Strong Python design and implementation skills. * Strong general analytical and debugging skills. * Prior experience working in agile environments. * Prior experience in collaborating with multi-disciplinary teams across time zones. * Strong leadership skills as a mentor, team player, communicator and presenter. * Experience deploying GenAI LLM/LVM models on edge devices. * Prior experience with model quantization, profiling and running models on edge devices. * Prior experience with frameworks like Huggingface Optimum, ONNX runtime and OpenVino. * Proven hands-on experience establishing a high-quality software delivery process using industry best-practices (code review, CI/CD, automation, etc.) is a plus. ## Description * Work in a dynamic research environment, * Be part of a multi-disciplinary team of researchers and software engineers who work with cutting edge AI frameworks and tools. * Architect, design, develop and test model optimization techniques that include - but are not limited to - graph optimization, pruning and quantization. ## Related Videos - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Architecting the Future: Leveraging AI, Cloud, and Data for Business Success](https://www.wearedevelopers.com/videos/1096-architecting-the-future-leveraging-ai-cloud-and-data-for-business-success) - [Machine learning in the browser with TensorFlowjs](https://www.wearedevelopers.com/videos/155-machine-learning-in-the-browser-with-tensorflowjs) - [Developing an AI.SDK](https://www.wearedevelopers.com/videos/198-developing-an-ai-sdk) - [From Model to Metal: An Open Source Stack for Accelerating Intelligence](https://www.wearedevelopers.com/videos/1636-from-model-to-metal-an-open-source-stack-for-accelerating-intelligence) - [Localized Open Models in Production: What Builders Need to Know](https://www.wearedevelopers.com/videos/100270-localized-open-models-in-production-what-builders-need-to-know) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [The Fastest-Growing Tech Sectors to Look Out for in 2025](https://www.wearedevelopers.com/magazine/373-the-fastest-growing-tech-sectors-to-look-out-for-in-2025) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers)