> Markdown version of [/jobs/ext/1300880-ai-performance-library-architect](https://www.wearedevelopers.com/jobs/ext/1300880-ai-performance-library-architect). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # AI Performance Library Architect - **Company:** Intel Corporation - **Location:** Folsom, CA, United States - **Salary:** $170,500.0 - $315,490.0 - **Contract:** Internship / Graduate position - **Skills:** Artificial Intelligence, Artificial Neural Networks, Basic Linear Algebra Subprograms, Nvidia CUDA, Data Centers, Linux, Microprocessors, X86 Assembly Languages, OpenCL, Performance Tuning, Tensorflow, Software Engineering, Graphics Processing Unit (GPU), Pytorch, Information Technology, ONNX (Open Neural Network Exchange) Format, Free and Open-Source Software, Software Performance, Software Library - **Published:** July 16, 2026 - **Apply:** https://dejobs.org/x/x/85BE3DBC3574407E8721091A7E4F66CB/job/ ## About the Role You must possess the below minimum qualifications to be initially considered for this position. Preferred qualifications are in addition to the requirements and are considered a plus factor in identifying top candidates., * Master's degree in Mathematics, Physics, Computer Science, or a relevant STEM field. OR * Ph.D. degree in Mathematics, Physics, Computer Science, or a relevant STEM field. * 5+ years of experience in the following areas: * C and C++ Maintaining or contributing to open-source software projects * Software libraries design and architecture * Implementation of linear algebra algorithms (functions from BLAS, LAPACK, or PyTorch) * Performance engineering and software performance optimizations * Floating point arithmetic and numerical stability * Software development on Linux * Low-level performance optimizations using CUDA, x86 assembly or intrinsics, or OpenCL, * 3 years+ Machine learning and deep learning algorithms or High-performance computing (HPC) applications development * 3 year+ Floating point implementations of transcendental functions (sin, cos, tanh, elu, etc) * 1 year+ Algorithms for non-IEEE low precision data types (bfloat16, fp8, fp4) * 1 year+ AI assisted software development Requirements listed would be obtained through a combination of industry relevant job experience, internship experiences and/or schoolwork/classes/research. ## Description Software and AI (SAI) organization is looking for a software development engineer to work on oneDNN project ( https://github.com/uxlfoundation/oneDNN ). oneDNN is a complex cross-platform open-source software project focusing on neural network performance. oneDNN is a critical and highly visible component of Intel AI strategy, powering key AI applications including OpenVINO, Tensorflow, Pytorch, ONNX Runtime, and more. In this role, you will be responsible for design, development, and maintenance of new functionality in oneDNN to enable performance critical portions of AI workloads. In this role you will be supporting software developers optimizing AI frameworks and workloads for Intel CPUs and GPUs, as well as cross-platform ecosystem of AI software developers contributing to oneDNN., The Software Team drives customer value by enabling differentiated experiences through leadership AI technologies and foundational software stacks, products, and services. The group is responsible for developing the holistic strategy for client and data center software in collaboration with OSVs, ISVs, developers, partners and OEMs. The group delivers specialized NPU IP to enable the AI PC and GPU IP to support all of Intel's market segments. The group also has HW and SW engineering experts responsible for delivering IP, SOCs, runtimes, and platforms to support the CPU and GPU/accelerator roadmap, inclusive of integrated and discrete graphics., This role will be eligible for our hybrid work model which allows employees to split their time between working on-site at their assigned Intel site and off-site. * Job posting details (such as work model, location or time type) are subject to change. ## Related Videos - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [From Model to Metal: An Open Source Stack for Accelerating Intelligence](https://www.wearedevelopers.com/videos/1636-from-model-to-metal-an-open-source-stack-for-accelerating-intelligence) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Coffee with Developers - Stephen Jones - NVIDIA](https://www.wearedevelopers.com/videos/1303-coffee-with-developers-stephen-jones-nvidia) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [Serverless deployment of (large) NLP models ](https://www.wearedevelopers.com/videos/158-serverless-deployment-of-large-nlp-models) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere) - [What Industries Outside of AI Are Hiring The Most AI Experts?](https://www.wearedevelopers.com/magazine/98-what-industries-outside-of-ai-are-hiring-the-most-ai-experts) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift)