AI Infrastructure Engineer

Intel Corporation
Folsom, CA, United States
3 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence General-Purpose Computing on Graphics Processing Units Python (Programming Language) Open Source Technology Performance Tuning Data Driven Tests Software Engineering AI Infrastructure Graphics Processing Unit (GPU) Pytorch Large Language Models C++14

Job description

Experteer Overview In this role you will push LLM inference on Intel GPUs to peak performance. You will work across the stack, profiling bottlenecks, developing high-performance kernels, and upstreaming optimizations into open-source projects like vLLM and SGLang. You’ll shape hardware roadmaps through data-driven analysis and cross-team collaboration. This is a hands-on, impact-driven position that ties AI workloads to Intel’s hardware strategy. Compensation / Benefits * Own end-to-end optimization of running state-of-the-art LLMs on Intel GPUs * Profile, diagnose, and resolve cross-stack performance bottlenecks * Design, write, and optimize custom high-performance kernels for attention, MoE, quantization, and operator fusions * Upstream architectural improvements into open-source repos like vLLM, SGLang, and PyTorch * Collaborate with architecture and compiler teams to shape future GPU roadmaps based on GenAI workloads * Promote AI infrastructure and performance optimization initiatives Tasks * 4+ years of software engineering experience in GPU computing, AI systems, or HPC (or 3+ years with a Masters, or PhD) * Proficiency in modern C++ and Python * Ability to read and modify complex systems-level code Key requirements * competitive pay * stock bonuses * health insurance * retirement plan * vacation

Requirements

GenAI Tasks * 4+ years of software engineering experience in GPU computing, AI systems, or HPC (or 3+ years with a Masters, or PhD) * Proficiency in modern C++ and Python * Ability to read and modify complex systems-level code Key requirements * competitive pay * stock bonuses * health insurance * retirement plan * vacation

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:24 min

Comprehensive AI infrastructure stacks at the Linux Foundation

Matt White Matt White · WWC 2025

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

3:26 min

Parameterizing test functions with different input datasets

Florian Bruhin · WWC 2021

1:29 min

Tech infrastructure capacity and AI product innovations

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · WWC Europe 2026

2:40 min

Building an automated workflow for AI-driven test generation

Alisa Hrustic Alisa Hrustic · Europe 2026 Virtual

Videos

See all

Related articles

See all