Senior/Principal AI Systems Engineer

Huawei
Dresden, Germany
29 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours
Languages
English

Tech stack

Abstraction Layers Artificial Intelligence Systems Engineering Program Optimization Profiling Distributed Systems Machine Learning Performance Tuning Cloud Platform System Large Language Models Multi-Agent Systems Information Technology
+2 more
Machine Learning Operations Microservices

Job description

Huawei’s Hilbert Research Center (DRC)’s mission is to explore programming models, operating system and virtualization technologies on multicore heterogeneous architectures and NVM/SCM platforms, aiming to provide a high-performance, reliable abstraction layer for efficient resource utilization. DRC focuses its technical research in the key areas of Smart Mobile, Telecom, Autonomous Driving, Internet of Things, and Industry 4.0. Join us as a Principal AI Systems Engineer (m/f/d) and lead the creation of next-generation agent systems for the AI PC era. In this role, you will drive research into intelligent, autonomous agents capable of operating across heterogeneous compute tiers - from on-device execution to near-device accelerators and cloud-scale reasoning. You will work closely with product teams to translate research breakthroughs into production-ready capabilities, while growing and mentoring a world-class research and development team. Your mission Architect hybrid agent systems - Design agent architectures that seamlessly combine on-device inference, near-device compute (local servers, LAN accelerators), and cloud-scale reasoning to deliver responsive, capable, and context-aware behavior. Lead multi-tier intelligence strategy - Define how agents decide where to run each component (LLM, vision, speech, planning, memory) based on latency, cost, privacy, and hardware availability. Build scalable agent runtimes - Own the architecture of runtimes that coordinate multiple models, sensor inputs, memory subsystems, and decision-making loops across heterogeneous compute tiers. Drive performance across CPU/GPU/NPU/cloud - Lead optimization efforts for latency, throughput, memory footprint, and energy efficiency across AI PCs, edge devices, and cloud clusters. Translate research into production agents - Partner with research teams to evaluate new model architectures and convert them into robust agent behaviors that run efficiently across distributed environments. Mentor and grow the engineering team - Provide technical leadership, architectural guidance, and mentorship to engineers working on inference, agent logic, distributed systems, and hardware-aware optimization. Your areas of expertise

Requirements

Advanced academic background - BSc/Master/PhD in Computer Science or related fields, with strong foundations in systems engineering or machine learning. ML systems engineering - Deep experience building and integrating inference pipelines into production systems. Agent architecture - Expertise in designing planning loops, memory subsystems, tool-use interfaces, and multi-step reasoning pipelines. AI PC optimization - Strong knowledge of CPU/GPU/NPU pipelines, memory hierarchies, and low-level performance tuning on PC-class hardware. Technical communication - Fluent in English with the ability to communicate complex systems and research concepts clearly to engineers, product teams, and leadership. Good-to-Have (Preferred Qualifications) Distributed systems - Familiarity with RPC frameworks, message buses, microservices, or distributed scheduling for multi-tier agent workloads. Hybrid inference deployment - Experience orchestrating model execution across on-device, near-device, and cloud environments. Agent architecture - Expertise in designing planning loops, memory subsystems, tool-use interfaces, and multi-step reasoning pipelines. Profiling & benchmarking - Skilled in designing and interpreting performance experiments across heterogeneous compute. Model optimization - Experience with quantization, pruning, distillation, and other efficiency-driven transformations.

About the company

Our culture is characterized by innovative power and team spirit as well as the intensive exchange of knowledge and experience within our global network. We offer healthy meals ranging from traditional Chinese to western delicacies in our famous company canteen. To keep your development ongoing, you will find a broad range of training opportunities. Many online and face-to-face training programs incl. language courses in German and Mandarin. Our diverse and welcoming environment is shaped by different backgrounds and around 40 individual nationalities. Self-responsible work in a competent, motivated and constantly growing team. Please send your application and CV (incl. cover letter and reference letters) in English. Huawei is a leading global information and communications technology (ICT) solutions provider. Our ICT solutions, products and services are used in more than 170 countries and regions, serving over one-third of the world’s population. With 208,000 employees, Huawei is committed to develop the future information society and build a Better Connected World. Department OS Kernel (Hilbert Research Center) Locations Huawei Hilbert Research Center (Dresden) Huawei Hilbert Research Center (Dresden) About Huawei Research Center Germany Huawei is a leading global information and communications technology (ICT) solutions provider. Our ICT solutions, products and services are used in more than 170 countries and regions, serving over one-third of the world’s population. With 208,000 employees, Huawei is committed to develop the future information society and build a Better Connected World. careersite–jobs–form-overlay#showFormOverlay”>

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.de

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

47 sec

Profiling native execution calls with async-profiler

Gonzalo Ortiz Jaureguizar Gonzalo Ortiz Jaureguizar · WWC Europe 2026

3:07 min

Transitioning architecture to microservices at Netflix

Steve Upton Steve Upton · WWC 2022

2:36 min

Applying supervised machine learning for practical rule extraction

Katja Träumner

6:17 min

Redefining engineering roles and system design responsibilities

Daniel Gebler Daniel Gebler +2 · WWC 2025

2:17 min

Comparing code profiling with surface level monitoring

Jérôme Vieilledent · LIVE

1:49 min

Augmenting junior and principal engineering roles with AI

Neel Sundaresan Neel Sundaresan +1 · WWC Europe 2026

Videos

See all

Related articles

See all