Senior Staff Engineer - AI Workloads & Storage

Samsung
San Jose, CA, United States
8 days ago
Apply on jobs.localjobnetwork.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$189,000.0 - $301,000.0
Working hours
Regular working hours

Tech stack

Adobe Flash Artificial Intelligence C++ (Programming Language) Computer Engineering Extract Transform Load (ETL) Linux Firmware PCI Express Remote Direct Memory Access Software Engineering SystemC Weka
+7 more
Ceph (Software) Large Language Models Perf (Linux) Information Technology Low Latency TensorRT Nvme

Job description

  • Own AI workload characterization.Profile production and emerging LLM inference, RAG, and training workloads to quantify their I/O, bandwidth, latency, and capacity demands, and turn those findings into concrete storage and memory-hierarchy design decisions.
  • Identify optimal data placement.Analyze workload access patterns to determine how data should be placed and separated on flash, and map those insights onto SSD data-placement technologies such asNVMe Flexible Data Placement (FDP)andstreamsto reduce write amplification and improve endurance, latency, and QoS.
  • Collaborate with key customersto identify differentiating SSD capabilities for AI workloads, and develop proof-of-concept implementations as part of those customer engagements - turning workload insights into demonstrable data-path, tiering, and data-placement wins.
  • Lead deep-dive performance analysisspanning the inference runtime, the Linux storage and networking stack, and the underlying hardware, tuning for latency, throughput, cost, and GPU utilization.
  • Build and evaluate transactional and system-level modelsof proposed architectures to de-risk decisions before hardware exists, and validate them against measured behavior.
  • Engage with the standards and open ecosystem- SNIA (including Storage.AI), MLCommons/MLPerf, and the open inference stack - to align our work with where the industry is heading and to shape it where we can.
  • Set technical direction others build on.Make build-vs-buy and architectural calls, establish benchmarking methodology and best practices, and mentor engineers across the org.
  • Partner cross-functionallywith product, hardware, and research teams, and with external vendors and partners, to bring architectures from concept to deployment., At Samsung Semiconductor, we use Artificial Intelligence (AI) tools in the recruitment process to enhance efficiency. However, AI is used as a support tool, not a final decision-maker. All hiring decisions are made by our human recruiting team and hiring managers to ensure every candidate is evaluated fairly and holistically.

Requirements

  • Bachelor’s degree 15+ years relevant industry experience or Master’s degree 13+ years’ experience or PhD with 10+ years relevant industry experience.
  • Extensive experience (typically10-15+ years) in systems, storage, or ML-systems software, with a track record of architecting systems that materially improved performance, reliability, or cost.
  • Demonstratedtechnical leadership and cross-team influence: setting direction, driving decisions across organizational boundaries, and mentoring senior engineers.
  • Working knowledge of modern AI inference, especially transformer architectures - attention, KV cache, batching, and the memory/compute trade-offs of serving large models.
  • Deep systems-level understandingof the Linux storage stack (block layer, I/O scheduling, NVMe) and ofNAND/SSD internals(flash-translation layer, garbage collection, endurance/write-amplification, latency behavior), plus hands-on performance analysis skill (e.g., perf, ftrace, eBPF, blktrace, fio).
  • Fluency inPython plus a systems language(C/C++, Rust, or Go).
  • MS or PhDin Computer Science, Electrical/Computer Engineering, or a related field preferred - or equivalent practical experience.

Preferred Qualification

  • Hands-on experience with the moderninference stack: vLLM, SGLang, LMCache, NVIDIA Dynamo, TensorRT-LLM, or Triton.
  • Familiarity withGPU-adjacent data movement and memory frameworks: NIXL, DOCA / DOCA MemOps, GPUDirect Storage, RDMA, NVMe-oF, and BlueField / DPU offload.
  • Understanding of GPU and TPU architecture(memory hierarchy, interconnects, and how accelerator design shapes I/O and data-movement demands) is highly desired.
  • Experience withuser-mode storage access frameworks: SPDK, uNVMe, libvfn, or similar.
  • SSD firmware experience- flash-translation layer, wear-leveling and garbage-collection algorithms, and data-placement features such asFDP / streams / ZNS- ideally paired with the ability to co-design firmware and host-side placement policy from workload characterization.
  • AI-workload characterization and benchmarkingexperience, and familiarity withSNIA Storage.AIandMLCommons / MLPerf.
  • Transactional / discrete-event or system-level modelingexperience in frameworks such as SystemC, SimPy, or similar.
  • Experience withSSD architecture and interfaces- NVMe (including ZNS, Flexible Data Placement / FDP), open-channel SSDs, computational storage - and with PCIe Gen5, CXL, and large-scale GPU-cluster storage (VAST, WEKA, Lustre, Ceph).

Benefits & conditions

The pay range below is for all roles at this level across all US locations and functions. Paywithin this range varies by work locationand may also depend on job-related knowledge, skills,and experience. We also offer incentive opportunities that reward employees based on individual and company performance.

This is in addition to our diverse package of benefits centered around the wellbeing of our employees and their loved ones. In addition to the usual Medical/Dental/Vision/401k, our inclusive rewards plan empowers our people to care for their whole selves. An investment in your future is an investment in ours.

Give Back With a charitable giving match and frequent opportunities to get involved, we take an active role in supporting the community. Enjoy Time Away You’ll start with 4+ weeks of paid time off a year, plus holidays and sick leave, to rest and recharge. Care for Family Whatever family means to you, we want to support you along the way-including a stipend for fertility care or adoption, medical travel support, and virtual vet care for your fur babies. Prioritize Emotional Wellness With on-demand apps and free confidential therapy sessions, you’ll have support no matter where you are. Stay Fit Eating well and being active are important parts of a healthy life. Our onsite Cafe and gym, plus virtual classes, make it easier. Embrace Flexibility Benefits are best when you have the space to use them. That’s why we facilitate a flexible environment so you can find the right balance for you. Base Pay Range $189,000-$301,000 USD

About the company

Please Note:

To provide the best candidate experience amidst our high application volumes, each candidate is limited to 10 applications across all open jobs within a 6-month period.

Advancing the World’s Technology Together

Our technology solutions power the tools you use every day–including smartphones, electric vehicles, hyperscale data centers, IoT devices, and so much more. Here, you’ll have an opportunity to be part of a global leader whose innovative designs are pushing the boundaries of what’s possible and powering the future.

We believe innovation and growth are driven by an inclusive culture and a diverse workforce. We’re dedicated to empowering people to be their true selves. Together, we’re building a better tomorrow for our employees, customers, partners, and communities.

Please Note:

To provide the best candidate experience amidst our high application volumes, each candidate is limited to 10 applications across all open jobs within a 6-month period.

Advancing the World’s Technology Together

Our technology solutions power the tools you use every day–including smartphones, electric vehicles, hyperscale data centers, IoT devices, and so much more. Here, you’ll have an opportunity to be part of a global leader whose innovative designs are pushing the boundaries of what’s possible and powering the future.

We believe innovation and growth are driven by an inclusive culture and a diverse workforce. We’re dedicated to empowering people to be their true selves. Together, we’re building a better tomorrow for our employees, customers, partners, and communities.

Our technology solutions power the tools you use every day–including smartphones, electric vehicles, hyperscale data centers, IoT devices, and so much more. Here, you’ll have an opportunity to be part of a global leader whose innovative designs are pushing the boundaries of what’s possible and powering the future.

We believe innovation and growth are driven by an inclusive culture and a diverse workforce. We’re dedicated to empowering people to be their true selves. Together, we’re building a better tomorrow for our employees, customers, partners, and communities.

Our technology solutions power the tools you use every day–including smartphones, electric vehicles, hyperscale data centers, IoT devices, and so much more. Here, you’ll have an opportunity to be part of a global leader whose innovative designs are pushing the boundaries of what’s possible and powering the future.

We believe innovation and growth are driven by an inclusive culture and a diverse workforce. We’re dedicated to empowering people to be their true selves. Together, we’re building a better tomorrow for our employees, customers, partners, and communities.

Our technology solutions power the tools you use every day–including smartphones, electric vehicles, hyperscale data centers, IoT devices, and so much more. Here, you’ll have an opportunity to be part of a global leader whose innovative designs are pushing the boundaries of what’s possible and powering the future.

We believe innovation and growth are driven by an inclusive culture and a diverse workforce. We’re dedicated to empowering people to be their true selves. Together, we’re building a better tomorrow for our employees, customers, partners, and communities.

At the Technology Enabling Development Lab (TED), our core development focus is the host interface firmware layer that sits in the intersection of system software and flash management firmware. This key host interface firmware technology drives Samsung’s breakthrough V-NAND technology and enables our customers to power performance-oriented, demanding, enterprise-class applications ranging from hyper-scale data centers, to big data processing, to software defined virtualized storage arrays and infrastructures.

We are building the next generation ofNAND/SSD storage systemsdesigned for the demands of large-scale AI. As the compute cost of transformer inference falls, the bottleneck is shifting to how quickly and efficiently we can move model weights, KV cache, and activations through the storage hierarchy - and NAND flash and SSDs are increasingly the tier where that data lives. Our focus is on making SSDs first-class citizens in the AI data path, from the NAND media and flash-translation layer up through NVMe and networked storage.

We are looking for aSr Staff Engineerwho lives at the intersection ofAI inference systemsandstorage/systems software. This is a hands-on technical leadership role: you will characterize real AI workloads, translate what you learn into architecture, and drive that direction across inference, platform, and hardware teams. This is a rare seat for someone who is equally comfortable reading a transformer serving stack and a Linux block-layer trace.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on jobs.localjobnetwork.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

1:12 min

Addressing the competitive landscape of specialized hardware demands

Hazal Mestci +1 · Coffee With Developers

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

1:58 min

Verifying hardware access and exploring AI inference scaling

Piotr Zaniewski Piotr Zaniewski · World Congress 2026 Europe

3:18 min

Understanding software complexity in modern computing ecosystems

Andrew Wafaa Andrew Wafaa · World Congress 2024

8:22 min

Simulating a Linux terminal and running Spring Boot

Jakov Semenski · LIVE

Videos

See all

Related articles

See all