Senior Engineer, Machine Learning Application Developer

Samsung
San Jose, CA, United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Compensation
$124,000.0
Working hours
Regular working hours
Job source

Tech stack

C (Programming Language) Application Programming Interfaces (APIs) Artificial Intelligence Application Performance Management C++ (Programming Language) Profiling Computer Programming Computer Engineering Software Debugging Python (Programming Language) Machine Learning OpenGL
+8 more
OpenCL Tensorflow Software Systems Graphics Processing Unit (GPU) Pytorch Information Technology Vulkan Graphics API Software Performance

Job description

As a Machine Learning Application Developer, you will develop neural rendering applications and machine learning (ML) software that enable efficient execution of AI workloads on Samsung’s premium mobile GPUs.

In this individual contributor role, you will contribute to the development of software solutions that bridge machine learning workloads and GPU hardware capabilities. Working closely with hardware, software, and architecture teams, you will help optimize performance, efficiency, and resource utilization to support next-generation intelligent computing experiences.

  • You help developing and optimizing neural rendering applications, API-level software, and ML operator implementations, including GEMM, convolution, activations, and related workloads, using Vulkan, OpenGL, and OpenCL to enable efficient execution of ML and graphics workloads on Samsung GPU platforms.
  • You analyze software performance and hardware resource utilization to identify bottlenecks and optimize application performance, efficiency, and scalability across a variety of ML workloads.
  • You proactively seek collaborations with GPU architects, software engineers, and hardware teams to understand underlying hardware constraints and translate performance insights into optimized software solutions.
  • You leverage low-level performance analysis techniques, including assembly-level investigation when needed, to help improve execution efficiency and maximize GPU utilization.
  • You take initiatives on moderate-to-complex projects and help advance best practices and methodologies by staying current with the latest advancements in machine learning, neural rendering, and GPU technologies.

Requirements

  • 3+ years of experience with a Bachelor’s Degree in Computer Science, Computer Engineering, or comparable field, or 2+ years of experience with a Master’s Degree, or Ph.D.
  • Strong programming skills in C, C++, and Python.
  • Proficiency with API-level programming using in Vulkan, OpenGL, OpenCL, and machine learning frameworks such as PyTorch and TensorFlow.
  • Understanding of GPU hardware architecture and experience with low-level performance profiling, analysis, and optimization.
  • Hands-on experience developing neural rendering applications at the API level.
  • Working knowledge of machine learning operators and workloads, including GEMM, convolution, activations, and related computational kernels.
  • Ability to analyze hardware resource constraints and bottlenecks and develop software optimizations that improve performance and efficiency.
  • Working knowledge of assembly-level analysis, debugging, or optimization is preferred.
  • Strong analytical and problem-solving skills, with the ability to identify bottlenecks and propose data-driven solutions.
  • Excellent communication and collaboration skills, with the ability to navigate ambiguity in a fast-paced, global team environment., This position requires the ability to access information subject to U.S. export control restrictions. Applicants must have the ability to access export-controlled information or be eligible to receive a government authorization to access export-controlled information.

Benefits & conditions

At Samsung - SARC/ACL, base pay is one part of our total compensation package and is determined within a range. This provides the opportunity to progress as you grow and develop within a role. The base pay range for this role is between $124,000 and $208,400. Your actual base pay will depend on variables that may include your education skills, qualifications, experience, and work location.

Samsung employees have access to benefits including: medical, dental, vision, life insurance, 401(k), onsite lunch, employee purchase program, tuition assistance (after 6 months), paid time off, student loan program, wellness incentives, and many more. In addition, regular full-time employees (salaried or hourly) are eligible for MBO bonus compensation, based on company, division, and individual performance.

Additionally, this role might be eligible to participate in long term incentive plan and relocation.

This is an exempt position, which is not eligible for overtime pay under the Fair Labor Standards Act (FLSA).

About the company

Samsung, a world leader in advanced semiconductor technology, is founded on a simple philosophy - the endless pursuit of excellence will create a better world for all. At Samsung Austin Research and Development Center (SARC) and Advanced Computing Lab (ACL), we are building a center of excellence for Intellectual Property (IP) that is applied to high-performance computing devices (mobile, automotive, and other custom market segments) consumed by millions of people around the world. Come build with us!

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on samsung.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:42 min

Navigating emerging hardware standardization in vendor programming ecosystems

Paul Graham Paul Graham · LIVE

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · WWC Europe 2026

4:41 min

Replacing PyTorch with ONNX runtime for AWS Lambda deployments

Marek Suppa · LIVE

Videos

See all

Related articles

See all