Software Engineer Graduate (Data-Speech-Product RD-Engineering-US) - 2027 Start

BYTEDANCE INC.
San Jose, CA, United States
about 1 month ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Starter
Compensation
$128,000.0 - $256,000.0
Working hours
Regular working hours
Job source

Tech stack

C++ (Programming Language) Compilers Nvidia CUDA Computer Engineering Distributed Computing Environment Distributed Systems Python (Programming Language) Performance Tuning Graphics Processing Unit (GPU) Deep Learning Parallel Computation Gpu Programming
+2 more
TensorRT Software Coding

Job description

The Speech team’s mission is to empower interaction and creation using speech & audio related technologies. The team focuses on cutting-edge R&D in areas like speech & audio, music processing, natural language understanding and multimodal deep learning. We are looking for top talents to work on these exciting technologies, integrate them into various products and ultimately bring joy to our global user base!, We are seeking a passionate AI Model Optimization Engineer to join our team. In this role, you will design and implement cutting-edge techniques to make AI models faster, more efficient, and easier to deploy at scale. You will collaborate across research and engineering to push the limits of AI performance in production environments., * Develop and implement algorithms for model optimization, including quantization, pruning, knowledge distillation, and efficient architectures.

  • Build and maintain performance benchmarking frameworks for large-scale training and inference.
  • Optimize training and inference pipelines on GPUs and across distributed systems.
  • Collaborate with ML researchers to transition optimized models into production.
  • Stay current with the latest research in model efficiency, compilers, and systems.

Requirements

  • Individuals who are completing or have recently completed a Bachelor’s/ Master’s degree in computer engineering or a related discipline.
  • Strong coding skills in Python and C++.

Preferred Qualifications

  • Experience with deep learning frameworks and distributed training systems.
  • Solid understanding of computer architecture, parallel computing, and GPU acceleration.
  • Familiarity with GPU programming (CUDA, Triton, or similar) is a plus.
  • Familiarity with ML compilers (e.g., TVM, XLA, TensorRT) is a plus.
  • Strong analytical skills and ability to work in a fast-paced team environment.

As a condition of employment, all successful candidates must be able to establish authorization to work in the United States. For this position, the Company does not provide sponsorship or any immigration-related benefits., Qualified applicants with arrest or conviction records will be considered for employment in accordance with all federal, state, and local laws including the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. Our company believes that criminal history may have a direct, adverse and negative relationship on the following job duties, potentially resulting in the withdrawal of the conditional offer of employment

Benefits & conditions

Compensation may vary outside of this range depending on a number of factors, including a candidate’s qualifications, skills, competencies and experience, and location. Base pay is one part of the Total Package that is provided to compensate and recognize employees for their work, and this role may be eligible for additional discretionary bonuses/incentives, and restricted stock units.

Benefits may vary depending on the nature of employment and the country work location. Employees have day one access to medical, dental, and vision insurance, a 401(k) savings plan with company match, paid parental leave, short-term and long-term disability coverage, life insurance, wellbeing benefits, among others. Employees also receive 10 paid holidays per year, 10 paid sick days per year and 17 days of Paid Personal Time (prorated upon hire with increasing accruals by tenure).

The Company reserves the right to modify or change these benefits programs at any time, with or without notice.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:37 min

Optimizing technical profiles for AI sourcing and recruitment

Mina Golesorkhi Mina Golesorkhi Ā· World Congress 2026 Europe

4:52 min

Essential phases in building and refining language models

Anshul Jindal Anshul Jindal +1 Ā· World Congress 2025

1:51 min

Evolution of custom compilers and virtual machines

Florian Rappl Ā· LIVE

6:21 min

Previewing upcoming hardware acceleration capabilities for Python environments

Chris Heilmann +2 Ā· LIVE

1:32 min

Driving tech talent attraction through deep individualization

Rudi Bauer +2 Ā· Cappuccino with HR

2:32 min

Core libraries driving inference engines and multi-GPU networking

Adolf Hohl Adolf Hohl Ā· World Congress 2024

Videos

See all

Related articles

See all