GPU Kernel Engineer

Typesafe AI Inc.
San Francisco, CA, United States
about 2 months ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$180,000.0 - $280,000.0
Working hours
Regular working hours
Job source

Tech stack

Nvidia CUDA Large Language Models

Job description

We’re looking for a GPU kernel engineer with deep, low-level CUDA expertise to make our training and inference faster and more efficient. You’ll write and optimize custom kernels, profile and eliminate bottlenecks, and work close to the metal across our model stack., * Write, optimize, and maintain high-performance GPU kernels (e.g., in CUDA / CuTe DSL) for training and inference

  • Profile end-to-end performance and eliminate bottlenecks across the stack
  • Partner with research and platform engineers to squeeze maximum throughput and minimum latency out of our hardware

Requirements

  • Have deep CUDA / GPU kernel expertise and a track record of real performance wins
  • Have built and optimized inference / training kernels
  • Have hands-on LLM training experience (real, not at a hobbyist level)
  • Reason from first principles about performance, memory, and parallelism
  • Are responsible, ownership-inclined team players who are mission aligned and excited to go all-in

Benefits & conditions

Pulled from the full job description

  • 401(k)
  • Health insurance, * Base salary of $180k-280k plus equity, based on leveling
  • 100% covered health insurance
  • Daily lunch and dinner
  • Visa sponsorships
  • 401K plans

Compensation Range: $180K - $280K

About the company

TypeSafe is a frontier model lab. We build reliable and general AI systems to power economically valuable automation. Our mission is to usher in a new era of Transformative Artificial Intelligence (TAI): technology with the power to drive a societal shift on the scale of the agricultural and industrial revolutions.

While others chase benchmarks and academic puzzles, we’ve been quietly rethinking the LLM stack from first principles - building a new kind of general frontier model designed for real-world reliability, decision-making, and autonomy in production.

We’re a small, fast-moving team from OpenAI, Google Brain, and Meta/FAIR, backed by top-tier investors. Since mid-2024, we’ve been engineering the foundation for what comes after the current “state-of-the-art” - a model that actually gets things done., We’re a small, flat, close-knit team dedicated to real-world impact preparing the world for Transformative AI. Our team works fully in-person in our San Francisco office near Embarcadero station. We love what we do and care about our work a lot.

We strive for excellence and craftsmanship and won’t stop until we get there. When the team wins, we all win, and we enjoy collaborating and inspiring each other to grow as a team and as individuals.

We also value emotional honesty, kindness, and bringing your whole self to work. We build machines; we don’t try to be machines.

We want you to be able to do the most impactful work of your career at TypeSafe and help define our future as a company.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 ¡ World Congress 2026 Europe

3:32 min

Fundamentals and limitations of large language models

Krzystof Czieslak ¡ LIVE

6:21 min

Previewing upcoming hardware acceleration capabilities for Python environments

Chris Heilmann +2 ¡ LIVE

1:37 min

Accelerating compute with focused developer tools

Julia Koch Julia Koch +1 ¡ World Congress 2026 Europe

3:30 min

Transitioning from CUDA software architect to user

Stephen Jones ¡ Coffee With Developers

2:32 min

Core libraries driving inference engines and multi-GPU networking

Adolf Hohl Adolf Hohl ¡ World Congress 2024

Videos

See all

Related articles

See all