Product Performance Engineer

Hark's Co.
San Jose, CA, United States
5 days ago
Apply on www.adzuna.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$120,000.0 - $300,000.0
Working hours
Regular working hours
Job source

Tech stack

C++ (Programming Language) Profiling Computer Programming Data Centers Python (Programming Language) Model Validation Information Technology

Job description

  • Build deep understanding of target workloads and characterize their compute, memory, bandwidth, and latency demands on a variety of target hardware platforms
  • Model performance at every level - silicon, system, service, and end-to-end user experience - and connect insights across layers
  • Develop and maintain analytical and simulation-based models that project system performance across design and assumption variants
  • Collaborate with silicon, software, hardware, model, and service teams to align assumptions and deliver insights
  • Validate models against measured workload behavior and continuously improve fidelity
  • Translate findings into clear, actionable recommendations for architects, engineers, and leadership
  • Proactively identify performance risks and design opportunities ahead of architectural commit points

Requirements

  • 5+ years analyzing workload behavior and modeling system performance for consumer or data center products
  • BS in Computer Science, Electrical Engineering, or related technical discipline
  • Hands-on experience with profiling, benchmarking, and tracing tools across the stack
  • Strong programming skills in Python and at least one systems language (C/C++/Rust)
  • Demonstrated ability to translate quantitative analysis into architectural recommendations

Bonus Qualifications

  • MS or PhD in a relevant technical discipline
  • Experience with hardware/software co-design and architectural exploration tooling
  • Familiarity with accelerator, SoC, or memory subsystem architecture
  • Comfort moving fluidly between rigorous quantitative work and cross-functional influence
  • Curiosity, technical depth, and a bias toward learning new system domains quickly

Benefits & conditions

The US base salary range for this full-time position is between $120,000 - $300,000 annually.

The pay offered for this position may vary based on several individual factors, including job-related knowledge, skills, and experience. The total compensation package may also include additional components/benefits depending on the specific role. This information will be shared if an employment offer is extended.

About the company

Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persistent memory.

We’re pairing that intelligence with next-generation hardware to create a universal interface between humans and machines. While today’s AI largely operates through chat boxes and decade-old devices, Hark is focused on what comes next: agentic systems that interact naturally with people and the real world.

To get there, we’re developing multimodal models and next-generation AI hardware together - designed from the ground up as a single, unified interface for a new era of intelligent systems.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.adzuna.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:31 min

Profiling computing workloads across environments using Performance Studio

Andrew Wafaa Andrew Wafaa · World Congress 2024

5:08 min

Validating requests and responses using data transfer objects

Roman Alexis Anastasini · World Congress 2021

51 sec

Repurposing hardware and operating underwater data centers

Chris Heilmann +1 · LIVE

47 sec

Profiling native execution calls with async-profiler

Gonzalo Ortiz Jaureguizar Gonzalo Ortiz Jaureguizar · World Congress 2026 Europe

56 sec

The negligible impact of AI model size on security

Julian Totzek-Hallhuber Julian Totzek-Hallhuber · World Congress 2026 Europe

4:03 min

Managing massive power consumption scaling in AI data centers

Stephan Gillich Stephan Gillich +3 · World Congress 2024

Videos

See all

Related articles

See all