Mid-Level Machine Learning Engineer
TETRAMEM INC
San Jose, CA, United States
3 months ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
5 years minimum
Compensation
$110,000.0 - $300,000.0
Working hours
Regular working hours
Job source
Tech stack
Artificial Intelligence
Audio Signal Processing
C++ (Programming Language)
Field-Programmable Gate Array (FPGA)
Python (Programming Language)
Machine Learning
Tensorflow
Systems Architecture
Application Specific Integrated Circuits
Pytorch
Information Technology
Low Latency
+3 more
ONNX (Open Neural Network Exchange) Format
Hardware Acceleration
TensorRT
Job description
- Develop, optimize, and deploy lightweight machine learning models for edge AI applications, particularly for audio processing.
- Implement and optimize ML models on embedded platforms, including FPGA and custom ASIC solutions.
- Work closely with hardware and software teams to integrate ML models into production systems.
- Research and implement state-of-the-art ML techniques to enhance model efficiency, latency, and power consumption for embedded AI applications.
- Improve inference efficiency and model compression techniques, including quantization, pruning, and knowledge distillation.
- Collaborate with cross-functional teams to drive innovation and contribute to the overall system architecture.
- Provide technical leadership and mentorship to junior engineers.
- Publish research findings, present at conferences, and contribute to open-source projects when applicable.
Requirements
- 5+ years of experience or PhD in Computer Science, Electrical Engineering, or related fields.
- Strong experience in machine learning, with a focus on edge AI and lightweight model deployment.
- Expertise in ML frameworks such as PyTorch, TensorFlow, JAX.
- Proficiency in programming languages such as C/C++, Python, and experience with ML model optimization.
- Ability to work independently and collaboratively in a fast-paced startup environment.
Experience in one or more of the following areas considered a strong plus:
- Understanding of ML compiler and runtime design.
- Experience working with tools such as Optimum, ONNX, TensorRT, TFLite/LiteRT, ncnn, or CoreML.
- Familiarity with hardware acceleration techniques.
- Experience in embedded system development.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on indeed.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
LM
Luis Minvielle
almost 3 years ago
BB
Benedikt Bischof
MLOps And AI Driven Development
over 4 years ago
EM
Eli McGarvie
Highest Paying Tech Companies for Developers
over 3 years ago
BB
Benedikt Bischof
MLOps – What’s the deal behind it?
almost 4 years ago
CH
Chris Heilmann
Dev Digest 120 - Apple and peers
about 2 years ago
CH
Chris Heilmann
Dev Digest 132 - Binging WADFlix?
almost 2 years ago