Machine Learning Infrastructure Engineer
Vision Technologies, LLC
Redwood City, CA, United States
3 days ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on www.careerbuilder.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Compensation
$105,000.0 - $143,000.0
Working hours
Regular working hours
Job source
Tech stack
Application Programming Interfaces (APIs)
Artificial Intelligence
Systems Engineering
C++ (Programming Language)
Cloud Computing
Distributed Systems
Python (Programming Language)
Machine Learning
Open Source Technology
Azure Machine Learning
Software Engineering
Data Logging
+10 more
Scripting
Graphics Processing Unit (GPU)
Autoscaling
Large Language Models
Caching
Kubernetes
Information Technology
Free and Open-Source Software
TensorRT
Programming Languages
Requirements
- Bachelor’s or Master’s degree in Computer Science or a related field.
- Six or more years of experience in distributed systems, infrastructure, or ML platform engineering.
- Strong proficiency in Python and a systems language such as Go, Rust, or C++.
- Deep experience operating high-throughput, low-latency services in production.
- Hands-on experience with LLM or large model inference frameworks such as vLLM or TensorRT-LLM.
- Strong understanding of GPU architecture, memory hierarchies, and accelerator utilization.
- Familiarity with Kubernetes, autoscaling, and modern cloud platforms.
- Experience with observability stacks including metrics, tracing, and structured logging.
- Solid grounding in performance engineering and capacity planning.
- Strong communication and incident response skills.
Preferred Qualifications
- Open-source contributions to model serving infrastructure.
- Experience with multi-region or globally distributed AI serving.
- Familiarity with model quantization, distillation, and compression techniques.
- Exposure to FinOps for AI workloads and cost-efficient serving design.
- Experience supporting external-facing AI APIs at scale., Application Programming Interface (API), Artificial Intelligence (AI), Autoscaling, C++ Programming Language, Caching, Capacity and Performance Management, Cloud Computing, Communication Skills, Computer Science, Consulting, Distributed Computing, EAD, GPU (Graphics Processing Unit), High Tech Industry, High Throughput, Incident Response, Machine Learning, Memory Hardware, Metrics, Open Source, Performance Engineering, Python Programming/Scripting Language, Rust Programming Language, Software Development, Systems Engineering
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.careerbuilder.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
LM
Luis Minvielle
almost 3 years ago
LM
Luis Minvielle
How to Become an AI Engineer
almost 3 years ago
BB
Benedikt Bischof
MLOps And AI Driven Development
over 4 years ago
BB
Benedikt Bischof
MLOps – What’s the deal behind it?
almost 4 years ago
BB
Benedikt Bischof
MLops – Deploying, Maintaining And Evolving Machine Learning Models in Production
about 4 years ago
EM
Eli McGarvie
Highest Paying Tech Companies for Developers
over 3 years ago