Machine Learning Engineer

ML Inc.
United States
1 day ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Airflow Data Structures Software Debugging Software Design Patterns Python (Programming Language) Linear Programming Machine Learning Language Modeling Open Source Technology Tensorflow Data Processing
+8 more
Pytorch Large Language Models Apache Spark Deep Learning Production Code Machine Learning Operations Multiaccess Edge Computing Data Pipelines

Job description

As a Machine Learning Engineer, you will play a central role in translating cutting-edge machine learning research into scalable, production-ready solutions. You will collaborate closely with cross-functional teams to identify opportunities where ML can drive product value, architect robust model-centric systems, and ensure their seamless integration into real-world applications. The role requires a strong balance between theoretical understanding and engineering execution, with a focus on building reliable, maintainable, and high-impact AI-driven features that align with Nace.AI’s strategic objectives., * Design, build, and maintain end-to-end ML systems, including synthetic data pipelines, model training, debugging, and performance evaluation.

  • Fine-tune large language models (LLMs) and implement meta-learning methods to enhance model generalization and efficiency.
  • Improve existing Nace.AI models by incorporating advancements from recent ML research.

Requirements

  • Hands-on experience training and fine-tuning large language models (LLMs) and vision-language models (VLMs), including practical work with pre-training, instruction tuning, and alignment techniques (GRPO,RLHF/DPO/PPO).
  • Hands-on Experience with Deep Learning Models, especially Transformers.
  • Ability to translate cutting-edge research from papers into clean, production-ready code (Paper to Code).
  • Proven experience scaling inference infrastructure for LLMs/VLMs, including expertise in model serving frameworks like vLLM, TGI.
  • Proficient in Python with a strong track record of building substantial projects.
  • Solid foundation in computer science fundamentals (data structures, algorithms, design patterns).
  • BS degree in CS or related technical field.
  • Solid Experience with ML frameworks and libraries (PyTorch, TensorFlow).
  • Self-starter comfortable working in a fast-paced, dynamic environment., * MS/PhD in CS or related technical field.
  • Familiarity with data processing stacks such as Spark and Airflow.
  • Experience with multi-node GPU training.
  • Contributor to open-source ML projects.
  • Deep knowledge in Linear Programming.
  • Experience with advanced NLP and Multimodal post-training experience (e.g., model distillation, quantization, deployment optimization).
  • Experienced in inference time optimization, deep understanding of LLM serving optimizations for LLMs/VLMs.
  • Hands on experience with quantization techniques (AWQ, GPTQ, FP8/GGUF).

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

1:39 min

Fundamentals of tensors and the TensorFlow library

Håkan Silfvernagel · LIVE

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

3:53 min

Architecting machine learning projects with the PAI platform

Qiyang Duan · LIVE

3:55 min

Evaluating central server APIs against edge deployment models

Hauke Brammer · World Congress 2021

Videos

See all

Related articles

See all