> Markdown version of [/jobs/ext/3556696-machine-learning-performance-engineer](https://www.wearedevelopers.com/jobs/ext/3556696-machine-learning-performance-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Machine Learning Performance Engineer - **Company:** Fintal Partners - **Location:** Stone Park, IL, United States - **Contract:** Permanent contract - **Skills:** Abstraction Layers, Artificial Intelligence, Artificial Neural Networks, C++ (Programming Language), Field-Programmable Gate Array (FPGA), High-Frequency Trading, Python (Programming Language), Machine Learning, Inference Optimization, Tensorflow, Signal Processing, SystemVerilog, VHSIC Hardware Description Language (VHDL), Pytorch, Discretization, Machine Learning Operations, MLIR (Multi-Level Intermediate Representation) - **Published:** October 2, 2026 - **Apply:** https://www.disabledperson.com/jobs/75658732-machine-learning-performance-engineer ## About the Role * Understanding of hardware constraints and design trade-offs (pipelining, resource utilization, fixed-point arithmetic) that shape how ML models map onto FPGAs or custom ASICs * Experience with hardware fundamentals, whether through VHDL/SystemVerilog development, HLS tools, or ML-to-hardware frameworks like hls4ml, FINN, or Vitis AI * Understanding of machine learning fundamentals: neural network architectures, inference optimization, quantization techniques, and ML frameworks such as PyTorch/TensorFlow * Proficiency in Python, C++, or similar languages for tooling, testing, and simulation * Strong communication skills and the ability to work collaboratively across disciplines with both technical and non-technical teams * An advanced degree (MS or PhD) in EE, CS, Physics, or a related field Particularly Relevant Experience Exposure to ML compiler infrastructure such as MLIR, TVM, or XLA; a background in latency-sensitive or resource-constrained systems including high-frequency trading, particle physics data acquisition, or real-time signal processing; familiarity with functional verification methodologies such as SystemVerilog, UVM, or Cocotb. Trading experience is a bonus, not a prerequisite. The firm is looking for researchers and engineers from any background who want to push the boundaries of what's computationally possible. ## Description A leading global trading firm is expanding its machine learning capabilities and looking for an experienced Hardware Machine Learning Engineer to help deploy ML directly onto custom hardware. This firm builds its own hardware, software, and infrastructure in-house, so when a model hits a latency wall or a resource ceiling, the engineer doesn't file a ticket and wait, they go fix it. There's no vendor to wait on and no abstraction layer they're not allowed to touch. This is a rare opportunity to architect solutions from scratch, influence technical research direction, and see the work drive real impact in one of the most demanding computing environments in the world. What You'll Work On * Architect and co-design ML models with traders, quant researchers, and software engineers, treating hardware constraints like latency budgets, resource limits, and numerical precision as first-class design inputs * Shape the custom hardware roadmap by translating ML model requirements into concrete architectural decisions * Work hands-on with hardware engineers to implement, verify, and deploy ML inference solutions from proof-of-concept through production * Track and evaluate emerging research in neural architecture search, machine learning systems, and quantization methods, and determine what translates to measurable improvements ## Related Videos - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Efficient deployment and inference of GPU-accelerated LLMs​](https://www.wearedevelopers.com/videos/929-efficient-deployment-and-inference-of-gpu-accelerated-llms) - [Machine learning in the browser with TensorFlowjs](https://www.wearedevelopers.com/videos/155-machine-learning-in-the-browser-with-tensorflowjs) - [Geometric deep learning for drug discovery](https://www.wearedevelopers.com/videos/264-geometric-deep-learning-for-drug-discovery) - [From Model to Metal: An Open Source Stack for Accelerating Intelligence](https://www.wearedevelopers.com/videos/1636-from-model-to-metal-an-open-source-stack-for-accelerating-intelligence) - [Getting Started with Machine Learning](https://www.wearedevelopers.com/videos/260-getting-started-with-machine-learning) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLops – Deploying, Maintaining And Evolving Machine Learning Models in Production](https://www.wearedevelopers.com/magazine/115-mlops-deploying-maintaining-and-evolving-machine-learning-models-in-production) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [How machine learning can help us tell fact from fiction](https://www.wearedevelopers.com/magazine/509-how-machine-learning-can-help-us-tell-fact-from-fiction)