Software Engineer - AI Systems

ProntoPro
Barcelona, Spain
1 day ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
2 years minimum
Working hours
Regular working hours
Job source

Tech stack

Adobe InDesign Application Programming Interfaces (APIs) Artificial Intelligence C++ (Programming Language) Profiling Code Review Continuous Integration Data Structures Software Debugging Memory Management Python (Programming Language) Machine Learning
+15 more
OpenCV Performance Tuning Tensorflow Software Engineering System Programming Pytorch Large Language Models Concurrency Caching Generative AI ONNX (Open Neural Network Exchange) Format HuggingFace Machine Learning Operations Stream Processing C++ Frameworks

Job description

Location: Barcelona, Spain (Remote) Job Type: Full-time About the role We are seeking a Software Engineer to join our high-performance team collaborating with leading AI chip companies. Our work focuses on developing software that enables users to run Vision and Generative AI inference workloads efficiently on custom accelerators. This is a hands-on engineering role where you will contribute to the development of frameworks, APIs, and runtime integrations that power AI models on next-generation hardware. You will work alongside senior engineers and coordinate with compiler/runtime and hardware architecture teams. This is not a traditional applied ML position. It suits someone who wants to build a strong foundation in AI/ML systems, performance optimization, and software engineering, while contributing to production-grade AI enablement. Key Responsibilities - Assist in implementing scalable software architecture and design patterns. - Help develop Python/C++ frameworks that integrate

Requirements

ML models with custom runtimes. - Contribute to building high-performance APIs, bindings, and libraries for Vision and Generative AI inference. - Support Model Zoo maintenance, model loaders, and optimization workflows for easier deployment. - Profile, debug, and optimize performance-critical sections in framework and runtime layers. - Contribute to real-time pipelines using GStreamer, OpenCV, and related frameworks. - Collaborate with compiler/runtime teams on graph-level and operator optimizations. - Apply best practices in design, testing, CI/CD, and code reviews. - Participate in end-to-end software delivery, including defining scope and meeting project timelines. Requirements - 2 - 3 years of professional or internship experience in software engineering, preferably in systems programming or performance-related domains. - Proficiency in C++ (C++11/14) and Python. - Familiarity with C++/Python bindings such as pybind11 or SWIG. - Strong understanding of Data structures and algorithms; Concurrency, threading, and synchronization; Memory management, caching, and performance profiling; Networking and streaming systems. - Some exposure to ML frameworks such as PyTorch, TensorFlow, or ONNX Runtime and their integration with hardware runtimes. - Interest in building frameworks, SDKs, or toolchains used by other developers. Bonus Points - Experience working with Vision or Generative AI models such as transformers, diffusion models, or LLM inference. - Familiarity with multimedia and vision pipelines (for example, GStreamer). - Contributions to open-source projects or personal ML systems experiments (for example, ONNX Runtime or Hugging Face). - Strong motivation to learn and work in a collaborative technical environment. Why Join 10xEngineers - Work with a world-class chip company on state-of-the-art AI systems. - Build a solid foundation in AI software and performance engineering. - Continuous exposure to Vision and

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on es.trabajo.org

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

3:15 min

Reversing the caching model for artifact delivery

Thijs Feryn Thijs Feryn · World Congress 2026 Europe

4:53 min

Achieving real-time tracking performance with OpenCV and segmentation

Thomas Endres Thomas Endres +2 · World Congress 2021

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · World Congress 2026 Europe

Videos

See all

Related articles

See all