MLOps & AI Infrastructure Engineer
NXAI
MLOps & AI Infrastructure EngineerNXAI
Linz, Austria
4 days ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on devjobs.at
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source
Tech stack
Artificial Intelligence
Computer Clusters
Continuous Integration
Linux
Distributed Systems
Github
General-Purpose Computing on Graphics Processing Units
Software Deployment
AI Infrastructure
Scripting
Pytorch
IT Architecture
+6 more
Reliability of Systems
Containerization
Kubernetes
Slurm
Machine Learning Operations
Docker
Job description
- As we expand our capabilities and deploy models at massive scale, we are looking for an MLOps Engineer (f/m/d) to design, scale and orchestrate the high-performance GPU clusters powering the next generation of AI.
- In this role, you will shape the future of NXAI’s AI infrastructure by:
- Scaling AI Architecture: Designing, operating and automating our high-performance infrastructure across cloud, HPC and on-premise environments.
- Empowering Research: Collaborating with Research and Applied Research teams to translate cutting-edge AI models into scalable, production-ready systems.
- Optimizing GPU Clusters: Managing and fine-tuning large-scale GPU compute, storage and networking resources for intensive workloads.
- Driving MLOps & Orchestration: Building robust machine learning pipelines and managing orchestration platforms using Kubernetes and Slurm.
- Ensuring Excellence: Implementing modern CI/CD workflows and maintaining the highest standards for system reliability, security, and observability., * Deep experience with Linux-based production environments, distributed systems, and containerization (Docker/Podman)., * The Mission & Impact: Shape the Future of AI: The chance to contribute to technology that has the potential to redefine how AI systems are built, trained and deployed.
- Challenging the Status Quo: Work directly with leading AI researchers, including the creators of xLSTM, on technologies that challenge the current transformer paradigm.
- A Key Position: A central role in enabling foundation model research, industrial AI applications and efficient AI systems that can run from cloud-scale infrastructure down to edge devices.
- Ownership & Autonomy: Significant ownership, autonomy and the ability to shape our infrastructure, tooling and engineering culture within a highly talented, research-driven team.
Requirements
- A hands-on problem solver passionate about building resilient systems that accelerate AI deployment.
- Proven track record with workload schedulers (Slurm), Kubernetes, CI/CD pipelines (e.g., GitHub Actions) and Python scripting.
- Hands-on expertise with GPU computing environments, storage optimization and building/monitoring MLOps pipelines., * Bonus Points: Experience with modern AI frameworks (like PyTorch) and supporting large-scale model training environments.
About the company
- Based in Linz, Austria, NXAI is a leading center for European AI innovation, turning advanced research into real-world industrial applications.
- Focused on next-generation xLSTM architectures, we offer companies a sovereign, on-premise alternative to US hyperscalers.
- A prime example is our zero-shot time series foundation model, TiRex, which delivers industrial enterprises a true technological edge - operating up to 50× more efficiently than traditional transformer models.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on devjobs.at
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
BB
Benedikt Bischof
almost 4 years ago
BB
Benedikt Bischof
MLOps And AI Driven Development
over 4 years ago
BB
Benedikt Bischof
MLops – Deploying, Maintaining And Evolving Machine Learning Models in Production
about 4 years ago
LM
Luis Minvielle
How to Become an AI Engineer
almost 3 years ago
LM
Luis Minvielle
What Are Large Language Models?
almost 3 years ago
CH
Chris Heilmann
Dev Digest 120 - Apple and peers
over 2 years ago