Senior Solutions Architect - Diffusion AI Models

NVIDIA Corporation
Valbonne, France
24 days ago
Apply on nvidia.wd5.myworkdayjobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Computer Vision Codecs Data Transformation High-Level Architecture Machine Learning Information Technology TensorRT Stable Diffusion Nim (Programming Language)

Job description

We are seeking a Senior Solutions Architect with deep expertise in large-scale training and inference optimization of computer vision models such as diffusion based. Working alongside leading AI Native companies building the next generation of image, video, and multimodal experiences, you will serve as the hands-on technical bridge between NVIDIA’s generative AI platform and the industry’s most demanding production pipelines. Your work will help customers unlock the full potential of NVIDIA’s accelerated computing stack, driving measurable gains in performance, scalability, and efficiency across the full content generation lifecycle, from diffusion model training and optimization to image/video synthesis pipelines and multimodal foundation world models.

What you will be doing:

  • Guide EMEA AI Native companies building image, video, and multimodal generation products in training and deploying their pipelines on NVIDIA infrastructure.
  • Provide deep technical guidance on diffusion model architectures (DiT, UNet, flow matching) and their efficient deployment across single and multi-GPU environments.
  • Optimize generation pipelines for
  • Guide customers through the full visual content generation stack: codec-aware preprocessing, temporal consistency, video token representation, efficient long-video inference.
  • Identify performance bottlenecks specific to vision workloads: memory-bound diffusion steps, attention scaling with resolution, and multi-GPU communication patterns for video.
  • Translate customer’s insights into actionable product feedback for NVIDIA’s research and engineering teams.
  • Contribute to the EMEA developer community through technical demos, workshops, and reference demos that showcase what is possible on NVIDIA stack.

Requirements

  • MS or PhD in Computer Science, Computer Vision, Machine Learning, or equivalent hands-on experience.
  • 5+ years in AI/ML with deep expertise Computer Vision models.
  • Experienced with diffusion model frameworks for image/video generation.
  • Understanding of vision encoder optimization, VAE architectures, and their performance tradeoffs at inference time.
  • Strong communication skills, effective with ML researchers, creative technologists, and infrastructure engineers alike.

Ways to stand out from the crowd:

  • Familiarity with NVIDIA’s inference stack: TensorRT, Triton Inference Server, and NIM.
  • Hand-son experience on video generation: Temporal attention, 3D convolutions, or causal video transformers.
  • Familiarity with codec-aware video pipelines and efficient video tokenization for generation at scale.
  • Published work or benchmarks in image/video generation, diffusion acceleration, or visual foundation models.

Benefits & conditions

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer www.nvidiabenefits.com, Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. For Poland: The base salary range is 292,500 PLN - 507,000 PLN.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on nvidia.wd5.myworkdayjobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · World Congress 2026 Europe

4:52 min

Essential phases in building and refining language models

Anshul Jindal Anshul Jindal +1 · World Congress 2025

5:55 min

Practical applications and use cases for computer vision

Flo Pachinger · LIVE

1:07 min

Hypothetical video codecs and AI agent services

Andrew MacLean Andrew MacLean +2 · LIVE

3:22 min

Evaluating advanced artificial intelligence platforms for daily recruitment

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

51 sec

Overcoming inefficiencies in computer vision modeling

Antonio Tavera Antonio Tavera · World Congress 2025

Videos

See all

Related articles

See all