Founding Real-Time Voice Ai Performance Engineer

Slng Ltd
Barcelona, Spain
about 1 month ago
Apply on www.buscojobs.com.es
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Microsoft Word Artificial Intelligence Nvidia CUDA Programming Tools Core Voice Platform Model Validation ONNX (Open Neural Network Exchange) Format

Job description

Our mission is simple: make AI voice relevant and available for the rest of the world.We usually respond within a daySLNG is building the backbone for real-time speech AI, enabling developers to run voice applications anywhere in the world with local compliance and ultra?low latency.Our founding team comes from the core of AI and developer tooling in the USA, with experience scaling platforms trusted by the world’s best builders.Meet Luke, Founder & CEO, and Ismael, Founder & CPO at SLNG.We’re bringing the San Francisco mindset to Europe with our first hub in Barcelona, and we have the support of leading international VCs and Angel investors who are backing the SLNG journey.Challenge ? Today, Most Speech Infrastructure Is US-centric, Unreliable, And Difficult To Deploy Globally.SLNG Is Changing That By Delivering a Platform That IsLocal: deployed close to users and compliant with regional regulationsFast: designed for real?time voice experiencesOpen: built to integrate with the tools and workflows developers already use.From real?time transcription to voice AI applications, SLNG is creating the Voice AI gateway that will empower the next generation of speech?powered products.As Speech Model Performance Engineer, you’ll work closely with Ismael and the founding tech team to shape the technical foundation of SLNG.From TTS voices to multilingual ASR, you’ll benchmark, optimise, and productionize speech inference at scale.What We OfferWe pay top of the market ? We want serious talent in the team, and we benchmark compensation accordingly.Equity ? All full?time, permanent roles include stock options - we want everyone to share in the upside as we build.Hybrid by design ? 3 days/week in the Barcelona office for collaboration and culture.Flexible benefits: Health insurance, gym, and more via Cobee.L&D: Annual budget (up to €**) for training, courses, or conferences.Remote work support: Monthly stipend (up to €50) to cover wifi or other costs.Great equipment: Laptop of your choice plus €500 to set up your workstation in year one, with top?ups in the following years.Time off: Additional +5 days vacation on top of the statutory minimumAbout The RoleYou’ll Make Speech Models Fast.Example InitiativesQuantise neural TTS and STT models to run with minimal latency on heterogeneous GPU hardware.Benchmark ASR models across dialectal variations, measuring Word Error Rate (WER) and latency trade?offs.Implement continuous batching and KV-cache reuse for streaming inference.Profile GPU utilisation with CUDA kernels to identify bottlenecks in large?scale inference.RequirementsStrong background in ASR/TTS.PyTorch/ONNX experience.Familiar with GPU profiling and optimisation.Fluency in English.Hiring Process (2 weeks)Intro call with one of the foundersCase StudyPanel conversation with members of the SLNG TeamFinal discussionsJob offerThe radically global speech AI gateway for simple, universal, real?time voice.#J-***-Ljbffr

Requirements

From TTS voices to multilingual ASR, you’ll benchmark, optimise, and productionize speech inference at scale.What We OfferWe pay top of the market ? We want serious talent in the team, and we benchmark compensation accordingly.Equity ? All full?time, permanent roles include stock options - we want everyone to share in the upside as we build.Hybrid by design ? 3 days/week in the Barcelona office for collaboration and culture.Flexible benefits: Health insurance, gym, and more via Cobee.L&D: Annual budget (up to €**) for training, courses, or conferences.Remote work support: Monthly stipend (up to €50) to cover wifi or other costs.Great equipment: Laptop of your choice plus €500 to set up your workstation in year one, with top?ups in the following years.Time off: Additional +5 days vacation on top of the statutory minimumAbout The RoleYou’ll Make Speech Models Fast.Example InitiativesQuantise neural TTS and STT models to run with minimal latency on heterogeneous GPU hardware.Benchmark ASR models across dialectal variations, measuring Word Error Rate (WER) and latency trade?offs.Implement continuous batching and KV-cache reuse for streaming inference.Profile GPU utilisation with CUDA kernels to identify bottlenecks in large?scale inference.RequirementsStrong background in ASR/TTS.PyTorch/ONNX experience.Familiar with GPU profiling and optimisation.Fluency in English.Hiring Process (2 weeks)Intro call with one of the foundersCase StudyPanel conversation with members of the SLNG TeamFinal discussionsJob offerThe radically global speech AI gateway for simple, universal, real?time voice.

About the company

Our mission is simple: make AI voice relevant and available for the rest of the world.We usually respond within a daySLNG is building the backbone for real-time speech AI, enabling developers to run voice applications anywhere in the world with local compliance and ultra?low latency.Our founding team comes from the core of AI and developer tooling in the USA, with experience scaling platforms trusted by the world’s best builders. Meet Luke, Founder & CEO, and Ismael, Founder & CPO at SLNG.We’re bringing the San Francisco mindset to Europe with our first hub in Barcelona, and we have the support of leading international VCs and Angel investors who are backing the SLNG journey.Challenge ? Today, Most Speech Infrastructure Is US-centric, Unreliable, And Difficult To Deploy Globally.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:16 min

Addressing latency and architecture in voice agents

Chris Heilmann +2 · LIVE

3:58 min

Mitigating diverse browser quirks and unsupported clipboard metadata formats

Markus Over Markus Over

6:21 min

Previewing upcoming hardware acceleration capabilities for Python environments

Chris Heilmann +2 · LIVE

5:08 min

Validating requests and responses using data transfer objects

Roman Alexis Anastasini · World Congress 2021

2:59 min

Utilizing live captions for context in video meetings

Florian Margaine Florian Margaine · Europe 2026 Virtual

1:37 min

Accelerating compute with focused developer tools

Julia Koch Julia Koch +1 · World Congress 2026 Europe

Videos

See all

Related articles

See all