Software Engineer - Akamai Inference Cloud - Remote/Poland

Akamai Technologies
Cambridge, MA, United States
13 days ago
Apply on arc.dev
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence C++ (Programming Language) Cloud Computing Profiling Code Review Encodings Software Debugging Linux Text Processing Python (Programming Language) Natural Language Processing Akamai
+8 more
System Programming Tokenization Data Processing Large Language Models Containerization AI Platforms Machine Learning Operations TensorRT

Job description

Are you excited about building the systems that process and optimize AI inference requests at scale?

Do you want to work hands-on with cutting-edge AI serving technologies and large language models?

Join the Akamai Inference Cloud Team!

The Akamai Inference Cloud team develops AI platforms for inference models and applications within Akamai’s Cloud Technology Group. The Inference Execution & Runtimes team manages the inference engine, AI runtime frameworks, model serving infrastructure, and execution environment. Their work impacts latency, throughput, and efficiency of AI workloads across Akamai’s extensive global infrastructure.

Partner with the best

As a Software Engineer on the Inference Execution & Runtimes team, contribute to prompt processing, tokenization, and request handling pipelines. Collaborate with experienced engineers to transform user requests into optimized model inputs. This role offers significant exposure to AI inference systems and opportunities to gain expertise in a rapidly evolving infrastructure engineering field.

As a Software Engineer, You Will Be Responsible For

  • Developing and maintaining prompt processing and tokenization pipelines that prepare inference requests for efficient model execution.
  • Implementing request routing, scheduling, and batching strategies to optimize throughput and latency across simultaneous inference workloads effectively.
  • Contributing to the integration of new model architectures and serving backends into the runtime framework.
  • Writing well-tested, well-documented code and participating in code reviews to maintain high engineering standards across the inference stack.
  • Supporting operational readiness through monitoring, debugging, and performance analysis of runtime components.

Requirements

  • Have relevant experience in related field.
  • Demonstrate a solid understanding of Python and familiarity with at least one systems programming language such as C++, Go, or Rust.
  • Show understanding of natural language processing concepts including tokenization, encoding, and text processing pipelines.
  • Have familiarity with AI inference, model serving, or LLM deployment including inference frameworks (TensorRT, vLLM, TorchServe, Triton).
  • Demonstrate experience working on or contributing to high-throughput, low-latency data processing systems or services.
  • Show familiarity with Linux systems, containerized environments, and profiling or debugging tools.
  • Demonstrate a keen willingness to learn and grow within the AI inference and model serving field.

About the company

At Akamai, we make life better for billions of people, trillions of times a day.

Whether you’re streaming live events, scrolling social media, watching your favorite series, or managing your savings, we’re the engine behind the scenes. We provide the world’s most distributed platform from Cloud to Edge to help the giants of the digital world work faster and stay more secure, making the internet a better experience for everyone.

Our Focus Is Simple

Cloud and Edge: Running apps closer to users for instant performance.

Security: Neutralizing threats before they ever reach your data.

Content Delivery: Scaling the world’s biggest moments without a glitch.

AI: Enabling our customers to build, secure, and scale AI apps on the world’s most distributed cloud platform.

At Akamai, we don’t just support the internet; we power and protect it, because behind every great digital experience is a massive hidden challenge. And we’re the ones who solve it. When millions of people hit play or pay, Akamai ensures it just works.

Benefits at Akamai: We support your health, well-being, finances, and life beyond work. See our benefits.

FlexBase adapts to your job’s needs

Akamai’s FlexBase program is yet another way we show our commitment to providing employees with an exceptional workplace experience. It’s not about telling employees where to work; it’s about supporting employees to do their best work.

We trust our incredible employees to work in ways that suit them best: at home, in an office, or a combination of both.

Connect with us on social and see what life at Akamai is like!

About Akamai Technologies

5001-10000

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on arc.dev
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:34 min

Leveraging Akamai edge workers for broad geographic scale

Austin Gil · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

4:52 min

Essential phases in building and refining language models

Anshul Jindal Anshul Jindal +1 · World Congress 2025

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

2:32 min

Core libraries driving inference engines and multi-GPU networking

Adolf Hohl Adolf Hohl · World Congress 2024

2:37 min

Optimizing technical profiles for AI sourcing and recruitment

Mina Golesorkhi Mina Golesorkhi · World Congress 2026 Europe

Videos

See all

Related articles

See all