> Markdown version of [/jobs/ext/2211043-software-engineer-akamai-inference-cloud-remote-poland](https://www.wearedevelopers.com/jobs/ext/2211043-software-engineer-akamai-inference-cloud-remote-poland). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Software Engineer - Akamai Inference Cloud - Remote/Poland - **Company:** Akamai Technologies - **Location:** Cambridge, MA, United States (Remote available) - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, C++ (Programming Language), Cloud Computing, Profiling, Code Review, Encodings, Software Debugging, Linux, Text Processing, Python (Programming Language), Natural Language Processing, Akamai, System Programming, Tokenization, Data Processing, Large Language Models, Containerization, AI Platforms, Machine Learning Operations, TensorRT - **Published:** August 24, 2026 - **Apply:** https://arc.dev/remote-jobs/j/redirect/peccf7ud9y ## About the Role * Have relevant experience in related field. * Demonstrate a solid understanding of Python and familiarity with at least one systems programming language such as C++, Go, or Rust. * Show understanding of natural language processing concepts including tokenization, encoding, and text processing pipelines. * Have familiarity with AI inference, model serving, or LLM deployment including inference frameworks (TensorRT, vLLM, TorchServe, Triton). * Demonstrate experience working on or contributing to high-throughput, low-latency data processing systems or services. * Show familiarity with Linux systems, containerized environments, and profiling or debugging tools. * Demonstrate a keen willingness to learn and grow within the AI inference and model serving field. ## Description Are you excited about building the systems that process and optimize AI inference requests at scale? Do you want to work hands-on with cutting-edge AI serving technologies and large language models? Join the Akamai Inference Cloud Team! The Akamai Inference Cloud team develops AI platforms for inference models and applications within Akamai's Cloud Technology Group. The Inference Execution & Runtimes team manages the inference engine, AI runtime frameworks, model serving infrastructure, and execution environment. Their work impacts latency, throughput, and efficiency of AI workloads across Akamai's extensive global infrastructure. Partner with the best As a Software Engineer on the Inference Execution & Runtimes team, contribute to prompt processing, tokenization, and request handling pipelines. Collaborate with experienced engineers to transform user requests into optimized model inputs. This role offers significant exposure to AI inference systems and opportunities to gain expertise in a rapidly evolving infrastructure engineering field. As a Software Engineer, You Will Be Responsible For * Developing and maintaining prompt processing and tokenization pipelines that prepare inference requests for efficient model execution. * Implementing request routing, scheduling, and batching strategies to optimize throughput and latency across simultaneous inference workloads effectively. * Contributing to the integration of new model architectures and serving backends into the runtime framework. * Writing well-tested, well-documented code and participating in code reviews to maintain high engineering standards across the inference stack. * Supporting operational readiness through monitoring, debugging, and performance analysis of runtime components. ## Related Videos - [Tour de Force: Open-Source LLM Inference Optimization from Simple to Sophisticated](https://www.wearedevelopers.com/videos/100099-tour-de-force-open-source-llm-inference-optimization-from-simple-to-sophisticated) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [What the Heck is Edge Computing Anyway?](https://www.wearedevelopers.com/videos/593-what-the-heck-is-edge-computing-anyway) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Efficient deployment and inference of GPU-accelerated LLMs​](https://www.wearedevelopers.com/videos/929-efficient-deployment-and-inference-of-gpu-accelerated-llms) - [Developing an AI.SDK](https://www.wearedevelopers.com/videos/198-developing-an-ai-sdk) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development)