> Markdown version of [/jobs/ext/2913335-software-engineer-inference](https://www.wearedevelopers.com/jobs/ext/2913335-software-engineer-inference). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Software Engineer - Inference - **Company:** Visionist, Inc - **Location:** Laurel, MD, United States - **Salary:** $85,000.0 - $120,000.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Continuous Integration, Python (Programming Language), Octopus Deploy, Cloud Services, Large Language Models, Kubernetes, Programming Languages - **Published:** September 15, 2026 - **Apply:** https://www.clearancejobs.com/jobs/9163105/software-engineer-inference ## About the Role Bachelor's degree in a technical discipline. (Additional 4 years of experience may substitute degree) - Experience with Python and/or other modern programming languages - Experience with Kubernetes/Helm - Strong communication skills and willingness to ask questions - Ability to learn new technologies quickly - Familiarity with AWS or other cloud service providers - Familiarity with Argo CD and/or other CI/CD frameworks ## Description Procure, configure, and test new inference models, preparing them for release to our user base - Develop in-house services and techniques to guarantee continual high-quality inference service for our customer - Work with model vendor teams and representatives to create reliable pipelines for closed-source model usage - Collaborate with teammates on surge efforts to support short-term, high-priority inference needs from our customer - Engage with other teams in our organization to establish solid infrastructure for our services and integrate LLM-powered tools for user needs ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Leverage Cloud Computing Benefits with Serverless Multi-Cloud ML ](https://www.wearedevelopers.com/videos/78-leverage-cloud-computing-benefits-with-serverless-multi-cloud-ml) - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [Machine Learning for Software Developers (and Knitters)](https://www.wearedevelopers.com/videos/154-machine-learning-for-software-developers-and-knitters) - [ZEISS & Microsoft - Building the Next Generation Medical Ecosystem in the Cloud](https://www.wearedevelopers.com/videos/424-zeiss-microsoft-building-the-next-generation-medical-ecosystem-in-the-cloud) - [Instant KAI Sandboxes with vCluster: Multi-Tenant, Multi-Scheduler GPU Sharing](https://www.wearedevelopers.com/videos/100333-instant-kai-sandboxes-with-vcluster-multi-tenant-multi-scheduler-gpu-sharing) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models)