> Markdown version of [/jobs/ext/2677134-backend-engineer_ai-gateway-korean-required](https://www.wearedevelopers.com/jobs/ext/2677134-backend-engineer_ai-gateway-korean-required). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Backend Engineer_AI Gateway (Korean Required) - **Company:** SBT Global, Inc. - **Location:** Plano, TX, United States - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Encodings, Communications Protocols, Databases, Continuous Integration, Dependency Injection, Github, Hypertext Transfer Protocols (HTTP), Identity and Access Management, Python (Programming Language), PostgreSQL, Nginx, OpenID, Query Optimization, Queueing Systems, RabbitMQ, Redis, Security Assertion Markup Language (SAML), Service Design, Data Streaming, Systems Integration, TCP/IP, Web Application Frameworks, Openapi, Large Language Models, Backend, Fastapi, Containerization, AI Platforms, Gitlab-ci, Kubernetes, Apache Kafka, Terraform, Docker - **Published:** September 2, 2026 - **Apply:** https://www.wayup.com/i-j-Backend-Engineer-AI-Gateway-Korean-Required-SBT-Global-Inc-984382264639123/ ## About the Role + Experience: Around 5 years of professional backend development experience, with a solid track record of building and operating production-grade services. + Async Python & Core Web Frameworks (3+ yrs): Deep understanding of asyncio and hands-on production experience with FastAPI (Pydantic, dependency injection, OpenAPI). + Network Protocol & Streaming Fundamentals: o Strong grasp of HTTP/1.1, HTTP/2, and SSE (Server-Sent Events) / Chunked transfer encoding essential for real-time LLM inference streaming. o Practical understanding of the TCP/IP lifecycle, connection pooling, keep-alive, and managing back-pressure. + Database & Async Messaging: o Proficiency in PostgreSQL (query tuning, connection pooling via PgBouncer). o Experience with message queues or event streams (e.g., Redis Streams, RabbitMQ, or Kafka) for decoupled task processing. + Containerization & CI/CD: o Docker, Kubernetes (basic workload management, deployment, HPA), and setting up CI/CD pipelines (GitHub Actions / GitLab CI). Work experience desired + Collaborative Security & Infra Awareness: Willingness to work with internal security/network teams on NGINX/Envoy reverse-proxy setup, mTLS, or IAM/SSO (OIDC/SAML) integrations. + AI/LLM Integration: Experience handling LLM API payloads, managing rate limits, or implementing simple proxy/gateway patterns for AI services. + Infrastructure as Code: Basic familiarity with Terraform for multi-env infrastructure management. ## Description We are building an AI Gateway-a high-performance intermediary that routes, inspects, and governs traffic between enterprise clients and LLM/AI service providers. In this role you will own the entire full-stack lifecycle of the gateway, from core Python API service design and implementation to containerization, deployment, monitoring, and day-to-day operations. While you can consult network and security specialists for architecture reviews, threat modeling, and policy definition, the application code, testing, CI/CD pipeline, and production operations will be developed and maintained solely by you. You'll have full technical autonomy to design, build, and scale a robust, production-grade system. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Post-Quantum Cryptography: Preparing for Q-Day](https://www.wearedevelopers.com/videos/100179-post-quantum-cryptography-preparing-for-q-day) - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) - [Streaming AI Responses in Real-Time with SSE in Next.js & NestJS](https://www.wearedevelopers.com/videos/1630-streaming-ai-responses-in-real-time-with-sse-in-next-js-nestjs) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [The 7 Most Popular Backend Frameworks for Developers](https://www.wearedevelopers.com/magazine/403-the-7-most-popular-backend-frameworks-for-developers) - [Dev Digest 132 - Binging WADFlix?](https://www.wearedevelopers.com/magazine/473-dev-digest-132-binging-wadflix) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud)