> Markdown version of [/jobs/ext/1183006-devops-engineer-bilingual-chinese](https://www.wearedevelopers.com/jobs/ext/1183006-devops-engineer-bilingual-chinese). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # DevOps Engineer (Bilingual Chinese) - **Company:** OMNI Solutions LLC - **Location:** United States (Remote available) - **Experience:** Experienced - **Salary:** $80,000.0 - $150,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Amazon Elastic Compute Cloud, Amazon S3, Big Data, Cloud Computing, Cloud Engineering, Nvidia CUDA, Continuous Integration, Data Centers, Linux, DevOps, Domain Name System (DNS), Github, Hypertext Transfer Protocols (HTTP), Python (Programming Language), Network Configuration and Change Management, Network Protocols, Remote Direct Memory Access, Prometheus, Webui, Shell Script, TCP/IP, AI Infrastructure, Data Storage Management, High Performance Computing, Grafana, Parallel Computation, Generative AI, Infrastructure as Code (IaC), Amazon Virtual Private Cloud (VPC), Backend, Gitlab-ci, Kubernetes, Information Technology, Machine Learning Operations, Stable Diffusion, Jenkins - **Published:** July 4, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=57dbef34abd449d2 ## About the Role * Background & Experience: Bachelor's degree or higher in Computer Science or a related field, with 3+ years of experience in DevOps or SRE. * Cloud & Containers: Proficient in core AWS services (EC2, EKS, S3, VPC); deep understanding of Kubernetes architecture and scheduling principles. * Development Skills: Must possess strong programming capabilities, proficiency in Python or Go, and hands-on experience in tool development or backend development. * Systems & Networking: Deep understanding of the Linux operating system, network protocols (TCP/IP, HTTP, DNS), and Shell scripting. * CI/CD: Highly proficient in designing pipelines using Jenkins, GitLab CI, or GitHub Actions. Bonus Points (Nice-to-Haves) * AI/GPU Operations: Experience in large-scale GPU cluster operations, familiar with GPU memory monitoring, resource partitioning, or Spot instance cost optimization. * AIGC Hands-on Experience: Practical experience with generative AI, familiarity with the deployment architecture of ComfyUI and Stable Diffusion WebUI, and experience in resolving dependency management, multi-user concurrency, or model loading acceleration. * MLOps/AIOps: Familiarity with Kubeflow, MLflow, or Triton Inference Server is a strong plus. * High-Performance Computing (HPC): Experience with RDMA networking or large-scale data parallel processing. ## Description The company is seeking a highly skilled and passionate DevOps Engineer to join our cutting-edge research and development team. In this role, you will bridge the gap between traditional cloud infrastructure and the rapidly evolving world of Generative AI. You will be responsible for scaling global containerized environments, automating complex workflows, and building the foundational infrastructure that powers our next-generation AI tools., * Cloud-Native Architecture: Responsible for the planning and management of AWS cloud and private data centers, implementing Infrastructure as Code (IaC). * Container Orchestration: Maintain large-scale Kubernetes (EKS) clusters, handling cluster upgrades, scaling, network configuration (CNI), and storage management. * DevOps Development: Develop automated operations tools, CLIs, or platforms using Python or Go to eliminate repetitive tasks and improve R&D efficiency. * Observability: Build full-link monitoring and logging systems based on Prometheus, Grafana, and ELK to ensure system stability and rapid troubleshooting. * AI Infrastructure: * Manage GPU server resources (Nvidia A100/T4, etc.), and optimize driver versions and CUDA environments. * Responsible for the containerized deployment and concurrency optimization of generative AI tools such as ComfyUI. ## Related Videos - [An Applied Introduction to eBPF with Go](https://www.wearedevelopers.com/videos/1075-an-applied-introduction-to-ebpf-with-go) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Turning Container security up to 11 with Capabilities](https://www.wearedevelopers.com/videos/718-turning-container-security-up-to-11-with-capabilities) - [DevOps for AI: running LLMs in production with Kubernetes and KubeFlow](https://www.wearedevelopers.com/videos/1222-devops-for-ai-running-llms-in-production-with-kubernetes-and-kubeflow) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) ## Related Articles - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers)