System Software Engineer for Kubernetes Fabric Integration

Cornelis Networks
San Jose, CA, United States
15 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Computing Platforms C++ (Programming Language) Cloud Engineering Code Generation Code Review Computer Programming Computer Engineering Continuous Integration Monitoring of Systems Python (Programming Language) Network Control
+19 more
Open Source Technology Performance Tuning Prometheus Software Engineering System Programming System Testing Systems Integration Multithreading Application Enhancement Tool Fluentd Delivery Pipeline Grafana Data Center Networking Containerization Kubernetes Information Technology Hardware Infrastructure Code Restructuring Docker

Job description

  • Architect and Design:Lead the design of robust, scalable solutions for integrating Cornelis Networks’platform and fabric management softwarewith Kubernetes.

  • Develop Kubernetes Operators:Build and maintain custom Kubernetes Operators and Controllers in Go to manage the lifecycle of our software and hardware components within a cluster.

  • Cloud-Native Integration:Develop solutions that allow for the seamless orchestration of our high-performancefabric services and platform management toolsalongside other containerized workloads.

  • Comprehensive Testing: Develop end-to-end automated validation frameworks to stress-test the product.
  • Cluster Management:Work on extending Kubernetes for managing specialized hardware, scheduling, and networking requirements unique to HPC and AI workloads.

  • Collaborate:Partner with the core platform, fabric, and hardware teams to ensure a cohesive and performant end-to-end solution.

  • Upstream Contribution:Engage with the open-source community and contribute to relevant projects within the cloud-native ecosystem.

  • Documentation and Best Practices:Author high-quality technical documentation and champion best practices for software development in a cloud-native environment.

  • Leverage AI-powered tools to accelerate software development workflows, including intelligent code generation, refactoring, and performance optimization.

  • Apply AI-driven techniques for automated code review, testing, and quality assurance to improve reliability and reduce development cycles.

Requirements

  • 5+ years of professional software development experience.

  • Proven experience in designing and developing solutions for Kubernetes, including building custom operators/controllers using tools like the Operator SDK or Kubebuilder.

  • Expert level proficiency in usingGo for systems programming. Experience with C++ or Python is also valuable.
  • Knowledge of Multithreaded Programming.
  • Deep understanding of Kubernetes architecture, including the control plane, networking (CNI), and storage (CSI) interfaces.

  • Hands-on experience with container technologies such as Docker or containerd.

  • Demonstrable experience in integrating existing software platforms or services with Kubernetes.
  • Solid understanding of high-performance data center networking environment.
  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related technical field.

Preferred Qualifications:

  • Experience with high-performance computing (HPC) or high-performance networking.

  • Familiarity with performance-sensitive environments and low-latency application requirements.

  • Experience with monitoring and observability stacks like Prometheus, Grafana, and Fluentd.

  • Knowledge of CI/CD principles and experience building deployment pipelines.

  • Contributions to open-source projects in the Kubernetes or cloud-native ecosystem.

Job Location: This role may work remote from Costa Rica via an EOR (Employer of Record).

Benefits & conditions

We offer a competitive compensation package that includes base salary, performance incentives and equity participation.

At Cornelis Networks, your compensation is only one component of your comprehensive total rewards package. Compensation will be determined by factors such as experience, qualifications, skills, and location relative to the hiring range for the position.

Depending on your geographic location, in addition to your total rewards package, you may also be eligible for additional benefits, paid holidays, and flexible work arrangements.

About the company

At Cornelis we’re building the future of AI and HPC networking with an AI-first approach to silicon and software development. We’re seeking engineers who are energized by working on cutting-edge ASIC design and distributed software systems, and who are motivated to push the boundaries on how AI can transform everything from chip architecture to system performance at scale.

Cornelis Networks delivers the world’s highest performance scale-out networking solutions for AI and HPC datacenters. Our differentiated architecture seamlessly integrates hardware, software and system level technologies to maximize the efficiency of GPU, CPU and accelerator-based compute clusters at any scale. Our solutions drive breakthroughs in AI & HPC workloads, empowering our customers to push the boundaries of innovation. Backed by top-tier venture capital and strategic investors, we are committed to innovation, performance and scalability - solving the world’s most demanding computational challenges with our next-generation networking solutions.

We are a fast-growing, forward-thinking team of architects, engineers, and business professionals with a proven track record of building successful products and companies. As a global organization, our team spans multiple U.S. states and six countries, and we continue to expand with exceptional talent in onsite, hybrid, and fully remote roles.

Cornelis Networks is seeking a talented and experienced Senior System Software Engineer for Kubernetes Fabric Integration to build the bridge between our high-performance interconnect hardware and the cloud-native orchestration ecosystem. In this role, you will be extending the capabilities of Kubernetes to manage highly specialized large scale network fabrics. You will design, develop and test advanced fabric management software that will be deployed in massive AI and HPC clusters. You will design, build, and maintain Kubernetes operators, controllers, and other components necessary to ensure our high-performance interconnect solutions can be seamlessly deployed, managed, and scaled in containerized environments. This is a critical role that will directly impact our customers’ ability to leverage Cornelis Networks’ technology in large-scale, modern data centers.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:54 min

Speaker background and open source Kubernetes edge computing projects

Gaurav Gahlot Gaurav Gahlot · World Congress 2026 Europe

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

1:04 min

Visualizing Keycloak performance via standard Grafana troubleshooting dashboards

Alexander Schwartz Alexander Schwartz · World Congress 2025

4:12 min

Utilizing pre-integrated observability and authentication platform tools

Aarno Aukia · LIVE

Videos

See all

Related articles

See all