CAREERS AT NETRIS

Shape the AI buildout

Join the team shaping the networking foundation behind the world's most demanding AI clouds.

Senior DevOps Engineer

About Netris

Netris is the leading network automation platform for AI clouds. AI is the largest infrastructure buildout in history, and networking is its bottleneck. Netris is the automation and multi-tenancy layer for AI cloud networking, enabling operators to launch GPU clouds in weeks, provision tenants instantly, and maximize GPU utilization. The company experienced 800% ARR growth in the last 12 months, with 35+ live GPU cluster deployments globally and growing, and is backed by Andreessen Horowitz. SDN reinvented data center networking, and Netris is doing it for the AI cloud with NAAM: Network Automation, Abstraction, and Multi-Tenancy.

The role

We are looking for a Senior DevOps/Integration Engineer to join the Netris Customer Success team and work directly with customers to deploy and integrate Netris across GPU cloud networking environments. This is a highly technical, customer-facing role that requires deep hands-on expertise in high-availability Kubernetes cluster administration, Linux systems and networking, Kubernetes operators, and Terraform.

What you’ll do

  • Work directly with customers to deploy and maintain the Netris controller in GPU cloud networking environments.
  • Deploy and operate high-availability Kubernetes clusters in customer environments.
  • Help customers perform ongoing maintenance, upgrades, and lifecycle management of Kubernetes-based Netris deployments.
  • Help customers troubleshoot issues across Kubernetes, Linux systems, containers, and Linux networking.
  • Help customers deploy, maintain, and troubleshoot Kubernetes operators and related platform components.
  • Use RestAPI, Terraform and Infrastructure as Code to automate and standardize customer deployments.
  • Investigate production issues, identify root causes, and work with Netris Engineering when deeper product-level troubleshooting is required.
  • Improve deployment and maintenance procedures based on real-world customer environments and recurring operational issues.

What we’re looking for

  • Strong hands-on experience administering and troubleshooting production Kubernetes environments, including high-availability clusters.
  • Deep understanding of Linux systems and Linux networking.
  • Experience deploying, operating, and troubleshooting Kubernetes operators and containerized applications.
  • Strong hands-on experience with RestAPI, Terraform, and Infrastructure as Code.
  • Experience with Kubernetes upgrades, maintenance, backup/restore, and cluster lifecycle management.
  • Strong troubleshooting skills across Kubernetes, Linux, networking, and distributed systems.
  • Ability to work directly with customers, communicate clearly during complex technical issues, and take ownership through resolution.
  • Comfortable working in customer production environments where reliability, change management, and attention to detail are critical.
  • Ability to collaborate closely with Engineering teams on root-cause analysis and product-level issues.

Preferred qualifications

  • Experience operating Kubernetes in bare-metal or on-premises environments.
  • Experience with K3s, Helm, and Kubernetes packaging/deployment workflows.
  • Strong understanding of Linux networking.
  • Experience working with GPU infrastructure, AI clusters, or GPU cloud environments.
  • Experience with Kubernetes observability and monitoring tools such as Prometheus and Grafana.
  • Experience with GitOps workflows.
  • Experience supporting production infrastructure in a customer-facing engineering role.
  • Experience working in fast-paced environments where engineers take end-to-end ownership of deployments and production issues.

Compensation, benefits, logistics

  • Competitive salary +  meaningful equity
  • Full time
  • Remote
  • Medical, vision, dental, 401(k), disability insurance, flexible PTO (US; benefits vary by location)

Apply for this job