Crusoe Energy Systems LLC logo

Engineering Manager, SDN Control Plane

Job Overview

Location

San Francisco, CA - US

Job Type

Full-time

Category

Engineering

Date Posted

August 8, 2026

Full Job Description

đź“‹ Description

  • • We're in the midst of the greatest industrial revolution of our time. The demand for AI compute is boundless, and power is a bottleneck. We're solving that — with an energy-first approach that makes AI infrastructure better for the world and faster for the people innovating with AI.
  • • We're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved — people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.
  • • If you want to do the most meaningful work of your career, help our customers and partners advance their AI strategies, and be part of a high-performing team that believes in each other, come build with us at Crusoe.
  • • In this hands-on leadership role, you will own the services on the network control plane that program multi-tenant virtual networks across large-scale GPU fleets.
  • • Unlike generic networking manager roles, this position focuses on delivering control-plane scale and reliability on aggressive timelines, often moving features from commit to production within days.
  • • Technical Ownership: You will help define the roadmap for the VPC control plane and drive the evolution of network virtualization systems (such as OVN/OVS) that deliver VPCs, subnets, security groups, load balancing, NAT, and internet/hybrid connectivity.
  • • Architecture & Scale: You will contribute to the design of distributed control-plane services: intent APIs, state reconciliation, network policy compilation, and southbound programming of hosts and DPUs, while driving scalability well beyond current fleet size.
  • • Production Execution: You will support reliability engineering, convergence and API-latency benchmarking, regression prevention, and incident response, ensuring operational excellence within 3-6 month execution cycles.
  • • People Leadership: You will mentor and grow a team of mid-level to senior distributed systems engineers, setting technical standards and fostering a high-performance culture of accountability.
  • • Collaboration: You will partner closely with data-plane teams (XDP/eBPF, DPDK, DPU offload) and cloud product teams to deliver low-latency, highly available networking for multi-tenant GPU clusters.
  • • Solid Experience: At least 6-8+ years in distributed systems or cloud networking engineering, with 2-4+ years specifically managing engineering talent.
  • • Technical Depth: Strong knowledge of SDN and network virtualization: overlay networking (VXLAN/Geneve), VPC constructs, routing (BGP/EVPN), and control-plane architectures such as OVN/OVS or equivalent.
  • • Distributed Systems Expertise: Hands-on experience building large-scale control planes: state reconciliation, consensus and consistency trade-offs, API design, and fleet-wide configuration propagation (Go, Kubernetes-style controllers, or similar).
  • • Reliability Focus: A strong understanding of operating multi-tenant cloud services: SLOs, convergence-time and scale benchmarking, graceful degradation, and blast-radius containment.
  • • Execution Mindset: The ability to resolve complex technical challenges in a fast-moving, execution-heavy environment.

🎯 Requirements

  • • At least 6-8+ years in distributed systems or cloud networking engineering, with 2-4+ years specifically managing engineering talent.
  • • Strong knowledge of SDN and network virtualization: overlay networking (VXLAN/Geneve), VPC constructs, routing (BGP/EVPN), and control-plane architectures such as OVN/OVS or equivalent.
  • • Hands-on experience building large-scale control planes: state reconciliation, consensus and consistency trade-offs, API design, and fleet-wide configuration propagation (Go, Kubernetes-style controllers, or similar).
  • • A strong understanding of operating multi-tenant cloud services: SLOs, convergence-time and scale benchmarking, graceful degradation, and blast-radius containment.

🏖️ Benefits

  • • Competitive compensation and equity packages
  • • Restricted Stock Units
  • • Paid time off, paid holidays & leave of absence programs
  • • Comprehensive health, dental & vision insurance
  • • Employer contributions to HSA account
  • • Paid parental leave
  • • Paid life insurance, short-term and long-term disability
  • • Professional development & tuition reimbursement
  • • Mental health & wellness support
  • • Commuter benefits (parking & transit)
  • • Cell phone stipend
  • • 401(k) Retirement plan with company match up to 4% of salary
  • • Volunteer time off
  • • Global travel insurance & emergency assistance
  • • Daily meals allowance
  • • Additional perks & programs specific to location

Skills & Technologies

Kubernetes
Hybrid
$215k-260k

Ready to Apply?

You will be redirected to an external site to apply.

AI Job Fit Analysis
Pro

See exactly how your profile matches this role — strengths, skill gaps, and what to do about them.

Crusoe Energy Systems LLC logo
Crusoe Energy Systems LLC
Visit Website

About Crusoe Energy Systems LLC

Crusoe Energy Systems is building Crusoe Cloud, an AI cloud platform that provides managed AI services and AI data center infrastructure. They cater to businesses seeking to accelerate AI solution development with optimized models and high-performance computing. The company utilizes environmentally aligned power sources, including wind, solar, and natural gas, to power its data centers. With features like managed Kubernetes and Slurm, Crusoe simplifies operations and ensures reliability with 24/7 support. Crusoe is expanding its reach, including a strategic European expansion with its first data center in Norway. Crusoe recently raised $1.375 billion at a valuation above $10 billion.

Get more remote jobs like this

Subscribe to the weekly newsletter for similar remote roles and curated hiring updates.

Newsletter

Weekly remote jobs and featured talent.

No spam. Only curated remote roles and product updates. You can unsubscribe anytime.

Similar Opportunities

Expired
Abridge AI, Inc. logo

Abridge AI, Inc.

SF Office
Full-time
Expired Jul 20, 2026
Python
Terraform
GitHub
+4 more

3 months ago

Polymarket Inc. logo

Polymarket Inc.

New York
Full-time
Expires Sep 19, 2026
Express
Senior
Onsite

28 days ago

Polymarket Inc. logo

Polymarket Inc.

New York
Full-time
Expires Oct 7, 2026
Express
Onsite

10 days ago

Remote
Full-time
Expires Oct 10, 2026
JavaScript
TypeScript
Rust
+3 more

7 days ago