RefolkCandidates
Open nowEngineeringCloud Networking

Staff Network Production Engineer, Deployment

Crusoe · San Francisco, California

Location
San Francisco, California, USA
Employment
Full time
Level
Staff
Posted
6 days ago

About this role

Crusoe is on a mission to accelerate the abundance of energy and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack - from electrons to tokens - to power the world's most ambitious AI workloads. When you join Crusoe, you join a team that is building the future, faster.

We're in the midst of the greatest industrial revolution of our time. The demand for AI compute is boundless, and power is a bottleneck. We're solving that - with an energy-first approach that makes AI infrastructure better for the world and faster for the people innovating with AI.

We're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.

If you want to do the most meaningful work of your career, help our customers and partners advance their AI strategies, and be part of a high-performing team that believes in each other, come build with us at Crusoe.

About This Role:

Crusoe Cloud is seeking a high-energy, detail-oriented Senior Network Production Engineer to support the physical and logical implementation of our global network. As we rapidly expand our footprint of high-performance compute (HPC) and GPU-based AI infrastructure, this role plays a key part in bringing new data centers and edge sites online, directly enabling the scale and reliability our AI cloud customers depend on.

This is a technical, hands-on, full-time role that sits at the intersection of Network Engineering, Data Center Operations, and Project Management. The ideal candidate is comfortable both on-site and remote, follows and helps refine deployment standards, and takes ownership of ensuring every switch, router, and fiber optic link is deployed to spec, validated, and seamlessly handed off to our Operations team.

What You'll Be Working On:

  • Network Deployment Execution: Execute on-site and remote network deployments for new data center and edge site builds, from rack-and-stack through final cutover.

  • Implementation Planning: Implement architecture designs into concrete deployment steps: cable maps, port assignments, device configs, and site-specific runbooks.

  • Automation Development: Write and maintain Python/Ansible automation and ZTP workflows to stage, configure, and validate batches of switches and routers without manual touch.

  • Acceptance Testing: Support burn-in and site acceptance testing (SAT) on new clusters, chase down failures, and help drive clean handoffs to Operations.

  • Fabric Configuration & Troubleshooting: Configure and troubleshoot Arista, Juniper, and NVIDIA/Mellanox gear in leaf-spine fabrics, including BGP, EVPN-VXLAN, and LLDP issues at the link and fabric level.

  • Vendor & Physical Layer Coordination: Work directly with structured cabling vendors, remote hands, and data center providers on-site to resolve physical layer issues - bad fiber runs, power/cooling deviations, mislabeled patch panels - before they block turn-up.

  • Physical/Link-Layer Diagnostics: Diagnose physical and link-layer problems using OTDRs, light meters, and packet captures when something doesn't come up clean.

  • Inventory & Capacity Tracking: Track hardware inventory and support turn-up of new backbone and edge interconnect capacity against deployment schedules.

  • Cross-Team Collaboration: Collaborate with other engineers on tricky sites or configs, and flag recurring issues back to the team so they get addressed at the process or automation level.

  • On-Call Support: Participate in an on-call rotation, responding to and troubleshooting network incidents affecting production infrastructure.

What You'll Bring to the Team:

  • 5+ years of experience in network engineering with a focus on large-scale data center deployments and infrastructure projects.

  • Strong knowledge of physical layer standards: solid experience with structured cabling (SMF/MMF, MPO/MTP), optical transceivers (400G/800G), and data center power/cooling requirements.

  • Solid routing and switching knowledge: hands-on experience configuring Arista (EOS), Juniper (Junos), and NVIDIA/Mellanox platforms in a leaf-spine architecture.

  • Protocol familiarity: working understanding of BGP, EVPN-VXLAN, and LLDP as they relate to large-scale fabric provisioning.

  • Automation-minded: proficiency in Python and Ansible for automating repetitive deployment tasks and validating configuration state.

  • Logistical skills: ability to manage multiple projects simultaneously across different time zones and physical locations.

  • Troubleshooting skills: ability to diagnose physical layer and link-layer issues using OTDRs, light meters, and packet captures.

  • Education: Bachelor's degree in a technical field or equivalent practical experience in hyperscale or ISP environments.

Bonus Points:

  • Experience working in hyperscale, ISP, or large multi-tenant data center environments.

  • Familiarity with GPU cluster networking (e.g., RDMA/RoCE, InfiniBand, or NVIDIA NCCL-aware fabric design).

  • Exposure to network monitoring and observability tooling (e.g., Prometheus, Grafana, telemetry-based fabric health checks).

  • Vendor certifications such as CCNP, JNCIP, or Arista ACE.

  • Prior experience mentoring junior engineers or contributing to deployment standards/documentation.

Benefits:

  • Competitive compensation and equity packages

  • Restricted Stock Units

  • Paid time off, paid holidays & leave of absence programs

  • Comprehensive health, dental & vision insurance

  • Employer contributions to HSA account

  • Paid parental leave

  • Paid life insurance, short-term and long-term disability

  • Professional development & tuition reimbursement

  • Mental health & wellness support

  • Commuter benefits (parking & transit)

  • Cell phone stipend

  • 401(k) Retirement plan with company match up to 4% of salary

  • Volunteer time off

  • Global travel insurance & emergency assistance

  • Daily meals allowance

  • Additional perks & programs specific to location

Compensation Range

Compensation will be paid in the range of up to $165,000 - $200,000 + Bonus. Restricted Stock Units are included in all offers. Compensation to be determined by the applicant's knowledge, education, and abilities, as well as internal equity and alignment with market data.

Crusoe is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.

As published by Crusoe. Applications are handled on their site.

Skills this posting mentions

AutomationJuniperCloud Services

About Crusoe

Crusoe is building the World’s Favorite AI-first Cloud infrastructure company. We’re pioneering vertically integrated, purpose-built AI infrastructure solutions trusted by Fortune 500 companies to power their most advanced AI applications. Crusoe is redefining AI cloud infrastructure, with a mission to align the future of computing with the future of the climate. Our AI platform is recognized as the “gold standard” for reliability and performance. Our data centers are optimized for AI workloads and are powered by clean, renewable energy.

All 363 openings at Crusoe

One click, then it is written

Apply to Crusoe with a resume written for this role.

Queue Staff Network Production Engineer, Deployment and I read the posting, rewrite your resume against it, draft the cover letter, and score the fit. Then you press send, or press one button and I fill in Crusoe’s form for you.

  1. 01Drop your resume

    A PDF or a LinkedIn URL. About a minute, once.

  2. 02I rank the openings

    Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.

  3. 03Each one is written up

    Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.

  • 25 sent a week, free
  • No card
  • Nothing sent until you say so

More roles at Crusoe

See all

Similar roles elsewhere

See more

Put this to work

Paste your career in once. Every application after that is written for you.

Drop a resume or a LinkedIn URL. I rank the live openings against it, rewrite the resume and write a cover letter for the best of them, and fill in the employer's form when you press the button. You read, you decide what goes out.

  1. 01Drop your resume

    A PDF or a LinkedIn URL. About a minute, once.

  2. 02I rank the openings

    Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.

  3. 03Each one is written up

    Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.

  • New matches ranked and written before you are up.
  • Every bullet stays inside what your history supports. Nothing invented.
  • Queued, submitted, interviewing, offer: one screen, not a spreadsheet.

500 free credits on sign-up. No card. Nothing is sent until you say so.

Listed from the job board Crusoe publishes. Refolk is not the employer and does not handle their hiring. Applications go to Crusoe directly.