Principal Product Manager, AI Infrastructure (Networking)

Crusoe
Sunnyvale +2 more
On-site

Who this role is best for

Geared toward technical product leaders comfortable with defining multi-year networking roadmaps and managing vendor strategies.

Best fit for

  • Candidates with 8+ years of technical product management experience in AI infrastructure or data center networking
    — “8+ years of technical product management experience, or product-minded infrastructure engineering experience
  • Individuals with a strong background in networking architecture and vendor management at scale
    — “Own Crusoe's long-term vendor management strategy for the network, including defining our cost structure and capital allocation
  • Professionals who can translate complex technical tradeoffs into strategic business decisions
    — “Represent networking foundations in leadership reviews, translating architecture and vendor tradeoffs into clear decision docs

Things to consider

  • The role requires a deep understanding of RDMA fabrics and GPU cluster networking for AI workloads
    — “Direct experience with RDMA fabrics such as InfiniBand or RoCE, and with GPU cluster networking for AI training or inference
  • Candidates must be able to work in high-performing teams with a strong sense of urgency
    — “problem-solving, opportunity-finding teammates with a sense of urgency

How to stand out

  • Emphasize your experience in defining fleet-wide network architecture and failure-domain strategies
    — “Direct experience with fleet-wide network architecture, topology design, and failure-domain strategy at scale
  • Highlight your analytical skills in making multi-year infrastructure decisions based on data
    — “Analytical mindset with the ability to use data to inform multi-year infrastructure decisions
Pace · SteadyCollaboration · HighAutonomy · HighDecision Impact · Company

Derived from job-description analysis by Serendipath's career intelligence engine.

What success looks like

  • define long-term vision and roadmap
  • set performance, reliability, scalability, and cost targets
  • own cross-DC and multi-site connectivity strategy
  • drive NIC, DPU, switch, and networking technology strategy
Typical background
Experience in product management, preferably in AI infrastructure or networking

Skills & requirements

Required

AI InfrastructureNetworking FoundationVendor Management StrategyPlatform TargetsConnectivity ArchitectureTechnology StrategyEngineering PartnershipExecutive Communication

Preferred

Cloud-native ArchitectureHyperscaler Portals

Stack & domain

Ai InfrastructureNetworking FoundationFleet-wide ArchitectureFabric DesignVendor Management StrategyNetwork Cost StructureCapital AllocationPerformance TargetsReliability TargetsScalability TargetsCost TargetsNetwork Cost Per GpuNetwork Cost Per ServerCross-dc ConnectivityMulti-site ConnectivityNetwork TopologyFailure-domain DesignNicDpuSwitchNetworking TechnologyHardware GenerationsVendorsStructural LimitationsLong-term InvestmentsEngineering PartnershipInfrastructure EngineeringArchitectureMajor Technical TradeoffsProduct RequirementsBusiness RationaleFoundational InvestmentsLeadership ReviewsArchitecture And Vendor TradeoffsLeadershipCollaborationCommunicationProblem-solvingTechnical ExpertiseTechnical GuidanceTechnical FeedbackTechnical ApproachesTechnical Apis

About the role

Original posting from Crusoe via Ashby

Crusoe is on a mission to accelerate the abundance of energy and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack — from electrons to tokens — to power the world's most ambitious AI workloads. When you join Crusoe, you join a team that is building the future, faster.

We're in the midst of the greatest industrial revolution of our time. The demand for AI compute is boundless, and power is a bottleneck. We're solving that — with an energy-first approach that makes AI infrastructure better for the world and faster for the people innovating with AI.

We're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved — people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.

If you want to do the most meaningful work of your career, help our customers and partners advance their AI strategies, and be part of a high-performing team that believes in each other, come build with us at Crusoe.

About the Role:

Crusoe is building the infrastructure layer that powers the most demanding AI workloads on the planet. Our networking stack is what connects thousands of GPUs into the high-performance clusters that some of the biggest names in AI rely on to train and run their models. As a Principal Product Manager on our AI Infrastructure team, you will own the long-term architecture and technology strategy for Crusoe's networking foundation: the fleet-wide design, the fabric, and the technology choices that determine whether our network can scale for years, not just for the next customer request.

This is a foundations role. You will work as a peer to our most senior infrastructure engineers and architects, and your decisions on architecture, vendor strategy, and technology direction will shape what every Crusoe customer builds on for years to come. You will partner closely with Infrastructure Engineering, Cloud Software Engineering, and Architecture to translate future workload, fleet, and customer requirements into the infrastructure Crusoe needs, and drive the long-term vendor and technology strategy that keeps our network world-class for AI at scale.

What You'll Be Working On:

  • Long-term vision: Define the 2-4 year vision and roadmap for Crusoe's networking foundation, from fleet-wide architecture to fabric design.
  • Vendor management strategy: Own Crusoe's long-term vendor management strategy for the network, including defining our cost structure and capital allocation.
  • Platform targets: Set the performance, reliability, scalability, and cost targets the network must hit, including network cost per GPU and per server.
  • Connectivity architecture: Own cross-DC and multi-site connectivity strategy, network topology, and failure-domain design.
  • Technology strategy: Drive NIC, DPU, switch, and networking technology strategy, including evaluating new hardware generations and vendors.
  • Structural investment: Identify structural limitations in the current architecture and prioritize the long-term investments that resolve them.
  • Engineering partnership: Partner with Infrastructure Engineering and Architecture on major technical tradeoffs, and define the product requirements and business rationale behind foundational investments.
  • Executive communication: Represent networking foundations in leadership reviews, translating architecture and vendor tradeoffs into clear decision docs.

What You'll Bring to the Team:

  • Bachelor's degree in electrical engineering, computer science, or a related field.
  • 8+ years of technical product management experience, or product-minded infrastructure engineering experience, with a track record of owning architecture-level outcomes at Principal or equivalent scope.
  • Subject matter expertise in networking, with the credibility to operate as a peer to senior infrastructure engineers and architects.
  • Direct experience with fleet-wide network architecture, topology design, and failure-domain strategy at scale.
  • Experience managing long-term vendor strategy and vetting in-house software engineering investment for core networking infrastructure.
  • A strong grasp of infrastructure economics at scale, including cost per unit and capacity planning.
  • Analytical mindset with the ability to use data to inform multi-year infrastructure decisions.
  • Exceptional communication skills, with the ability to translate technical tradeoffs into clear narratives for engineering leadership and executives.
  • Direct experience with RDMA fabrics such as InfiniBand or RoCE, and with GPU cluster networking for AI training or inference.

Bonus Points

  • Experience with data center network design, including cross-DC and multi-site connectivity.
  • Familiarity with the current landscape of networking silicon, NICs, and switch vendors serving AI infrastructure.
  • Experience managing or mentoring other product managers.

Benefits:

  • Competitive compensation and equity packages
  • Restricted Stock Units
  • Paid time off, paid holidays & leave of absence programs
  • Comprehensive health, dental & vision insurance
  • Employer contributions to HSA account
  • Paid parental leave
  • Paid life insurance, short-term and long-term disability
  • Professional development & tuition reimbursement
  • Mental health & wellness support
  • Commuter benefits (parking & transit)
  • Cell phone stipend
  • 401(k) Retirement plan with company match up to 4% of salary
  • Volunteer time off
  • Global travel insurance & emergency assistance
  • Daily meals allowance
  • Additional perks & programs specific to location

Compensation Range

Compensation will be paid in the range of up to $285,000 - $335,000 + Bonus. Restricted Stock Units are included in all offers. Compensation to be determined by the applicant's knowledge, education, and abilities, as well as internal equity and alignment with market data.

Crusoe is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.

Source: Crusoe careers (Ashby)

Similar roles