Fluidstack Logo

Fluidstack

Network Engineer, Design & Engineering

Reposted 11 Days Ago
Hybrid
New York, NY, USA
202K-261K Annually
Senior level
Hybrid
New York, NY, USA
202K-261K Annually
Senior level
Design end-to-end datacenter network architectures for AI training and inference: topology, IP/routing, RDMA/lossless fabrics, physical integration (rack, power, cabling), and comprehensive HLD/LLD documentation. Collaborate with hardware, operations, cabling, software, and validation teams to produce deployable, scalable, and performance-validated designs.
The summary above was generated by AI
About Fluidstack

We exist to make humanity more free. For most of human history, you farmed or you starved. Technology gave people more time for the things they wanted to do, instead of things they had to do. Powerful AI will be the biggest lever for human choice we've ever built - but only if models are aligned with what humanity actually wants. There are groups building AI who don't share these goals. Whoever deploys frontier compute infrastructure fastest will decide whether AI expands human freedom or shrinks it.

We're singularly focused on delivering 10 to 100s of GWs of compute faster than anyone else, rethinking every layer of the stack. We acquire power, design and build data centers, and operate them - with teams spanning hardware and software. Speed and scale are our key differentiators. Come be a part of building civilization-scale infrastructure for AI.


We hire people who care deeply about this problem space. If that is you, please apply!

How We Operate
  • Be a barrel. Full autonomy. Own things end to end, take on scope without being asked, no permission required to operate outside your core role.

  • Insane urgency. We drive everything forward as fast as possible.

  • Reason from first principles. Challenge every assumption. Zero analogy thinking, no egos, the best idea wins.

  • Love of the game. The frontier of AI is the most interesting problem of our time. We put in long hours at high intensity to push the frontier forward.

  • Build something that actually matters. If you're going to spend your time, spend it on something that matters to the world.

The Infrastructure Team

Examples of key problems the team is working on

  • Design the fabrics the frontier trains on. Lossless, non-blocking backend networks for clusters of 100k+ accelerators, re-derived for every new generation of silicon, often before the chip is public.

  • Multiple fabrics, one system. Frontend, backend, backbone, management, enterprise, and BMS, designed as a single coherent architecture.

  • Generate the design, don't draw it. Topologies, addressing, BGP/ASN schemas, and golden configs produced from a source-of-truth model, a full site network design in days instead of quarters, correct by construction.

     
Role Scope
  • Own the network design lifecycle from customer requirements (GPU shape, workload, scale, tenancy) through deployable, validated architectures for AI training and inference.

  • Produce topology designs, IP/addressing schemes, routing policy, and fabric configuration specs across front-end, back-end (GPU-to-GPU training fabric), and storage networks.

  • Adapt architectures to different GPU platforms (NVIDIA, AMD, custom accelerators), form factors, and workload profiles, each with its own rack layout, power envelope, and cabling approach.

  • Translate logical designs into physical reality: rack elevations, power constraints, structured cabling and fiber budgets, pathway routing, and airflow that affect equipment placement.

  • Design lossless Ethernet fabrics for RDMA (RoCEv2): PFC, ECN tuning, traffic classes, and congestion management, reasoning about ECMP and collective-communication patterns in distributed training.

  • Produce HLDs, LLDs, cutsheets, BOMs, cabling matrices, and design decision records, and lead design reviews and reusable reference architectures.

     
What We're Looking For

The below is a starting point. We always make space for exceptional people, so if you don't fit this role exactly, tell us where you would.

  • You've designed data center network fabrics from requirements through deployment, not just configured them, and can explain the tradeoffs behind every decision.

  • You have deep L1 to L3 expertise: CLOS/fat-tree topologies, BGP, EVPN/VXLAN, and the fundamentals underneath them.

  • You design lossless RDMA (RoCEv2) fabrics and understand congestion management at training scale.

  • You reason from first principles through novel design challenges rather than pattern-matching to one reference architecture.

  • You partner across Hardware, DC Operations, ICT, Software, and Validation so designs are buildable and operationally sound.

  • Bonus: Source-of-truth-driven design generation. Multi-vendor GPU platform integration. Large-cluster (100k+ accelerator) experience.

     
Salary & Benefits
  • Competitive total compensation package (salary + equity).

  • Retirement or pension plan, in line with local norms.

  • Health, dental, and vision insurance.

  • Generous PTO policy, in line with local norms.

 

We are committed to pay equity and transparency.

Fluidstack is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Fluidstack will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

You will receive a confirmation email once your application has successfully been accepted. If there is an error with your submission and you did not receive a confirmation email, please email [email protected] with your resume/CV, the role you've applied for, and the date you submitted your application-- someone from our recruiting team will be in touch.

Similar Jobs

5 Minutes Ago
Remote or Hybrid
17 Locations
150K-240K Annually
Expert/Leader
150K-240K Annually
Expert/Leader
Information Technology • Productivity • Software • Infrastructure as a Service (IaaS)
Develop and architect a cross-platform C++ endpoint patching agent for Windows, macOS, and Linux. Responsibilities include code reviews, mentoring engineers, improving development processes, collaborating across patching teams, and translating product needs into resilient technical designs. The role involves system-level APIs, patch installation and rollback logic, embedded databases, and client-server integration.
Top Skills: AWSC++C++17CmakeGoGrpcJavaKotlinLinuxmacOSPosixPostgresProtocol BuffersQtSqliteVcpkgWin32 ApiWindows
11 Minutes Ago
Easy Apply
Hybrid
Brooklyn, NY, USA
Easy Apply
240K-280K Annually
Senior level
240K-280K Annually
Senior level
Artificial Intelligence • Machine Learning • Software
Lead a distributed cloud engineering team responsible for regulated and sovereign AWS and Azure environments. Own platform architecture, infrastructure as code, Kubernetes, release tooling, accreditation readiness, security controls, reliability, operational practices, cost management, and customer-facing technical engagements. Standardize environment deployments, maintain compliance and SLAs, and provide hands-on technical leadership. Manage hiring, performance, career development, planning, incident response, and cross-functional coordination across engineering, security, compliance, implementation, and product teams.
Top Skills: AWSAzureBlue/Green DeploymentsFedrampInfrastructure As Code (Iac)KubernetesOpentofuTerraform
13 Minutes Ago
In-Office
New York City, NY, USA
209K-246K Annually
Senior level
209K-246K Annually
Senior level
Artificial Intelligence • Legal Tech • Software
Own pipeline and revenue across strategic U.S. accounts, develop and scale go-to-market plays, advance complex enterprise deals, engage senior legal stakeholders, translate product capabilities into commercial value, analyze conversion and win-rate data, and codify successful enterprise sales strategies.
Top Skills: Artificial IntelligenceB2B Saas

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account