Senior / Staff Infrastructure Engineer (Compute)
About FluidStack
Fluidstack is the AI Cloud Platform. We build GPU supercomputers for top AI labs, governments, and enterprises. Our customers include Mistral, Poolside, Black Forest Labs, Meta, and more.
Our team is small, highly motivated, and focused on providing a world class supercomputing experience. We put our customers first in everything we do, working hard to not just win the sale, but to win repeated business and customer referrals.
We hold ourselves and each other to high standards. We expect you to care deeply about the work you do, the products you build, and the experience our customers have in every interaction with us.
You must work hard, take ownership from inception to delivery, and approach every problem with an open mind and a positive attitude. We value effectiveness, competence, and a growth mindset.
About the Role
We are looking for an Senior / Staff Infrastructure Engineer (Compute) to design, deploy, and manage the compute infrastructure powering Fluidstack's GPU clusters. You will be responsible for ensuring the performance, scalability, and reliability of our compute resources, working closely with hardware and software teams to support our AI workloads.
Focus
Design and implement GPU/ASIC infrastructure at the server, rack, and system level.
Troubleshoot complex GPU and compute system related failures.
Develop and maintain hardware/firmware management services.
Automate all aspects of the server lifecycle.
Own end-to-end compute lifecycle, including partnering with vendors on RMAs.
Serve as the main point of contact for hardware escalation and troubleshooting.
Monitor system performance, identifying and resolving bottlenecks.
Automate deployment and management tasks to improve efficiency.
Collaborate with storage and network teams to ensure cohesive infrastructure operations.
About You
5+ years of experience in compute infrastructure engineering.
Strong knowledge of Linux systems administration and performance tuning.
Experience with bare metal provisioning tools (MaaS, Metal3, Tinkerbell, or other).
Familiarity with GPU hardware and workload optimization, especially kernel and driver level requirements.
Proficiency in automation tools (e.g., Ansible, Terraform).
Experience operating Kubernetes and SLURM clusters.
Benefits
Competitive total compensation package (salary + equity).
Retirement or pension plan, in line with local norms.
Health, dental, and vision insurance.
Generous PTO policy, in line with local norms.
Fluidstack is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Fluidstack will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.
Recommended Jobs
Farm Equipment Operator
Farm Equipment Operator (6151) Location: New York State Job Number: 6151 Farm Equipment Operator Position available on 1,200 cow dairy and 8,000 acre crop farm in North Central New York. MUS…
Commis Chef
JOB DETAILS JoJo Restaurant is eeeking experienced Commis Chefs! THE BRAND Jean-Georges Management is built on a powerful culinary foundation and has evolved into a reputable and award-winn…
Service Technician
Joe Cecconi's Chrysler Complex is currently looking to hire an auto mechanic to join our service department. Joe Cecconi's has been in business for over 50 years and has the largest Chrysler, Dodge Je…
OPEN JOB: Emergency Medicine BC/BE - to $400K + $100K Starting Bonus | 272 Hours of PTO
OPEN JOB: Emergency Medicine BC/BE - to $400K + $100K Starting Bonus | 272 Hours of PTO LOCATION: Corning New York or Sayre Pennsylvania (candidate choice) ***Relocation Assistance Available …
Licensed Practical Nurse (LPN)
CNY Family Care, LLP, located in the heart of East Syracuse, is seeking a dedicated and compassionate Licensed Practical Nurse (LPN) to join our dynamic healthcare team. As a community-centered medica…
Infrastructure and Cloud Engineer - NBS
Nelnet Business Services (NBS), a division of Nelnet, Inc., provides payment technology, education services, and learning management solutions to education and faith-based organizations, serving mor…
Project Coordinator
Local Foreigner is a boutique consultancy specializing in high-end curated travel. Whether orchestrating a romantic weekend in Paris or planning an epic Patagonian expedition, we transform travel aspi…
Embrace Adventure in Albany: Heal, Deliver, Explore!
Registered Nurse - Labor & Delivery - Travel - (LD RN) Embark on a rewarding travel nursing opportunity as a Registered Nurse in Labor and Delivery at Albany Medical Center, a regional perinatal cent…
Software Engineer, Agent Platform
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a qui…
Manager - International Tax Services - Generalist
Manager - International Tax Services - Generalist, PwC US Tax LLP, New York, NY. Help multi-national businesses achieve their business goals in a tax-efficient manner and address their cross-border …