Post a job

Senior Compute Infrastructure Engineer

StackYak · United States · posted today ago

Apply for this role

engineeringoperationsremote jobs in united states
Listing supplied by Himalayas. 101 Careers did not originate this post and applications are handled by the employer.

About this role

  1. About StackYak

StackYak is building the infrastructure layer for AI.

We are an early-stage, funded company building software that brings compute, GPU infrastructure, networking, and inference together into one product. The opportunity is large, the market is moving quickly, and we are building for real production workloads from the start.

This is not internal IT. This is not a slow-moving infrastructure team maintaining someone else's platform. The infrastructure is the product.

We are a small, senior team with very little bureaucracy. This is a founding role in its discipline. You will be the first person here whose primary responsibility is the systems layer beneath the product, and the shape it takes will largely be yours to decide.

Treat this document as a starting point rather than a boundary. The people who do well here take ground early and are not asked to give it back.

We move quickly. We do not have months for someone to learn the fundamentals of their discipline. You should already be very good at what you do, be able to ramp into adjacent areas quickly, and be comfortable operating without perfect requirements or neatly defined boundaries.

The Role

We need someone who knows what stands between GPU capacity and production infrastructure, and can build it.

That capacity does not arrive in one form. Some of it comes as hardware, with everything that implies: firmware, drivers, hosts, and physical work carried out by people you will never meet in buildings you will rarely visit. Some of it comes from neoclouds and other providers, where you control the software and very little else, on terms set by someone whose interests are not yours.

The interesting problem is neither of those on its own. It is the gap between them — how much of it can honestly be abstracted away, how much has to stay visible, and how much the rest of the company should ever have to think about. Those are open questions here, and the person in this role is the one who gets to answer them.

You should be comfortable at both ends. On one side: Linux, kernel, firmware, drivers, virtualization, host hardening, and debugging a machine you cannot walk up to. On the other: provider APIs, capacity that is not permanent, and automation at a scale where nothing gets configured by hand.

Workloads will run on bare metal, in VMs, and in containers. The right answer will not be the same one twice, and choosing is part of the job.

This is not an architecture-only role. You will design systems and then build, debug, and operate them.

What You Will Own

Not tasks. Outcomes, and the authority that comes with them.

What Success Looks Like

You can be handed GPU capacity — some of it ours, some of it rented, none of it uniform — along with a set of business and security requirements, and turn it into production infrastructure the rest of the company can trust and operate.

You will help us answer questions such as:

What We Need

You Will Be Especially Strong If

This Is Probably Not For You If

How We Work

Compensation

Competitive compensation plus meaningful equity. Exact structure will depend on location, engagement model, and experience.

A Note For Agencies

We are not using external recruiters or agencies for this role, and we will not be persuaded otherwise by an email. We do not want your spam. We will not read the CVs you send, we will not reply to your follow-up, and no candidate you put in front of us creates a fee obligation of any kind. Do not contact us.

Originally posted on Himalayas

Similar remote engineering jobs

Browse all remote engineering jobs →

Senior Compute Infrastructure Engineer at StackYak — Remote | 101 Careers