thealleninstitute

Senior Software Engineer, AI Infrastructure

Apply Now

At a Glance

Location
Seattle, Washington, United States
Experience
8+ years
Compensation
ter. Our base salary range is $126,000 - $189,000, and in addition we have gener
Posted
2026-06-08T17:40:41-04:00

Key Requirements

Required Skills

DockerKubernetesLinuxPython

Domain Knowledge

  • Engineering

Benefits & Perks

Time Off

er year, up to 20 vacation days per year and twelve paid holi

Requirements

You are motivated by the idea that world-class infrastructure should be a catalyst for public good, not a proprietary secret.

You are as comfortable designing a resource allocation algorithm in Go as you are debugging a NCCL timeout.

Not only do you build systems, but you also ensure they thrive under the pressure of training world-class AI models.

of professional experience developing business-critical software and operating large-scale compute infrastructure.

Proficiency in Go and/or Python preferred.

Expert-level knowledge of Linux internals, and container runtimes like Docker.

Compensation & Benefits

Team members and their families are covered by medical, dental, vision, and an employee assistance program.

Team members are able to enroll in our health savings account plan, our healthcare reimbursement arrangement plan, and our health care and dependent care flexible spending account plans.

Team members are able to enroll in our company’s 401k plan.

Team members will receive $125 per month to assist with commuting or internet expenses and will also receive $200 per month for fitness and wellbeing expenses.

Team members will also receive up to ten sick days per year, up to seven personal days per year, up to 20 vacation days per year and twelve paid holidays throughout the calendar year.

Team members will be able to receive annual bonuses and can participate in the long-term incentive plan.

Responsibilities

Full-Stack Ownership:

Independently design and deliver critical systems that span the entire stack—from the Beaker job scheduler to the execution runtime.

Build innovative tooling and software-defined infrastructure to accelerate researcher velocity and automate cluster health management.

Conduct root-cause analysis on complex distributed system failures and implement optimizations for distributed workloads.

Technical Input & Ownership:

Provide valuable input into the roadmap for managing large-scale HPC systems, including the deployment of compute, networking, and storage in partnership with leadership.

About the Company

While much of the AI industry has moved behind closed APIs, proprietary datasets, and "black box" infrastructure, Ai2 remains a lighthouse for Open Science. Founded by the late Paul Allen, we are a non-profit research institute dedicated to building AI for the common good.

We don't have a stock price to defend or a walled garden to protect. Instead, we have a mission: to provide the global research community with the transparent, high-performance foundations they need to achieve humanity-enriching breakthroughs.

What Makes Us Different:

Radical Transparency:

We don't just release model weights; we release the data, the training code, and the infrastructure insights. We believe the "how" is just as important as the "what."

Mission over Margin: