A logo

Software Engineer - AI Infrastructure

Andromeda ClusterSan Francisco, California

Automate your job search with Sonara.

Submit 10x as many applications with less effort than one manual application.1

Reclaim your time by letting our AI handle the grunt work of job searching.

We continuously scan millions of openings to find your top matches.

pay-wall

Overview

Schedule
Full-time
Career level
Senior-level
Remote
Remote
Benefits
Career Development

Job Description

Software Engineer - AI Infrastructure

Location: North America Remote / San Francisco · Full-Time

About Andromeda

Andromeda Cluster, founded by Nat Friedman and Daniel Gross, is on a mission to democratize access to cutting-edge AI infrastructure previously reserved for hyperscalers. What began with a single managed cluster has quickly evolved into a global platform, connecting leading AI labs, data centers, and cloud providers.

Our orchestration layer seamlessly routes training and inference jobs across the world, unlocking flexibility and efficiency in one of the fastest-growing sectors on earth. Our long-term vision is to establish a global marketplace for AI compute—powering AGI with the same fluidity as world financial markets.

We are scaling rapidly and seeking exceptional talent in AI infrastructure, research, and engineering.

The Role

As an Infrastructure Product Engineer, you will play a pivotal role in building the backbone of Andromeda’s platform. You'll transform complex, real-world infrastructure challenges into scalable product capabilities that benefit our customers.

Positioned at the intersection of infrastructure and product engineering, this role is deeply technical and systems-oriented, yet laser-focused on building solutions with broad leverage.

What You'll Do

  • Design and develop core platform components, including infrastructure orchestration, provisioning, and lifecycle management solutions.

  • Build robust APIs, services, and control planes that abstract over diverse infrastructure types (VMs, Kubernetes, bare metal, schedulers).

  • Translate customer usage patterns into product requirements, delivering impactful features and improvements.

  • Create automation and internal tooling to eliminate manual or ad-hoc operational work.

  • Enhance reliability, performance, and observability at the platform level, emphasizing durable improvements over quick fixes.

  • Collaborate with peer teams to define clear ownership boundaries between platform capabilities and customer-specific solutions.

  • Write clean, maintainable, and well-documented code with a focus on long-term sustainability.

  • Participate in technical design discussions and contribute to the architectural evolution of our platform.

What We're Looking For

  • 5+ years of experience in Infrastructure, Platform, or Backend Engineering roles.

  • Strong systems fundamentals: deep understanding of Linux, networking, storage, and distributed systems.

  • Proven expertise with Kubernetes, VMs, or bare-metal environments.

  • Advanced software engineering skills; capable of building production-grade APIs and services (Python, Go, or similar).

  • Extensive experience with infrastructure as code and automation tools (Terraform, Ansible, Helm, etc.).

  • Demonstrated ability to navigate ambiguity and distill complex problems into clear, maintainable abstractions.

  • Product-focused mindset: care about interfaces, defaults, reliability, and sustainable operations.

  • Excellent written and verbal communication skills; effective collaborator across engineering and product functions.

Nice to Have:

  • Hands-on experience with GPU or AI infrastructure.

  • Experience with control-plane or orchestration systems.

  • Background spanning both infrastructure and application/backend engineering.

  • Experience architecting multi-tenant systems.

  • Strong skills in technical writing and design documentation.

  • Early-stage startup experience.

Why You’ll Love It Here

This is a true builder’s opportunity: you’ll have ownership and autonomy to shape our systems, engage directly with customers and providers, and lay the foundations for scalable, reliable AI infrastructure. Join us at Andromeda and help power the future of AI.

Automate your job search with Sonara.

Submit 10x as many applications with less effort than one manual application.

pay-wall

FAQs About Software Engineer - AI Infrastructure Jobs at Andromeda Cluster

What is the work location for this position at Andromeda Cluster?
This job at Andromeda Cluster is located in San Francisco, California, according to the details provided by the employer. Some roles may also include multiple work locations depending on the requirement.
What pay range can candidates expect for this role at Andromeda Cluster?
Employer has not shared pay details for this role.
What employment applies to this position at Andromeda Cluster?
Andromeda Cluster lists this role as a Full-time position.
What experience level is required for this role at Andromeda Cluster?
Andromeda Cluster is looking for a candidate with "Senior-level" experience level.
Does Andromeda Cluster allow remote work for this role?
Yes, this position at Andromeda Cluster supports remote work, giving candidates the flexibility to work outside the primary office location.
What benefits are offered by Andromeda Cluster for this role?
Andromeda Cluster offers Career Development for this position. Actual benefits may vary depending on the employer's policies and employment terms.
What is the process to apply for this position at Andromeda Cluster?
You can apply for this role at Andromeda Cluster either through Sonara's automated application system, which helps you submit applications 10X faster with minimal effort, or by applying manually using the direct link on the job page.