Computer Software Jobs 2026 (Now Hiring) – Smart Auto Apply

We've scanned millions of jobs. Simply select your favorites, and we can fill out the applications for you.

Fluidstack logo

Software Engineer, Inference Platform

Fluidstack
San Francsisco, California

$165,000 - $500,000 / year

About Fluidstack At Fluidstack, we build the compute, data centers, and power that will fuel artificial superintelligence. We work with Anthropic, Google, Meta, AMI Labs, and Black...

Posted 30+ days ago

Blackstone logo

Software Engineering Manager, SVP – AI Platform Development

Blackstone
Miami, Florida

$175,000 - $250,000 / year

Blackstone is the world’s largest alternative asset manager. We seek to create positive economic impact and long-term value for our investors, the companies we invest in, and the c...

Posted 30+ days ago

Skydio logo

Senior Software Engineer, Data Platform

Skydio
San Mateo, California

$180,000 - $240,000 / year

Skydio is the leading US drone company and the world leader in autonomous flight, the key technology for the future of drones and aerial mobility. The Skydio team combines deep exp...

Posted 30+ days ago

Glimpse logo

Senior Software Engineer

Glimpse
New York, New York
ABOUT GLIMPSE Glimpse is the leading AI platform for CPG brands — automating critical back-office workflows like deductions management, revenue recovery, and cash application. Sinc...

Posted 30+ days ago

T logo

Software Engineer

Traba
New York City, New York

$140,000 - $200,000 / year

Traba is building the autonomous future of industrial labor. We are the AI operating layer for the industrial supply chain. We started in workforce—temp staffing, the biggest opera...

Posted 30+ days ago

V logo

Senior Software Engineer, ML Infrastructure

Voxel
San Francisco, California
Who We Are Voxel is building the future of Computer Vision and Machine Learning for operations, risk, and safety. We use computer vision and AI to enable existing security cameras...

Posted 30+ days ago

R logo

Sr. Fullstack Software Engineer

Reliable Robotics Corporation
Mountain View, California
We're building safety-enhancing technology for aviation that will save lives. Automated aviation systems will enable a future where air transportation is safer, more convenient and...

Posted 30+ days ago

iHeartMedia logo

Software Engineer

iHeartMedia
Nashville, Tennessee
iHeartMedia Current employees and contingent workers click here to apply and search by the Job Posting Title. The audio revolution is here – and iHeart is leading it! iHeartMedia,...

Posted 30+ days ago

Generalist logo

Software Engineer: Infrastructure

Generalist
San Francisco, California
About the Role This role broadly owns infrastructure across the stack. If it’s running in the cloud, you probably care about it. The life blood of robotics is data, and we need hig...

Posted 30+ days ago

Plaid logo

Senior Software Engineer - ML Infrastructure

Plaid
San Francisco, California
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools...

Posted 30+ days ago

Collective logo

Senior Software Engineer

Collective
San Francisco, California
About Collective: Collective is on a mission to redefine the way businesses-of-one work. Our technology and team of trusted advisors help members achieve financial independence by...

Posted 30+ days ago

Generalist logo

Software Engineer: Robotics Controls

Generalist
San Francisco, California
About the Role You will own the software that controls our robots. You will collaborate with ML to turn output intentions into smooth, efficient, safe, and responsive actions on a...

Posted 30+ days ago

S logo

Sr. Software Engineer - Trading Infrastructure

Sei Labs
New York, New York
About Sei Labs Sei Labs builds open sourced technology for the high-performance Sei Blockchain, the first parallelized EVM Layer 1 blockchain designed to scale with the industry. T...

Posted 30+ days ago

External logo

Lead Software Engineer

External
Frisco, Texas

$140,000 - $165,000 / year

Toshiba Global Commerce Solutions is seeking a hands-on Lead Software Engineer to drive end-to-end solution delivery for major retail platforms. In this role, you will own hands-on...

Posted 30+ days ago

S logo

Senior Software Engineer (Product)

Seven AI
Boston, Massachusetts
We are seeking a product-minded Software Engineer to join our core team. In this role, you’ll take ownership across the entire product stack designing and building backend systems,...

Posted 30+ days ago

B logo

Perception Sensor Software Engineer

Bedrock Robotics
San Francisco, California
Join the team bringing advanced autonomy to the built world At Bedrock, we’re moving AI out of the lab and into the real world. Our team is composed of industry veterans who helped...

Posted 30+ days ago

Nimble Robotics logo

Senior Robotics Software Engineer

Nimble Robotics
San Francisco, California

$190,000 - $250,000 / year

About Nimble Nimble is an AI robotics company building the autonomous supply chain to power fast, efficient and economical commerce. We’re training robot AGI to power a proprietary...

Posted 30+ days ago

B logo

Software Engineer, Product & Growth

Blockit AI
San Francisco, California
About Blockit Time is the the most valuable resource we have, yet coordinating it remains stuck in the dark ages. At Blockit, we're building the AI that finally fixes this: an auto...

Posted 30+ days ago

R logo

Flight Software Manager

Reliable Robotics Corporation
Mountain View, California
We're building safety-enhancing technology for aviation that will save lives. Automated aviation systems will enable a future where air transportation is safer, more convenient and...

Posted 30+ days ago

Viant Technology logo

Principal Software Engineer

Viant Technology
Los Angeles, California

$190,000 - $260,000 / year

WHAT YOU’LL DO Viant’s customers use the Demand Side Platform (DSP) to set up, run and monitor ad campaigns. The platform team owns a complex set of backend services and the fronte...

Posted 30+ days ago

Fluidstack logo

Software Engineer, Inference Platform

FluidstackSan Francsisco, California

$165,000 - $500,000 / year

Automate your job search with Sonara.

Submit 10x as many applications with less effort than one manual application.1

Reclaim your time by letting our AI handle the grunt work of job searching.

We continuously scan millions of openings to find your top matches.

pay-wall

Overview

Schedule
Full-time
Career level
Senior-level
Compensation
$165,000-$500,000/year
Benefits
Health Insurance
Dental Insurance
Vision Insurance

Job Description

About Fluidstack

At Fluidstack, we build the compute, data centers, and power that will fuel artificial superintelligence. We work with Anthropic, Google, Meta, AMI Labs, and Black Forest Labs to deploy gigawatts of compute at industry defining speeds. We are investing tens of billions of dollars in US infrastructure. In 2026, we will deploy 1GW. In 2027, 10GW.

Our team is small, fast, and obsessed with quality. We own outcomes end-to-end, challenge assumptions, and treat our customers' problems as our own. No task is beneath anyone here.

There are a few thousand people who will shape the trajectory of superinteligence. Come and be one of them.

About the Role

Inference is now the defining cost and latency bottleneck for frontier AI. Fluidstack’s Inference Platform team owns the serving layer that sits between our global accelerator supply and the production workloads our customers run on it: LLM serving frameworks, KV cache infrastructure, disaggregated prefill/decode pipelines, and Kubernetes-based orchestration across multi-datacenter footprints.

This is a hands-on IC role at the intersection of distributed systems, model optimization, and serving infrastructure. You’ll own end-to-end inference deployments for frontier AI labs and our inference product, drive measurable improvements in throughput, cost-per-token, and time-to-first-token, and contribute to the platform architecture choices that determine how Fluidstack deploys across tens of thousands of accelerators.

You will:

  • Own inference deployments end-to-end: from initial configuration and performance tuning to production SLA maintenance and incident response.

  • Drive measurable improvements in throughput, TTFT, and cost-per-token across diverse model families (dense transformers, mixture-of-experts, multi-modal) and customer workload patterns.

  • Build and operate KV cache and scheduling infrastructure to maximize utilization across concurrent requests.

  • Implement and validate disaggregated prefill/decode pipelines and the Kubernetes orchestration that supports them at scale.

  • Profile and resolve bottlenecks at the compute, memory, and communication layers; instrument deployments for end-to-end observability.

  • Partner with customers to translate their model architectures, access patterns, and latency requirements into deployment configurations and upstream platform improvements.

  • Contribute to inference platform architecture and roadmap, with a focus on reducing deployment complexity, improving hardware utilization, and expanding support for new model classes and accelerators.

  • Participate in an on-call rotation (up to one week per month) to maintain the reliability and SLA commitments of production deployments.

Basic Qualifications

  • 5+ years of professional software engineering experience with a track record of shipping production-quality systems.

  • Strong programming skills in Python and/or Go.

  • Hands-on production experience with at least one LLM serving framework (vLLM, SGLang, TensorRT-LLM, TGI, or equivalent).

  • Working knowledge of PyTorch or JAX and an understanding of how model architecture choices affect inference characteristics.

  • Experience deploying and operating GPU workloads on Kubernetes at production scale, including autoscaling and resource scheduling.

  • Solid understanding of GPU memory hierarchies, compute parallelism, and the tradeoffs across tensor, pipeline, and expert parallelism strategies.

  • Ability to create structure from ambiguity and communicate technical tradeoffs clearly to both engineering peers and customers.

  • Great written and verbal communication skills in English.

Preferred Qualifications

  • Production experience with disaggregated prefill/decode architectures (NVIDIA Dynamo, LLM-d, or equivalent), including scheduling policies and network fabric configuration.

  • Deep familiarity with KV cache strategies: RadixAttention, slab-based memory allocators, cross-request prefix sharing, and cache-aware scheduling.

  • Experience with multi-node GPU inference across InfiniBand or RoCE fabrics, including NCCL collective communication tuning.

  • Custom kernel or operator development experience (e.g., CUDA, Triton, torch.compile, Pallas, or equivalent)

  • Contributions to open-source inference engines (vLLM, SGLang, TGI, TensorRT-LLM, or similar).

  • Hands-on experience with quantization tooling: GPTQ, AWQ, FP8 via llm-compressor, or AutoGPTQ.

  • Knowledge of speculative decoding implementations (Medusa, EAGLE-3, draft-model approaches) and their performance/quality tradeoffs.

  • Experience optimizing and adapting model implementations for non-NVIDIA accelerators and their ecosystems: AMD, TPU, Trainium/Inferentia, Cerebras, Groq, and other custom ASICs.

Salary & Benefits

  • Competitive total compensation package (salary + equity).

  • Retirement or pension plan, in line with local norms.

  • Health, dental, and vision insurance.

  • Generous PTO policy, in line with local norms.

The base salary range for this position is $165,000 – $500,000 per year, depending on experience, skills, qualifications, and location. This range represents our good faith estimate of the compensation for this role at the time of posting. Total compensation may also include equity in the form of stock options.

We are committed to pay equity and transparency.

Fluidstack is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Fluidstack will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

You will receive a confirmation email once your application has successfully been accepted. If there is an error with your submission and you did not receive a confirmation email, please email careers@fluidstack.io with your resume/CV, the role you've applied for, and the date you submitted your application-- someone from our recruiting team will be in touch.

Automate your job search with Sonara.

Submit 10x as many applications with less effort than one manual application.

pay-wall