About The Company
Our hiring partner is on a mission to make AI compute ubiquitous, seamless, and limitless. They’re building a cloud where AI just works—anywhere, anytime. “AI Power. Everywhere.” Join the team designing the infrastructure for an AI-first world.
Why this role exists
Our hiring partner is looking for a Backend Engineer to build the systems that orchestrate GPU clusters for AI workloads. You’ll develop APIs that manage GPU allocation, memory, compute scheduling, and multi-tenant isolation—challenges unique to AI infrastructure that go far beyond standard backend engineering. On their backend team, you’ll tackle questions like: How can high-cost GPU resources be efficiently shared among users? How do we handle memory constraints for large AI models? How do we maintain quality of service when workloads compete for compute? This is an opportunity to build infrastructure where every API call could allocate thousands of dollars of compute per hour, and where your optimizations directly influence whether AI startups can train their models affordably.
What you’ll do
You’ll thrive here if you
Bonus qualifications
Details