Andromeda (andromeda.ai) Logo

Andromeda (andromeda.ai)

Software Engineer - AI Infrastructure

Reposted One Month Ago
In-Office or Remote
Hiring Remotely in Canada
Senior level
In-Office or Remote
Hiring Remotely in Canada
Senior level
As a Software Engineer in AI Infrastructure, you will design and develop core platform components, build APIs and services, enhance performance, and automate tooling while collaborating across teams and improving system reliability.
The summary above was generated by AI
Software Engineer - AI Infrastructure

Location: North America Remote / San Francisco · Full-Time

About Andromeda

Andromeda is a market and infrastructure platform to buy, sell, and operate compute.
We believe demand for compute will grow exponentially. So fast that a handful of vertically integrated providers won't be able to scale across operations, capital, supply chains, and politics to serve it. The result is a massive wave of fragmentation, with AI factories of every shape and size coming to market to fill this demand. Our job is to enable all of that fragmented compute to flow through one platform, delivering reliable capacity to model builders, research labs, and inference providers when they need it. We believe every spare electron should be made productive for AI and we're building the platform that makes that possible.


We sit at the center of three forces:

  • Companies that need reliable, high-performance compute fast

  • A fragmented global supply of GPUs across hyperscalers, neoclouds, and independent data centers

  • Capital, risk, and operational complexity that most teams are not equipped to manage

When we succeed, trillions of dollars of compute will flow through Andromeda. Builders get capacity when they need it. Providers get a reliable way to monetize, operate, and finance infrastructure at scale. Capital gets an easy way to deploy, hedge, and underwrite.
In five years, Andromeda won't just participate in the AI infrastructure market. We will shape it.

The Role

As an Infrastructure Product Engineer, you will play a pivotal role in building the backbone of Andromeda’s platform. You'll transform complex, real-world infrastructure challenges into scalable product capabilities that benefit our customers.

Positioned at the intersection of infrastructure and product engineering, this role is deeply technical and systems-oriented, yet laser-focused on building solutions with broad leverage.

What You'll Do
  • Design and develop core platform components, including infrastructure orchestration, provisioning, and lifecycle management solutions.

  • Build robust APIs, services, and control planes that abstract over diverse infrastructure types (VMs, Kubernetes, bare metal, schedulers).

  • Translate customer usage patterns into product requirements, delivering impactful features and improvements.

  • Create automation and internal tooling to eliminate manual or ad-hoc operational work.

  • Enhance reliability, performance, and observability at the platform level, emphasizing durable improvements over quick fixes.

  • Collaborate with peer teams to define clear ownership boundaries between platform capabilities and customer-specific solutions.

  • Write clean, maintainable, and well-documented code with a focus on long-term sustainability.

  • Participate in technical design discussions and contribute to the architectural evolution of our platform.

What We're Looking For
  • 5+ years of experience in Infrastructure, Platform, or Backend Engineering roles.

  • Strong systems fundamentals: deep understanding of Linux, networking, storage, and distributed systems.

  • Proven expertise with Kubernetes, VMs, or bare-metal environments.

  • Advanced software engineering skills; capable of building production-grade APIs and services (Python, Go, or similar).

  • Extensive experience with infrastructure as code and automation tools (Terraform, Ansible, Helm, etc.).

  • Demonstrated ability to navigate ambiguity and distill complex problems into clear, maintainable abstractions.

  • Product-focused mindset: care about interfaces, defaults, reliability, and sustainable operations.

  • Excellent written and verbal communication skills; effective collaborator across engineering and product functions.

Nice to Have:

  • Hands-on experience with GPU or AI infrastructure.

  • Experience with control-plane or orchestration systems.

  • Background spanning both infrastructure and application/backend engineering.

  • Experience architecting multi-tenant systems.

  • Strong skills in technical writing and design documentation.

  • Early-stage startup experience.

Why You’ll Love It Here

This is a true builder’s opportunity: you’ll have ownership and autonomy to shape our systems, engage directly with customers and providers, and lay the foundations for scalable, reliable AI infrastructure. Join us at Andromeda and help power the future of AI.

Similar Jobs

3 Hours Ago
Remote or Hybrid
2 Locations
Entry level
Entry level
Big Data • Food • Hardware • Machine Learning • Retail • Automation • Manufacturing
The People Experience Advisor supports service operations by ensuring compliance with policies, improving processes, and enhancing employee experience through effective onboarding and policy interpretation.
3 Hours Ago
Remote or Hybrid
Senior level
Senior level
eCommerce • Fintech • Hardware • Payments • Software • Financial Services
Own the full outbound B2B sales cycle for mid-market merchants, including prospecting, pipeline development, discovery, product demonstrations, negotiations, and closing complex multi-product deals. Build net-new business, tailor Square ecosystem solutions, forecast accurately in Salesforce, collaborate cross-functionally, and consistently achieve revenue targets.
Top Skills: Salesforce
7 Hours Ago
Remote
Trenton, ON, CAN
Senior level
Senior level
Aerospace • Information Technology • Software • Cybersecurity • Design • Defense • Manufacturing
Provides aerospace quality assurance and QMS oversight for C-17 maintenance operations. Performs maintenance and records audits, failure and forensic analysis, root cause analysis, compliance reporting, and quality checks against technical orders and regulations. Trains personnel on QMS and SMS, addresses audit findings, promotes safety and first-time quality, and advises management on compliance risks. The role is fully onsite in Trenton, Ontario, with up to 10% travel.
Top Skills: As9100Automated Data Aircraft Management (Adam)Electronic Record Keeping System (Erks)ExcelMicrosoft OutlookMicrosoft PowerpointMicrosoft WordQuality Management System (Qms)Safety Management System (Sms)

What you need to know about the Calgary Tech Scene

Employees can spend up to one-third of their life at work, so choosing the right company is crucial, not just for the job itself but for the company culture as well. While startups often offer dynamic culture and growth opportunities, large corporations provide benefits like career development and networking, especially appealing to recent graduates. Fortunately, Calgary stands out as a hub for both, recognized as one of Startup Genome's Top 100 Emerging Ecosystems, while also playing host to a number of multinational enterprises. In Calgary, job seekers can find a wide range of opportunities.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account