CDIT Logo

CDIT

Principal Dev-Ops Architect

Posted 8 Days Ago
Be an Early Applicant
In-Office or Remote
Hiring Remotely in Toronto
Expert/Leader
In-Office or Remote
Hiring Remotely in Toronto
Expert/Leader
Principal-level individual contributor architect responsible for designing and operating a global AWS SaaS platform. Owns Terraform infrastructure, CI/CD, Kubernetes, observability, SLOs, reliability, release management, security controls, and AI/ML platform operations. Establishes engineering standards and reference implementations, governs LLMOps, and ensures HIPAA and SOC 2 Type 2 readiness through secrets management, supply-chain security, audit controls, and FinOps.
The summary above was generated by AI

This is a remote position.

Senior technical authority for a cloud platform running global, multi-tenant SaaS services. This is a hands-on individual-contributor architect role — not people management. The architect designs the platform, sets standards and reference implementations other teams build on, and still writes Terraform, builds CI/CD pipelines, and stands up the AI/ML platform personally. Defines how reliability is measured against SLOs, how releases ship, and how the AI/ML platform is built and governed while keeping the environment HIPAA-compliant and SOC 2 Type 2 audit-ready. Influences products, software, and QA through architecture and example.

Core Responsibilities

     Own platform architecture and technical roadmap for infrastructure, deployment, observability, and the AI/ML platform.

     Set engineering standards, patterns, and golden paths for Infrastructure as Code (IaC), CI/CD, and AI tooling; drive adoption through reference implementations and architecture reviews.

     Manage all cloud infrastructure as code in Terraform — reusable modules, remote state, peer-reviewed PRs, drift detection, and automated plan/apply in CI/CD.

     Enforce policy-as-code (OPA, Sentinel, or equivalent) so infrastructure changes meet security and cost guardrails before merge.

     Design and deploy AWS infrastructure across dev, UAT, staging, and production for performance, availability, recoverability, and security (CIS Critical Security Controls).

     Build and operate CI/CD pipelines for large-scale applications on AWS; own release management, rollback, blue/green, canary, and release gates.

     Package and run containerized workloads on Docker and Kubernetes (EKS).

     Lead the SLI/SLO/SLA program and modern observability using OpenTelemetry; drive down MTTD and MTTR; lead blameless post-incident reviews and participate in on-call.

     Provision and operate the AI/ML platform — Anthropic Claude via AWS Bedrock and internal MCP services — all managed as IaC.

     Build LLMOps practices: prompt versioning, evaluation pipelines, token cost attribution, guardrails, and audit logging of agent actions; enforce the PHI data boundary to BAA-covered providers only.

     Operate and evidence the platform controls required for SOC 2 Type 2 and HIPAA; own secrets management, supply-chain security (SBOM, image and dependency scanning), and FinOps.

Required Qualifications

     Bachelor's degree in Software Engineering or equivalent combination of technical education and work experience.

     10+ years in SRE / DevOps / Platform Engineering delivering CI/CD, REST API deployment, containerization, IaaS/PaaS, data pipelines, and application observability — including time at a senior IC or architect level (Staff, Principal, or Architect).

     Proven technical authority across teams: sets architecture and standards and influences delivery through expertise and example rather than direct management.

     Demonstrated experience driving adoption of a new practice or platform (IaC, CI/CD overhaul, or an AI/ML platform) across multiple teams.

     Hands-on Terraform, including reusable modules other teams consume via self-service, remote state, and change management in a CI/CD pipeline.

     Building and operating CI/CD pipelines for large-scale applications on AWS (GitHub Actions, Jenkins, GitLab, or AWS-native).

     Running containerized workloads on Docker and Kubernetes.

     Monitoring and troubleshooting using cloud-native tooling and OpenTelemetry.

     Linux system administration, Unix scripting, and automation.

     Experience working in a HIPAA / HITECH / HITRUST / PHI / PII or PCI DSS environment.



Similar Jobs

4 Hours Ago
Remote or Hybrid
CA
Expert/Leader
Expert/Leader
eCommerce • Fintech • Hardware • Payments • Software • Financial Services
Leads ambiguous, high-impact revenue strategy initiatives, including sales forecasting, pipeline health, capacity planning, organizational sizing, partnership modeling, and commercial deal economics. Builds repeatable operating processes, analyzes complex data, and delivers executive-ready recommendations. Partners across Sales, Finance, Partnerships, Services, Data, and Systems, influencing outcomes without direct authority.
Top Skills: ExcelGoogle SheetsSQL
4 Hours Ago
Remote or Hybrid
CA
Expert/Leader
Expert/Leader
eCommerce • Fintech • Hardware • Payments • Software • Financial Services
Leads sales compensation operations, including quota and crediting processes, commission calculations, data pipelines, audits, controls, governance, and payout accuracy. Manages and develops a high-performing team while administering Pigment and Salesforce, improving automation and data integrity, resolving exceptions, and partnering with Sales, Finance, and Revenue Operations leadership. Requires strong analytical, systems, stakeholder-management, and sales compensation expertise.
Top Skills: AnaplanCaptivateiqEtl PipelinesGoogle SheetsLookerExcelPigmentSalesforceSQLVaricentXactly
4 Hours Ago
Remote or Hybrid
CA
Expert/Leader
Expert/Leader
eCommerce • Fintech • Hardware • Payments • Software • Financial Services
Leads quota, capacity, and headcount planning across Block’s commercial sales organizations. Builds scalable planning methodologies, systems, data, governance, and processes; owns quota-setting and hiring models; conducts scenario planning; and advises Sales and Finance executives on growth investments. Partners with Compensation Design, Recruiting, Strategy, and Operations to align assumptions, incentives, role mix, hiring timing, productive capacity, and performance management across direct, channel, advertising, media, and account management motions.
Top Skills: LookerPigmentSalesforceSnowflake

What you need to know about the Calgary Tech Scene

Employees can spend up to one-third of their life at work, so choosing the right company is crucial, not just for the job itself but for the company culture as well. While startups often offer dynamic culture and growth opportunities, large corporations provide benefits like career development and networking, especially appealing to recent graduates. Fortunately, Calgary stands out as a hub for both, recognized as one of Startup Genome's Top 100 Emerging Ecosystems, while also playing host to a number of multinational enterprises. In Calgary, job seekers can find a wide range of opportunities.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account