Location Details: Canada - BC or ON (remote) At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join our team GoDaddy - Global Production Engineering looks after GoDaddy's global infrastructure, in the cloud and on-premises. We are hiring an experienced Technical Program Manager, focused on our AWS cloud infrastructure, to plan, lead and deliver complex cross-team initiatives. This is a heavily coordination-focused role: you will own the execution of a portfolio of AWS cloud platform and cost-savings programs, working hands-on with software engineers, engineering managers and SREs to achieve outcomes aligned with the strategy. You will drive dependencies end-to-end, facilitate trade-off decisions, and give collaborators and leadership clear, reliable access to status and risk. You will be an integral part of the Technical Program Management team, partnering closely with engineering leads to ensure GoDaddy delivers on its planned objectives and global strategy. What you'll get to do... Own end-to-end delivery of a portfolio of concurrent cloud platform programs, coordinating across engineering and partner teams to manage scope, schedule and dependencies against the critical path. Drive cost-savings program coordination, including tracking, reporting and surfacing risks to goals and achievements proactively. Run intake and prioritization processes and keep priority pages and status sources current and trustworthy. Serve as the central coordination point across teams, facilitating trade-off and negotiation discussions, driving alignment, and resolving roadblocks with minimal issues. Build reports, scorecards and dashboards to c
Jobs in Canada
Cloud Infrastructure Engineer Manager Manager Manager in Canada
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current cloud infrastructure engineer manager manager manager jobs across Canada. Filter by work mode, employment type, experience, department, date posted and distance.
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! We're looking for an Engineering Manager to lead our Deployment Engineering team. This isn't a typical management role — we need someone who leads from the front, gets their hands dirty, and drives impact. You'll manage a team of Forward Deployed Engineers who are on the front lines of deploying Cohere's North platform into customer environments. You should be ready to be a force to be reckoned with. Location: North America (remote-first) What You'll Do Lead and mentor a team of Forward Deployed Engineers Drive end-to-end deployment of North in private cloud and on-premises environments Take ownership of customer success from technical implementation through delivery Collaborate closely with Product, Engineering, and Sales to shape how we deliver AI to enterprises Mentor your team on cloud infrastructure, Kubernetes, and enterprise-grade deployments Optimize performance for OpenSearch, databases, and other K8s services Define scaling guidelines for GPU and CPU compute resources Build processes & technology that scales — we're growing fast. What We're Looking For 5+ years of experience in software engineering with demonstrate
At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 ClickUp is looking for an experienced Engineering Manager to lead our fullstack team responsible for building and scaling our flagship products. As the leader of the team that owns the APIs and core experiences powering ClickUp, you will play a pivotal role in shaping the future of our platform. You will guide engineers working across the stack, from frontend experiences to backend infrastructure. Your focus will be on driving the development of new features, addressing performance and reliability challenges, and ensuring operational excellence as we continue to grow. This is an opportunity to make a significant impact on our core product while fostering a culture of technical excellence and collaboration. The Role: Technical Leadership : Provide hands-on technical guidance to the team, ensuring best practices in software development, architecture, and design. Team Management : Lead, mentor, and grow a team of engineers, fostering a culture of collaboration, innovation, and continuous improvement. Product Development : Drive the development of new features and enhancements, ensuring high performance, scalability, and reliability. Collaboration : Work closely with product managers, designers, and other engineering teams to align on goals, prioritize initiatives, and deliver exceptional user experiences. Code Quality : Oversee code reviews, ensure adherence to coding standards, and advocate for clean, maintainable, and testable code. Innovation : Stay up-to-date with the latest trends and technologies in collaborative editing, cloud infrastructure, and web development, and apply them to improve our produ
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Opportunity: Okta Access Gateway (OAG) Enterprises run on a mix of modern cloud services and mission-critical on-premises systems (such as Oracle E-Business Suite, SAP, PeopleSoft, and custom legacy web apps). Okta Access Gateway (OAG) solves the enterprise hybrid cloud challenge by extending Okta’s cloud identity, Adaptive MFA, and Zero Trust security policies to on-premises and legacy applications without requiring custom code changes or traditional VPNs. As the Engineering Manager for Okta Access Gateway in Toronto, you will lead and grow a team of software engineers building the next generation of our hybrid access and gateway infrastructure. You will partner closely with Product Management, Architecture, Security, and Quality teams to deliver high-throughput, mission-critical security software deployed across multi-cloud and enterprise datacenters globally. What You’ll Do People Leadership & Team Growth Lead, mentor, and empower an engineering team, fostering an inclusive, high-performance, and psychologically safe engineering culture. Drive career progression, goal setting, regular 1:1s, and continuous feedback to help engineers grow their technical and leadership skills. Attract, interview, and hire diverse engineering talent to scale Okta’s engineering presence in Toronto. Delivery & Operational Excellence Own the end-to-end execution and delivery of key product roadmap initiatives, balancing feature velocity, technical debt, and softwar
From $252K/yr
About Scale AI Scale AI is the data foundation for AI, helping organizations build and deploy reliable production AI applications. We partner with leading enterprises and government organizations to accelerate their AI initiatives through our data annotation platform, generative AI solutions, and enterprise AI capabilities. Role Overview As a Forward Deployed AI Engineering Manager on our Enterprise team, you'll be the technical bridge between Scale AI's cutting-edge AI capabilities and our most strategic customers. You'll work with enterprise clients to understand their unique challenges, lead a team that architects specific AI solutions, and ensure successful deployment and adoption of AI systems in production environments. This is a Management role that combines deep engineering and AI expertise, leading a team, and working on customer-facing problems. You'll work directly with customer engineering teams to integrate AI into their critical workflows. Key Responsibilities Customer Integration & Deployment Partner directly with enterprise customers to understand their technical infrastructure, data pipelines, and business requirements Design and implement custom integrations between Scale AI's platform and customer data environments (cloud platforms, data warehouses, internal APIs) Build robust data connectors and ETL pipelines to ingest, process, and prepare customer data for AI workflows Deploy and configure AI models and agents within customer security and compliance boundaries AI Agent Development Develop production-grade AI agents tailored to customer use cases across domains like customer support, data analysis, content generation, and workflow automation Architect multi-agent systems that orchestrate between different models, tools, and data sources Implement evaluation frameworks to measure agent performance and iterate toward business objectives Design human-in-the-loop workflows and feedback mechanisms for continuous agent improvement Prompt Engineeri
From $24K/yr
Amplitude is the leading AI analytics platform, helping over 4,700 customers—including Atlassian, Burger King, NBCUniversal, and Square—build better products and digital experiences. With powerful AI Agents embedded across our platform, teams can analyze, test, and optimize user experiences faster than ever. Ranked #1 across multiple categories in G2’s Winter 2026 Report, Amplitude is the best-in-class solution for product, data, and marketing teams. Learn more at amplitude.com . As an organization, we deliver for our customers by living our values. We operate from a place of humility, take ownership of problems and successes, approach challenges with a growth mindset, and put our customers at the center of everything we do. Amplitude’s Commitment to Diversity Equity & Inclusion (DEI): Amplitude believes that diversity enables the creation of better products, improves the ability to solve complex problems, and drives more powerful solutions. We strive to create an environment of inclusion—one focused on psychological safety, empathy, and human connection—that will allow employees of all backgrounds to thrive. About the Role & Team We’re looking for an Engineering Manager to lead the Data Infrastructure team within Statsig Experiment at Amplitude. You will lead a multidisciplinary team of software engineers, data engineers, and data scientists responsible for the systems that power experimentation at scale. The team owns three critical areas: Data ingestion: Collecting and importing experiment exposures, custom events, OpenTelemetry data, and real user monitoring data across SDKs, streaming systems, cloud storage, and customer data warehouses. Data computation: Building distributed computation systems that transform raw data into accurate, timely experiment results. Stats engine: Developing and productionizing the statistical methods that help customers make trustworthy decisions from their experiments. This is not a traditional data engineering management ro
About the Team DoorDash Labs is an independent team within DoorDash. We explore robotics and automation to transform last mile logistics in the long term. If you have a passion for applying robotics solutions in a service used by millions of people, then we want to talk to you! About the Role We're hiring a Robotics Infrastructure Engineer in our Autonomy Software team. In this role, you'll own, build, and manage the infrastructure that makes aerial autonomy development possible. You'll work on the onboard systems that keep a drone alive (process management, health monitoring, parameterization) and the development environment that makes the team fast (build systems, CI/CD, logging, debugging, regression testing). This is not cloud infrastructure. This is real-time, fault-tolerant, onboard software for vehicles that cannot gracefully restart at 50 meters altitude. You're excited about this opportunity because you will… Play an integral role on a small and focused team Develop and own critical onboard components: process management, health monitoring, configuration management, and message passing Own the build system (C++, Python) and middleware layer (ROS2), including cross-compilation for Jetson targets Design and maintain CI/CD pipelines and regression testing infrastructure Build and manage the parameterization system, including schema definition, validation, migration, and deployment Build robotics logging, plotting, and debugging tools that make the entire team more productive Work closely with the simulation team to support SIL/HIL development workflows Define reliability standards for onboard software: watchdogs, failover, and graceful degradation We're excited about you because… You have prior experience at a robotics company in a similar infrastructure role You have experience with robotics middleware (ROS2, LCM, eCal, Apex.AI) You have experience with build systems and package managers (CMake, Bazel, Nix, Conan) You have experience with NVidia Jetson and Je
From C$107K/yr
Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. Remote: This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. About the Team Global Compute builds and operates the core cloud infrastructure that engineering teams rely on every day. We provision and manage AWS accounts across the company, operate the network backbone that connects them, and maintain the security guardrails that keep those environments safe, compliant, and scalable. We believe reliability is an engineering challenge, not an operations task. We automate repetitive work, build for scale before it becomes a problem, and invest heavily in observability to identify issues before they impact the business. What you'll get to do... Operate and scale AWS production infrastructure, owning the health of services that provision, secure, and manage accounts across GoDaddy AWS organisations. Design, build, and maintain cloud platform capabilities using Python, CloudFormation, AWS CDK, and automation-first practices. Drive cost optimisation initiatives that improve efficiency and deliver measurable business impact. Improve observability through monitoring, alerting, dashboards, and operational tooling. Participate in on-call rotations, lead incident response efforts, and drive long-term reliability improvements through blameless post-incident reviews. Support strategic AWS initiatives across networking, identity, governance, and multi-account architecture. Review code and designs, contribute documentation and operational runbooks, and mentor fellow engineers. Leverage AI-assisted tooling to improve engineering productivity, accelerate automation, and reduce operati
About Forma.ai: Forma.ai is a Series B startup that's revolutionizing how sales compensation is designed, managed and optimized. We handle billions in annual managed commissions for market leaders like Edmentum, Stryker, and Autodesk. Our growth has been fuelled by our passion for fundamentally changing and shaping how companies use sales intelligence to drive business strategy. We’re welcoming equally driven individuals who are excited about creating something big! The Opportunity As a Staff Security Engineer, you will be a hands-on technical leader strengthening security across Forma's application, cloud infrastructure, development lifecycle, internal systems, and incident-response practices. Security today is shared across Engineering and DevOps. You'll work closely with both teams and have real room to shape how Forma approaches security as we grow. Depending on your interests and the needs of the business, the role could develop into a deeper individual-contributor position or help build a dedicated security team. You'll work directly with Engineering, DevOps, IT, Product, Legal, and Privacy to identify risks, design practical controls, automate security processes, and help teams ship secure and reliable software. What you'll do Cloud and infrastructure security Design and implement security controls across Forma's AWS environments, with a focus on IAM, least-privilege access, service identities, and account boundaries. Embed security requirements into Terraform and other Infrastructure as Code, and improve secrets, certificate, encryption-key, and credential management. Build automated checks for insecure configurations, excessive permissions, exposed resources, and configuration drift across Kubernetes, containers, serverless workloads, networking, and data services. Application, data, and AI security Run threat modelling and security architecture reviews for new products, services, APIs, data pipelines, and third-party integrations
Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Corporate Systems team focuses on maintaining secure, reliable systems that support employees across the company. This team works closely with Security, IT, and Engineering partners to manage identity systems, endpoint devices, and cloud infrastructure. The goal is to ensure systems are scalable, dependable, and prepared to address evolving security risks. As an Application Engineer, you will manage and improve systems that support identity, device management, and cloud infrastructure. You will build automation to streamline account and device lifecycle processes, respond to technical issues, and contribute to engineering standards through code reviews and documentation. You will also use AI-assisted tools to enhance development workflows and improve efficiency. This role is based in our Menlo Park, CA office(s), with in-person attendance expected at least 3 days per week. At Robinhood, we believe in the power of in-person work to accelerate progress, spark innovation, and strengthen community. Our office experience is intentional, energizing, and designed to fully support high-performing teams. What you’ll do You manage endpoint devices, identity systems, Okta, Go
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! As a Senior Security Operations Engineer you will: Serve as trusted advisor to team’s leadership and partner teams by clearly articulating business risks associated with security issues Harden our cloud-native environments (AWS, OCI, GCP) by introducing secure by default designs and features into network, tooling, and processes Own and drive resolutions for enabling engineers to design, build, and use infrastructure securely at scale by deploying secure architectures using infrastructure-as-code and reusable code libraries Manage IAM / RBAC for cloud infrastructure, and partner with IT on streamling authentication/authorization to ensure unified access control across the board Deploy and operationalize some of the security services and tools (eg: SIEM, SOAR, domain monitoring, endpoint tooling, cloud security tooling) Respond to security incidents and harden environments post-incidents. Support control monitoring and remediation for compliance initiatives Gather and analyze security metrics to address security issues with cross-team dependencies Be a problem solver who is empathetic to developer concerns and will employ construc
C$125K – C$200K/yr
We deliver foundational systems that shape the future of how technology is used in a top performing quantitative equity fund that manages over $78+ billion USD in financial assets. This is a fantastic opportunity in the exciting intersection of finance and technology where investment decisions are made using technology. Quantitative equity funds use programmed investment strategies and as a result, our technology team is crucial to its success. The team is headquartered and deeply rooted in West Coast Vancouver. We place high value on maintaining an entrepreneurial spirit and creating a culture where each of us has opportunities to succeed. What You Will Do The technology infrastructure team plays an essential role through innovative technologies on our hybrid (on-premise and cloud based) platform: distributed computing, petabyte-scale data storage, containerization, non-traditional high-performance databases, process orchestration, monitoring, data visualization and DevOps. You own the entire technology infrastructure life cycle: Engineer and support software and systems infrastructure. Introduce new foundational technologies that advance our software engineering capabilities to the next level. Collaborate with our software development teams on support issues and improvements to our infrastructure tools, processes, and software. Act as a conduit between our application development teams, and IT, network security, and other stakeholders to align priorities and translate business requirements into technical designs. Improve systems infrastructure reliability. Gather and analyze metrics from operating systems and applications to assist in performance tuning, fault finding and business continuity planning. Design, plan and implement solutions in an entrepreneurial spirit. What You Bring Programming Knowledge – you have an undergraduate, graduate, or post-graduate degree in a computer-related field OR exceptional programming skills gain
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Security Clearance: Active Secret+ clearance strongly preferred; candidates eligible and willing to obtain clearance will also be considered. More information about Canadian Security Clearance is available here . As an Infrastructure Security Engineer, your key responsibilities include: Deploy, and manage infrastructure for Protected B classified environments, ensuring compliance with ITSG-33 and Canadian government standards Design and implement security controls for cloud (AWS, GCP, Azure) and hybrid/multi-cloud deployments Evaluate, implement, and manage security tools and technologies for training cluster and inference infrastructure hardening Implement security best practices including IAM, encryption, logging, and monitoring Participate in security incident response activities, including detection, analysis, containment, and remediation Conduct regular vulnerability assessments and penetration testing of infrastructure components Maintain comprehensive security documentation, procedures, and configurations for classified environments Maintain active Secret+ security clearance and adhere to all Canadian government security
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this team? The internal infrastructure team is responsible for building world-class infrastructure and tools used to train, evaluate and serve Cohere's foundational models. By joining our team, you will work in close collaboration with AI researchers to support their AI workload needs on the cutting edge, with a strong focus on stability, scalability, and observability. You will be responsible for building and operating superclusters across multiple clouds. Your work will directly accelerate the development of industry-leading AI models that power Cohere's platform North. Please Note: All of our infrastructure roles require participating in a 24x7 on-call rotation, where you are compensated for your on-call schedule. As a Staff Software Engineer, you will: Build and scale ML-optimized HPC infrastructure : Deploy and manage Kubernetes-based GPU/TPU superclusters across multiple clouds, ensuring high throughput and low-latency performance for AI workloads. Optimize for AI/ML training : Collaborate with cloud providers to fine-tune infrastructure for cost efficiency, reliability, and performance , leveraging technologies like R
From $216K/yr
The Public Sector software engineers (SWEs) create the core product building blocks forward-deployed teams use to develop agentic capabilities that function across multiple domains. SWEs responsibilities include building the systems required to ingest and process federal datasets to support real-time decision-making in contested environments. We develop novel agentic enabling capabilities that includes: Create multi-layered guardrails around agents Optimize data retrieval for agents Orchestrate fleets of asynchronous agents Automatically alerts users to deviations in data Illustrating how an agent reached a decision As a Senior Software Engineer, you will lead the development of a vertical feature or a horizontal capability to include defining requirements with stakeholders and implementation until it is accepted by the stakeholders. You will: Lead the design and implementation of scalable backend systems and distributed architectures for Federal customers. Manage the full lifecycle of feature development from requirement definition to deployment on classified networks. Direct the orchestration of asynchronous agent fleets to meet mission requirements. Lead customer engagements to translate mission needs into technical requirements. Own the communication with stakeholders to ensure implementation meets defined acceptance criteria. Conduct technical reviews and identify risks within machine learning infrastructure and model serving. Drive the platform roadmap by providing technical specifications for Federal product offerings. Ideally you will have: Full Stack Development: Proficiency in front-end, back-end development and infrastructure, including experience with modern web development frameworks, programming languages, and databases Cloud-Native Technologies: Familiarity with cloud platforms (e.g., AWS, Azure, GCP) and experience in developing and deploying applications in a cloud-native environment. Understanding of containerization (e.g., Docker) and contai
Other cities to consider
More places hiring for this role
Get new cloud infrastructure engineer manager manager manager jobs in Canada by email
Daily job updates · Unsubscribe anytime