Jobiba hiring network

Lead Cloud Operations Engineer Jobs

6,876 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current lead cloud operations engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

O
Okta
📍 India• Full-time
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. We’re redefining Privileged Access Management (PAM) from the ground up, purpose-built for Cloud, SaaS, Databases, Containers, and any virtualized environment. Our mission is to simplify and secure workforce access with seamless, secure-by-default workflows that adapt dynamically to modern infrastructure. We eliminate standing privileges, enforce least privilege, and embed Zero Trust principles into every access workflow by default. About the Role We are seeking a Principal Backend Engineer (P5) to serve as the technical leader and compass for our newly established engineering pod in India. Operating at the intersection of identity, networking, and security infrastructure , you will be responsible for tackling highly complex, vaguely specified problems without day-to-day oversight. In this role, you will champion the technical execution of your team. You will lead the design and implementation of secure database and network device connectors (routers, switches, firewalls) on top of our core Zero Standing Privileges (ZSP) platform. You will work closely with our local Technical Team Lead to mentor mid-level engineers, while collaborating closely with global Tech Leads to ensure architectural alignment. What You’ll Be Doing Lead Technical Strategy & Execution: Turn high-level, complex PAM and network access problems into clear, modular technical designs that your team can execute. Own Connector Ecosystems: Design, architect, and optimize resilient connecto

javasqlaws
View job →
N
Notion
📍 San Francisco• Full-time• $185K – $220K/yr
1mo ago

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: We are seeking a strategic and technically fluent Lead, IT Audit to join our Finance team reporting to the Head of Internal Audit. This is a broad, high-impact role spanning both IT SOX compliance and operational IT audits. You will help establish and elevate our technology controls program end to end — owning the IT SOX lifecycle, designing the IT general and application controls framework, embedding AI and automation into how we test and monitor controls, and delivering value-added operational IT and cybersecurity audits that strengthen how the company builds and runs its systems. You will partner with leaders across Engineering, Security, IT, Finance, and the business to ensure sound technology controls are built into how the company operates as we scale. This role is ideal for someone who thinks like a builder, not just an auditor — someone who can translate complex control and security requirements into practical, scalable processes in a fast-moving SaaS environment with modern cloud architecture and complex data flows. This role can be based in either San Francisco or New York City. We work from our offices on M

awsazuregcp
View job →

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Staff Software Engineer - External Observability Platform Location: Bellevue, WA (Hybrid: 3 days/week in-office) Team: Infrastructure & Observability Platform Engineering About the Role Snowflake’s Data Cloud processes exabytes of data across multi-cloud global environments every day. Delivering seamless reliability and real-time visibility to thousands of global enterprise customers requires an Observability Platform built on hyper-scalable backend distributed systems. We are seeking a Staff / Lead Software Engineer to architect, design, and scale our External Observability Platform . In this role, you will lead the technical strategy for customer-facing telemetry, system metrics, audit logs, distributed tracing, and actionable operational insights. You will build high-throughput, low-latency infrastructure capable of ingesting, processing, and serving petabytes of telemetry data with strict SLA guarantees. You will join a team of world-class engineers in our Bellevue, WA office. To be successful, you must be deeply technical, capable of leading complex cross-functional architecture initiatives, and skilled at mentoring senior engineers while holding your own with the brightest technical minds in the industry. Key Responsibilities Architect & Scale Distributed Infr

javavueaws
View job →

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is an AI-powered observability platform built on the Snowflake AI Data Cloud and engineered for scale. We ingest and store logs, metrics, traces, and events on an open, scalable data lakehouse, using open formats like Apache Iceberg, at dramatically lower cost. A dynamic Context Graph and chat-based AI SRE provide rich context and automated workflows so teams can move from detection to root cause of production issue and resolution 10x faster. Leading engineering teams at companies like Capital One, Topgolf, and Dialpad rely on Observe to troubleshoot hundreds of terabytes of telemetry daily while maintaining reliability at enterprise scale. As part of Snowflake, Observe combines startup-style ownership and velocity with the global reach, operational excellence, and ecosystem of one of the world’s leading data platforms. In this role you will: Develop interactive, data-rich user interfaces using React, TypeScript, and Vega, with a focus on integrating LLM-driven features (e.g., natural language querying, generative UI, and AI-assisted data storytelling). Lead the end-to-end delivery of substantial product features, ensuring AI outputs are presented with high reliability and low latency. Work closely with PMs, UX designers, and AI/ML engineers to bridge t

javascripttypescriptjava
View job →
E
9 days ago

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE: Everpure is looking for a dynamic Systems Engineering Manager to lead and inspire a high-performing team in our SLED organization. In this pivotal role, you’ll guide talented pre-sales Systems Engineers who help some of the world’s largest and most innovative companies modernize their data strategies with industry-leading all-flash storage, cloud, and AI-driven data solutions. If you’re passionate about coaching technical talent, elevating customer outcomes, and shaping the future of data-driven enterprises, this is your opportunity to make a visible impact at one of the fastest-growing companies in the industry. WHAT YOU’LL DO: Demonstrate effective people development in a rapidly evolving technical space Act as an employee advocate who is focused on team and people development, demonstrating high EQ Have a desire to drive operational excellence; create a team support creation and validation of technical demand Lead by example, living the Everpure Values: Persistence, Creativity, Teamwork, Ownership, Customer-first Develop an exhaustive understanding of what drives a customer’s business and what motivates their decision making Passionately bring to light the advantages of a Pure Storage solution Refine sales strategy and tactics, taking command of technical responsibilities Delight customers and teammates with your technical leadership and domain expertise on storage products, distributed storage architect

REMOTEaisupply chain
View job →
W-
Wolt - English
📍 Helsinki• Full-time
16 days ago

About Wolt At Wolt, we create technology that brings joy, simplicity and earnings to the neighborhoods of the world. In 2014 we started with delivery of restaurant food. Now we’re building the delivery of (almost) everything and you’ll find us in over 500 cities in 30 countries around the world. In 2022 we joined forces with DoorDash and together we keep on dreaming big and expanding across the globe. Working at Wolt isn’t always easy, but it’s definitely exciting. Here you’ll learn more, build more, and ship more than in most other companies. You’ll be challenged a lot, but also have a lot of fun on the way. So, if you’re a self-starter with drive and entrepreneurial spirit, this could be the ride of your life. What you’ll do: Build and maintain high-throughput backend services using Go . Collaborate with product managers, designers, and frontend developers to ship features that support internal support agents across the globe. Design systems that are scalable , resilient , and easy to maintain. Lead and contribute to architectural discussions and technical decision-making. Write well-tested code and help the team maintain high code quality standards. Our humble expectations: 7+ years of professional software engineering experience, with a proven track record of building and scaling complex systems. 2+ years of production experience in Golang , with the ability to mentor others and drive best practices across the team. Strong hands-on experience with both SQL and NoSQL databases — especially Cassandra. Solid understanding of designing and operating low-latency, high-throughput distributed systems . Nice to have Background in Node.js or other backend languages. Familiarity with cloud infrastructure (AWS, GCP) and event-driven architectures. Previous on-call experience , with a pragmatic approach to reliability and incident management. What we value A product-oriented mindset — you think beyond the ticket, understand the “why” behind the work, and aim to create real

node.jssqlaws
View job →
S
1mo ago

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are looking for a talented and passionate Staff Software Engineer for our Snowpark Container Service , part of our Snowflake Compute Platform - to build our elastic, high-scale, high-performance, cloud native compute platform to enable bringing Compute to Data effortless and simple. Snowpark Container Services is a fully managed container offering that helps our customers easily deploy, manage, and scale containerized applications without having to move data out of Snowflake. As a fully managed service, it comes with Snowflake security, configuration, and operational best practices built in. You will be part of this highly productive, fast moving, and growing team that is critical to realizing Snowflake’s Data Cloud Mission. AS A STAFF SOFTWARE ENGINEER, YOU WILL: Design and develop features, understand customer requirements and meet business goals. Lead a team of engineers, including mentoring and guiding them, and build technical direction and strategy for large and critical parts of the product surface area. Manage all aspects of the Project, including Design, Coding, Reviews, Testing, Observability, Tooling and On-Call support. Build highly reliable and fault-tolerant software to meet the needs of the largest customers. Ensure operational readiness and maintainabilit

REMOTEjavavueaws
View job →

Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity Postman is seeking a strategic and results-driven engineering leader who is passionate about cloud agnostic infrastructure, operational excellence, and enabling engineering teams to operate autonomously and build with confidence. As Head of Infrastructure, you'll lead a talented and geographically distributed team of engineers across the SF Bay Area, India, and Europe, fostering a culture of collaboration, ownership, and continuous improvement. You'll own the infrastructure that underpins one of the world's most widely used API platforms, an environment handling ~80,000 requests per second at the front door, and be responsible for its reliability, scalability, and evolution. In addition to infrastructure, you'll own the Site Reliability Engineering (SRE) function at Postman, setting the standards and practices that keep the platform reliable at scale. You'll work closely with engineering managers, product managers, and platform teams to drive the technical roadmap for our cloud agnostic infrastructure and reliability practices, ensuring we can support a large and rapidly growing engineering organization. If you're p

awsazurekubernetes
View job →
C
Coinbase
📍 - USA• Full-time• Remote• From $218K/yr
1mo ago

Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . Team/ Role: Coinbase is building the future of institutional trading as part of the Institutional Everything Exchange, powering the systems that let the world's largest financial institutions trade across markets with speed, depth, and trust. As an Infrastructure Lead on the Exchange team, you'll own the infrastructure, deployment, and operational tooling behind a unified trading platform that runs across cloud and on-prem/colocated venues, keeping latency-sensitive, multi-node trading environments fast, reliable, and easy to build on. It's a rare chance to shape trading infrastructure from the ground up, with the ownership of a startup and the reach of Coinbase. What you’ll do: Own infrastructure, deployment, and operational tooling for latency-sensitive, multi-node trading environments across cloud and on-prem. Drive reliability and developer velocity, owning observability, deploy safety, and incident response for systems that run 24/7. Set the operational bar through standards, reviews, and automation, and reduce single-point-of-failure risk across the trading stack. Partner with engineers building the trading platform to make latency-sensitive systems operable and performant. Mentor engineers and build team resilience across a lean, high-impact environment. Partner cross-functionally with Product, Institutional Markets, and SRE to turn platform needs into a

REMOTEawslinuxai
View job →
NR
New Relic
📍 Atlanta• Full-time• From $186K/yr
1mo ago

We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity Are you ready to step into a pivotal leadership role where your engineering depth directly shapes the future of our core platform? As our new Engineering Manager, you will lead a talented, distributed team across US and EU time zones, acting as the critical manager bridging regional collaboration. Our Cloud Foundation team is the backbone of the New Relic platform. In this role, you won't just manage tasks; you will mentor and empower engineers, transitioning our operational framework from a reactive state to a culture of proactive ownership and engineering excellence. You will oversee critical global initiatives, including major regional expansions into FedRAMP High / IL4, India, and Australia. If you thrive on solving complex multi-cloud challenges at an exabyte scale while helping engineers grow in their careers, this is your opportunity to make a lasting impact. What you'll do Empower & Mentor: Lead and nurture a high-performing engineering team across the US and EU, facilitating career development, performance growth, and a collaborative team culture. Drive Strategic Ownership: Champion a shift from reactive delivery to proactive technical ownership, establishing best practices for platform reliability and cross-regional alignment. Lead Regional Expansions: Architect and execute key global infrastructure expansions across complex environments (including FedRAMP High / IL4, India, and Australia). Architect for Extreme Scale: Guide decisions around micr

reactawsazure
View job →
M
Mindbody
📍 United States• Full-time• $170K – $250K/yr
1mo ago

At Playlist, life's richest moments happen when people step away from screens to move, connect, explore, and play. We're building the definitive platform for intentional living, connecting people with inspiring experiences in fitness, wellness, and beyond. With popular brands like Mindbody and ClassPass, Playlist empowers businesses and individuals, making it effortless for aspirations to become actions. Join us in reshaping technology's role to foster meaningful, real-world connections. Mindbody equips wellness entrepreneurs with technology to support thriving businesses and create exceptional experiences. Innovation and curiosity drive our culture, connecting businesses and individuals through cutting-edge solutions. Join us if you're passionate about enhancing wellness through technology. The Role You'll Play At Mindbody, Core Engineering builds and evolves the foundational systems that help our products run reliably at scale. In this staff role, you’ll bring clarity to complex technical problems, guide architecture, and strengthen how we design, deliver, and operate the backend services that power real-world experiences. Lead cross-team technical execution, aligning architecture and delivery across multiple squads and core domains Design and evolve microservices patterns that improve reliability, performance, and maintainability Drive cloud and deployment improvements across AWS, Mindbody’s cloud platform, and our containerized deployment environment Partner with engineering and product leaders to turn ambiguous problems into clear technical plans and milestones Establish and socialize standards for service design, APIs, and relational data modeling (SQL) Strengthen monitoring and operational visibility using New Relic and Kibana, turning insights into durable system improvements Mentor and unblock engineers through design reviews, pairing, and practical guidance Reduce technical risk and complexity while balancing

pythonreactsql
View job →
M
Mindbody
📍 Brazil• Full-time
1mo ago

At Playlist, life's richest moments happen when people step away from screens to move, connect, explore, and play. We're building the definitive platform for intentional living, connecting people with inspiring experiences in fitness, wellness, and beyond. With popular brands like Mindbody and ClassPass, Playlist empowers businesses and individuals, making it effortless for aspirations to become actions. Join us in reshaping technology's role to foster meaningful, real-world connections. Mindbody equips wellness entrepreneurs with technology to support thriving businesses and create exceptional experiences. Innovation and curiosity drive our culture, connecting businesses and individuals through cutting-edge solutions. Join us if you're passionate about enhancing wellness through technology. The Role You’ll Play As a Senior Platform Engineer on Playlist, you’ll design and deliver Infrastructure-as-Code solutions that empower developer teams. You’ll drive cloud architecture, iterate with squads on their workloads, and build self-service tools to speed delivery and improve quality. Your work will help design, implement, and operate the cloud infrastructure that powers the Mindbody ecosystem and supports millions of users. Partner with Product and Engineering to design, build, and operate the cloud platform that enables squads to deliver reliably and autonomously. Own and evolve our production Kubernetes platform and core cloud primitives, driving safe, automated, and observable infrastructure-as-code. Deliver self-service tooling and IaC patterns so teams can provision and run workloads without platform intervention. Lead cross-team projects from problem definition through architecture, implementation, and launch while reducing operational toil and improving reliability and security. Be the go-to engineer for production incident response, runbook automation, and platform change management across segmented and compliance-bound environments. Experience You Bring

typescriptpythonaws
View job →
M
Mindbody
📍 Brazil• Full-time
1mo ago

At Playlist, life's richest moments happen when people step away from screens to move, connect, explore, and play. We're building the definitive platform for intentional living, connecting people with inspiring experiences in fitness, wellness, and beyond. With popular brands like Mindbody and ClassPass, Playlist empowers businesses and individuals, making it effortless for aspirations to become actions. Join us in reshaping technology's role to foster meaningful, real-world connections. Mindbody equips wellness entrepreneurs with technology to support thriving businesses and create exceptional experiences. Innovation and curiosity drive our culture, connecting businesses and individuals through cutting-edge solutions. Join us if you're passionate about enhancing wellness through technology. The Role You’ll Play As a Senior Platform Engineer on Playlist, you’ll design and deliver Infrastructure-as-Code solutions that empower developer teams. You’ll drive cloud architecture, iterate with squads on their workloads, and build self-service tools to speed delivery and improve quality. Your work will help design, implement, and operate the cloud infrastructure that powers the Mindbody ecosystem and supports millions of users. Partner with Product and Engineering to design, build, and operate the cloud platform that enables squads to deliver reliably and autonomously. Own and evolve our production Kubernetes platform and core cloud primitives, driving safe, automated, and observable infrastructure-as-code. Deliver self-service tooling and IaC patterns so teams can provision and run workloads without platform intervention. Lead cross-team projects from problem definition through architecture, implementation, and launch while reducing operational toil and improving reliability and security. Be the go-to engineer for production incident response, runbook automation, and platform change management across segmented and compliance-bound environments. Experience You Bring Senior

typescriptpythonaws
View job →
D
1mo ago

Husky is what we call the distributed, petabyte-scale columnar event store at the heart of our Event Platform, which powers dozens of Datadog’s most popular products – Logs, RUM, APM , Cloud Network Monitoring, Netflow, and many more. Husky was built from the ground up at Datadog to store and query massive volumes of event data at low cost, with data fully queryable within seconds of arrival. As an Engineering Manager on Husky, you will lead one of a few closely collaborating teams that work directly on Husky’s internals. We own the full lifecycle of events stored by our Event Platform – whether it’s managing the ingestion of over 100 million events per second exactly-once , optimizing the persistence layer with lightning-fast compaction , tweaking our columnar format, timely deletion across hundreds of thousands of customer tables, or building the metadata service that enables hundreds of thousands of queries per second – you will be responsible for the growth and success of a team constantly meeting new scalability requirements driven by Datadog’s growing customer base and suite of products. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Manage a team of software engineers ranging from new grads to senior engineers, focusing on career development, inclusivity, and high impact Be a technical leader of the team; while the role may not involve frequent hands on coding, you’ll be reviewing architecture decisions, RFCs, and pull requests and ensuring what the team ships is high quality Partner with product teams to evaluate use cases, ensure smooth implementation, and prioritize new platform capabilities. Share on-call responsibilities with the rest of the team and ensure a culture of operational exc

javarestai
View job →
F
Figma
📍 Ca New York• Full-time• From $258K/yr
1mo ago

Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI. From idea to product, Figma empowers teams to streamline workflows, move faster, and work together in real time from anywhere in the world. If you're excited to shape the future of design and collaboration, join us! Figma's AI Tools team builds the AI-powered workflows, platforms, and tooling that make every engineer at the company more productive. From agentic CI auto-fixing and AI-assisted code review to background cloud agents that turn a Slack message into a pull request, this team owns the systems that are transforming how Figma builds software. The team operates across a three-layer platform stack—sandbox runtime, cloud agents, and workflow orchestration—while shipping and maintaining a growing portfolio of org-wide developer workflows that compound across hundreds of engineers. This is a full time role that can be held from one of our US hubs or remotely in the United States. What you’ll do at Figma: Lead and grow a team of engineers responsible for building and operating Figma's AI developer workflows and the cloud agent platform that powers them Own the technical strategy and roadmap for AI Developer Experience, spanning sandbox runtime infrastructure, cloud agent reliability, workflow orchestration, and org-wide agentic workflows Hire and scale the team - establishing team culture, execution cadence, and operational processes from the ground up Drive the reliability, observability, and scalability of our cloud agent platform, ensuring it meets production-grade standards as adoption grows across the engineering organization Partner with product engineering, security, infrastructure, and DevEx teams to identify the highest-leverage opportunities for AI-assisted developer workflows and drive adoption

awsci/cdai
View job →
🔔

Get new lead cloud operations engineer jobs by email

Daily job updates · Unsubscribe anytime