Jobiba hiring network

Lead Cloud Operations Engineer Jobs

6,753 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current lead cloud operations engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

M
Mongodb
📍 United States• Full-time• From $151K/yr
1mo ago

MongoDB’s Security Product Management team is seeking a Staff Product Manager to own security, compliance, and public sector product strategy within Atlas for Government (A4G). In this role, you will be a key member of the security product management team, helping make data security a market differentiator that enables MongoDB to win in enterprise and regulated industries. You will lead product strategy and execution for capabilities and programs that support federal compliance requirements and broader regulated-industry needs. You will work across Engineering, Compliance, Legal, Security, and Go-to-Market teams to translate complex regulatory and customer requirements into clear product investments, roadmaps, and outcomes. This role is suited for a candidate who can orchestrate multiple interlinked product areas, own cross-team product and architectural trade-offs, and align senior stakeholders around multi-year bets. You will be expected to identify opportunities and dependencies that cut across team boundaries, bring clarity to ambiguous problem spaces, and establish reusable ways of working that help MongoDB deliver secure, compliant platform capabilities for public sector and other regulated markets. You will partner closely with senior engineering and business leaders to evaluate trade-offs, sequence investments, and bring clarity to decisions that balance customer impact, execution risk, and long-term business value. This role can be based out of of our offices or remotely in the United States. The Federal Risk and Authorization Management Program (FedRAMP) is a US government-wide program that provides a standardized approach to security assessment, authorization, and continuous monitoring for cloud products and services. Our FedRAMP program requires that anyone who is accessing customer data or metadata inside the Authorization Boundary be a US Person on US Soil. Responsibilities Security and Compliance Product Strategy Own the public sector compliance roadm

mongodbawsazure
View job →

As a TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB’s cloud products. You will drive program execution, strengthen production reliability practices, and coordinate cross-functional efforts across US and EMEA teams. Success in this role means smoother launches, clearer roadmaps, stronger reliability metrics and an SRE organization that's better-equipped to deliver predictability at scale. This role can be based remotely on the East Coast What You'll Do Drive Program Planning & Execution – Define program scope, milestones, and success criteria with SRE engineers and leaders. Manage dependencies across platform teams, keep work clearly tracked in Jira, and deliver on time Strengthen Production Reliability – Lead change management and launch readiness programs. Partner with SREs and product teams to define and operationalize SLOs/SLIs, and use incident data, metrics, and capacity signals to drive prioritization and continuous improvement Lead Cross-Functional Coordination – Align SRE with Security, Compliance, Cloud platform, and other engineering teams. Coordinate cross-team incident response, ensure clear follow-through, and build trust as the go-to driver of complex, multi-team efforts Build Scalable Systems & Processes – Design lightweight frameworks and communication patterns that help SRE deliver reliably at scale. Work yourself out of the "hero" role by leaving teams better-equipped to execute independently Requirements 8+ years in technical program management, engineering management, or a comparable technical role partnering with software engineering teams Proven track record leading large-scale, cross-team platform initiatives through ambiguity and change Strong knowledge of production change management, software development lifecycle, and reliability metrics (SLOs, SLIs) Skilled at shaping roadmaps and managing dependencies Able to query and interpret metrics, logs, or other data s

mongodbawsazure
View job →
M
Mongodb
📍 San Francisco• Full-time• From $126K/yr
1mo ago

Join the Atlas Search Query team to design and develop the next generation of Search query architecture, optimization, and execution. Atlas Search is a growing cloud service that allows users to execute complex search and vector search queries using the MongoDB Query Language. Our users can focus on relevance and data retrieval instead of the machinery needed to search data at scale. Our team is building a cloud-based distributed system responsible for the core components of search including data ingestion, performance, query language, query execution, for both relevance-based search and vector search. Our product is being adopted quickly and there are many interesting projects. This is a technical role where you will be responsible for the success of complex Search Query feature development. We are looking to speak to candidates who are based in San Francisco, CA for our hybrid working model. What You’ll Do Lead complex projects across the MongoDB ecosystem, for instance, development of a new Search aggregation framework within the MongoDB aggregation framework. Set project level strategy, architect features, and lead projects to successful execution Identify, design, and implement features enhancing our query language, performance, and operability Perform code reviews with peers and make recommendations on how to improve our software development processes Influence and grow team members through active mentoring and leading by example What We Look For 5+ years experience in data management/search systems, ideally with a strong query processing and optimization background Experienced in the development and maintenance of stateful distributed systems Eager to shape the technological direction of a complex system and have the ability to lead initiatives through collaboration with others Experienced in debugging and profiling multithreaded applications written in Java and Rust Bonus: experience with designing high-volume query engines, such as a datab

javamongodbaws
View job →
M
Mongodb
📍 San Francisco• Full-time• From $126K/yr
1mo ago

Join the Atlas Search team to design and develop the next generation of Semantic and Vector Search infrastructure. Atlas Search is a growing cloud service that allows users to execute complex search and vector search queries using the MongoDB Query Language. Our users can focus on relevance and data retrieval instead of the machinery needed to search data at scale. Our team is building a cloud-based distributed system responsible for the core components of search including data ingestion, performance, query language, query execution, for both relevance-based search and vector search. Our product is being adopted quickly and there are many interesting projects. This is a technical role where you will be responsible for the infrastructure and features enabling our at-scale cloud service powering vector and semantic search. We are looking to speak to candidates who are based in the San Francisco Bay Area for our hybrid working model. What You’ll Do Lead complex projects across the MongoDB ecosystem, for instance, development of a new Search deployment framework within the MongoDB managed cloud Set project level strategy, architect features, and lead projects to successful execution Identify, design, and implement features enhancing our reliability, performance, security and efficiency Perform code reviews with peers and make recommendations on how to improve our software development processes Influence and grow team members through active mentoring and leading by example What We Look For 5+ years experience in data management/search systems, ideally with a strong distributed systems and infrastructure background Experienced in the development and maintenance of concurrent, stateful services Eager to shape the technological direction of a complex system and have the ability to lead initiatives through collaboration with others Experienced in writing features, debugging and optimizing multithreaded applications written in Java Familiarity with LLM

javamongodbaws
View job →
M
Mongodb
📍 Austin• Full-time• From $90K/yr
1mo ago

MongoDB Professional Services serves as the delivery owner for complex customer initiatives, translating strategic goals into disciplined execution and measurable business outcomes across modernization, migration, cloud, data, and AI-related programs. The team drives delivery accountability across customers and internal stakeholders through strong governance, proactive risk management, and sustained alignment to customer priorities. About the Role MongoDB is looking for a Senior Project Manager to lead complex customer-facing Professional Services engagements and serve as the accountable delivery owner for customer requirements, execution, and business outcomes. The role translates customer objectives into a clear delivery strategy, aligns execution to expected outcomes, and drives alignment across scope, schedule, risk, dependencies, stakeholder actions, governance, and success measures. The person acts as a consultative partner to the customer’s business. helping shape priorities, tradeoffs, and decisions while keeping delivery focused on value realization. They must bring strong services delivery leadership, executive communication skills, and the ability to operate effectively in technically complex environments with sound judgement. The role partners closely with Consulting Engineers, Customer Success, Sales, and customer stakeholders to maintain momentum, align resources, and deliver measurable value. It also operates at both the engagement and broader account level, including leading regular quarterly business reviews and helping teams reassess priorities when needed. This role will be based remotely in the United States. What you will do: Lead complex Professional Services engagements from kickoff through closeout, maintaining continuity, accountability, and focus on value realization Own delivery accountability for customer requirements, execution, and business outcomes, keeping work aligned to customer prior

mongodbawsazure
View job →

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Mission: Building the Data Foundations for AI We are the Snowflake Interoperable Foundations organization - the foundational layer that powers Snowflake’s AI, Analytics and Data Engineering capabilities. We lead innovations across open table formats such as Apache Iceberg, helping customers build peta-byte scale multi-cloud data lakes on Snowflake. We deliver core Metadata capabilities that power Snowflake’s industry-leading performance, AI, governance and platform features. We are embarking on a 0->1 redesign of our core systems across Interoperable Foundations. While we already manage exabyte-scale data supporting Snowflake’s AI capabilities, the next frontier is providing the foundational data layer that accelerates agentic innovation in an open, multi-format data world, You will be setting the technical vision across our investments in metadata platforms, Apache Iceberg and AI-ready storage. Your Impact: From Redesign to Reality 0->1 Architectural Leadership: Lead the ground-up redesign of our core Metadata systems, influencing the transaction frameworks that power query, DML, and AI-driven data interactions in addition to extending our lead on platform capabilities such as Zero Copy Cloning and Cross-Region / Cross-Cloud Replication. Iceberg Innovation: Drive

aigorust
View job →
C
9 days ago

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Senior Manager, Platform Engineering / DevOps Who Are You You are an experienced Senior Manager / emerging Staff-level leader in DevOps and Platform Engineering with strong technical depth and demonstrated leadership in delivering enterprise-scale cloud platforms. You bring a balanced mix of hands-on engineering expertise, team leadership, and execution rigor. You excel in driving outcomes in complex, multi-stakeholder environments, guiding teams to deliver secure, scalable, and high-quality platform solutions. You are comfortable leading engineers, managing stakeholders, and owning delivery across multiple workstreams. You demonstrate: A strong ownership mindset with accountability for delivery and outcomes Ability to translate business needs into actionable engineering roadmaps Solid expertise in cloud-native platforms, DevOps practices, and SRE principles Capability to lead teams and influence without requiring extensive tenure Role Responsibilities Development & Enforcement Own and execute the H100 platform engineering roadmap, aligned to enterprise priorities and program milestones Drive delivery of GCP-based platform capabilities (GKE, networking, IAM, CI/CD, observability) Establish and enforce engineering standards, best practices, and ADR compliance <li

sqlmongodbgcp
View job →

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Role Summary At CVS Health®, you’ll be working with a team of passionate colleagues who care deeply, innovate with purpose, hold themselves accountable and prioritize safety and quality in everything we do. Do you enjoy innovation while having fun doing it? If the answer yes, then this role might be for you! Join us and be part of something bigger, innovative and simplification in healthcare. We are seeking a highly experienced and innovative Principal (Director Level) Software Development Engineer to lead the application architecture, design, development, delivery of next-generation digital applications (including Reporting and financial solutions), and optimization of scalable, secure, and high-performance solutions leveraging AI across all major cloud platforms (AWS, Azure, and GCP). This role requires deep technical expertise, AI-enabled solutions, strategic thinking, scalable digital platforms, and enterprise integrations that power critical healthcare and pharmacy experiences and a passion for driving excellence in software engineering practices. This is a senior technical leadership role for a hands-on engineer who can operate across the full stack—from intuitive front-end applications to resilient backend services—while setting architectural direction, influencing engineering standards, and mentoring teams. The ideal candidate combines deep technical expertise, platform thin

typescriptpythonjava
View job →
N
12 days ago

We are now looking for a dynamic business leader to grow NVIDIA's Host Networking business for AI infrastructure with AI Labs and Hyperscalers! This leader will drive strategic direction, customer engagement, and multi-year growth for networking products such as NVIDIA DPUs, SuperNICs, and their associated software and ecosystem. Success in this role will be measured by the level of adoption and integration of our Host Networking products with our end customers' workflows and workloads. Success is contingent upon building trust with executives, architects, product leaders, and platform teams across NVIDIA and our largest customers. This leader will lead the go-to-market motion, connecting customer AI factory needs to NVIDIA's networking portfolio and aligning product, sales, engineering, architecture, marketing, and partner teams to secure design wins and scale deployments. What you'll be doing: Identify, develop and close strategic design wins for DPU and SuperNIC with top AI labs and Cloud Service Providers! Build and implement the segment sales growth strategy for host networking across hyperscaler and frontier model AI labs building large scale AI infrastructure. Define customer-specific DPU and SuperNIC value propositions and deployment motions, and lead a matrixed team across product, architects, engineering, sales and marketing teams. Promote NVIDIA host networking products externally and internally, positioning their value for AI workloads and other infrastructure products from NVIDIA, in a collection of use-cases in Networking, Security and Storage. Build a robust opportunity pipeline with segment sales and account teams, including account mapping, customer requirements, proof points, executive engagement, and partner alignment. Track and drive quarterly business reporting, forecast accuracy, design-win progress, roadmap asks, and

aiprocurement
View job →
RS
13 days ago

OUR MISSION At Redwood, we empower our customers with lights-out automation for their mission-critical business processes. ABOUT US Redwood Software is the leader in full stack automation fabric solutions for mission-critical business processes. With the first SaaS-based composable automation platform specifically built for ERP, we believe in the transformative power of automation. Our unparalleled solutions empower you to orchestrate, manage and monitor your workflows across any application, service or server — in the cloud or on premises — with confidence and control. CORE VALUES One Team. One Redwood Make Your Own Weather Obsess over Customer Success Work the Problem Be Curious Own the Outcome Respect Each Other YOUR IMPACT We are looking for an Engineering Manager, Products & Platforms to join our Product engineering team to lead a high-performing software engineering team while driving technical strategy across Redwood’s Workload Automation Platform. You will be instrumental in mentoring engineers, managing team deliverables, and guiding the design, development, and enhancement of scalable, secure software that powers enterprise data exchange for more than 1,000 customers worldwide. As an Engineering Manager, you will: People Leadership & Mentorship: Manage, coach, and grow a team of talented software engineers, supporting career development, conducting performance reviews, and fostering an inclusive, collaborative team culture. Technical Strategy & Architecture: Provide hands-on technical guidance, participate in design reviews, and define technical roadmaps for Java/Spring Boot microservices while ensuring high standards for architecture, security, and observability. Delivery & Platform Ownership: Oversee team execution, Sprint planning, and delivery timelines to ensure resilient, scalable core platform features and infrastructure. AI Integration: Research and apply AI/ML concepts and their usage to innovate and enhance the MFT (Managed

javaawskubernetes
View job →
I
13 days ago

Job Details: Job Description: The Role and Impact As a GPU Platform Hardware Design Engineer, you will play a pivotal role in designing and developing high-quality GPU hardware platforms that drive innovation in high-performance computing, graphics, and visualization technologies. You will lead the design process from initial feasibility studies through board layout, tapeout, and platform power-on, ensuring robust functionality and compatibility with industry standards. Your expertise in platform-level requirements, electrical engineering applications, and system bring-up will directly contribute to delivering cutting-edge GPU systems that accelerate Intel's leadership in computing. Business group The Data Center Group (DCG) is dedicated to advancing Intel's role in powering the digital world with leading-edge technologies. Focused on delivering innovative solutions for data center and cloud environments, DCG supports high-performance computing and graphics to enable capabilities such as AI, machine learning, and advanced visualizations. As part of the GPU IP Engineering team within DCG, you'll contribute to developing GPU systems that meet the evolving demands of the industry while supporting Intel's broader mission to create world-changing technology. Key Responsibilities - Design, develop, and evaluate electronic components, PCBs, and integrated circuits for GPU hardware platforms. - Translate platform-level requirements into detailed specifications and ensure adherence throughout the design process. - Define component placement and trace routing rules to optimize board layouts for performance, power, and signal integrity. - Conduct feasibility studies, board layout, tapeout, and platform power-on activities. - Perform functionality tests and utilize tools to verify platform configurations and compatibility. - Research, develop, and validate firmware, hardwa

machine learningairecruitment
View job →
EI
Eltropy Inc.
📍 India• Full-time
17 days ago

Role: Engineering Manager – Communications Location: Remote Team size: 15 engineers (backend + frontend) Type: Full-time About the Role We’re looking for an Engineering Manager who thrives at the crossroads of leadership, hands-on engineering, and solving problems that don’t come with an instruction manual. You’ll lead a team of talented backend and frontend engineers who are building the bridges between systems that power the full customer journey for financial institutions. In this role, you won’t just be overseeing work, you’ll be rolling up your sleeves, writing and reviewing production code, guiding architecture decisions, and coaching engineers to deliver their best work. The systems you’ll help build will connect modern cloud APIs with decades-old banking platforms, bringing reliability and elegance to what often starts as messy complexity. Our integration platform touches everything—from communication stacks to core banking systems, lending and mortgage platforms, payment gateways, and AI Platform. Every integration is an opportunity to shape how our customers experience our products end-to-end. And because we take an AI-augmented approach to software development, you’ll be part of a team that uses AI tools to augment SDLC and write better code, automate testing, and ship faster without compromising quality. What You’ll Be Doing Lead by Example – Stay hands-on with coding, designing architectures and reviewing code while guiding the team toward engineering excellence. Own the Integration Layer – Architect and scale connections across diverse systems, from sleek modern APIs to finicky legacy protocols. Champion the Customer Experience – Partner with Product, Implementation, and customer teams to ensure integrations truly solve real-world challenges. Collaborate Without Boundaries – Work closely with other engineering leaders to ship features that feel seamless across products. Build for the Long Run – Keep systems observable, reliable, and perform

javascriptjavareact
View job →
E
17 days ago

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As a Software Engineer on the DX Security Components team , you will architect the backend services that safeguard Pure’s core authentication and remote access infrastructure. You’ll act as a security-minded engineer within a high-impact platform team, ensuring our cloud offerings and customer appliances remain auditable and resilient. Collaborating across the DX organization, you will bridge the gap between robust security protocols and seamless developer integration to protect our global customer base. WHAT YOU’LL DO Engineer Security Infrastructure: Design and operate high-availability backend services that manage authentication, authorization, and certificate lifecycles to ensure secure access across all Pure1 cloud and appliance environments. Drive End-to-End Ownership: Lead the full service lifecycle—from initial architectural design and threat modeling (STRIDE) to deployment, observability, and long-term cost efficiency. Champion Secure Integration: Partner with Security Governance and product teams to streamline remote-access flows, translating complex security requirements into pragmatic, automated workflows for other engineering squads. Ensure System Resilience: Maintain the integrity of security-sensitive systems by participating in a global follow-the-sun on-call rotation, performing root-cause analysis, and hardening infrastructure against emerging threats. Automate Trust: Evolve Infrastructure as Cod

pythonawsci/cd
View job →
E
Everpure
📍 Bengaluru• Full-time
17 days ago

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Join the Exa team and lead the charge in redefining enterprise storage by unifying block, file, and object protocols across hybrid-cloud environments. You will combine deep technical expertise in distributed systems with hands-on people leadership to guide architectural decisions and mentor high-impact engineers. This is a unique opportunity to build new engineering teams from the ground up and drive industry-leading innovation alongside Product and Architecture partners. Your work will directly impact how customers consume, scale, and operate mission-critical storage infrastructure. WHAT YOU'LL DO Drive End-to-End System Architecture: Lead the architectural evolution and end-to-end delivery of high-performance, resilient storage systems from initial design concepts to high-quality shipped products. Optimize for Modern Data Workloads: Design and implement robust algorithms and concurrent platform solutions engineered for modern data pipelines, AI infrastructure, distributed computing, and enterprise analytics. Resolve Complex System Engineering Challenges: Apply deep root-cause analysis and system-level insight to solve multi-threaded, high-concurrency performance and reliability issues across Linux platform internals. Cross-Functional Ownership & Leadership: Collaborate across product management, validation, and support teams to align technical roadmaps, establish architectural standards, and drive ent

pythonjavaaws
View job →
E
Everpure
📍 Bengaluru• Full-time
17 days ago

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Join the Exa team and lead the charge in redefining enterprise storage by unifying block, file, and object protocols across hybrid-cloud environments. You will combine deep technical expertise in distributed systems with hands-on people leadership to guide architectural decisions and mentor high-impact engineers. This is a unique opportunity to build new engineering teams from the ground up and drive industry-leading innovation alongside Product and Architecture partners. Your work will directly impact how customers consume, scale, and operate mission-critical storage infrastructure. WHAT YOU'LL DO Drive End-to-End System Architecture: Lead the architectural evolution and end-to-end delivery of high-performance, resilient storage systems from initial design concepts to high-quality shipped products. Optimize for Modern Data Workloads: Design and implement robust algorithms and concurrent platform solutions engineered for modern data pipelines, AI infrastructure, distributed computing, and enterprise analytics. Resolve Complex System Engineering Challenges: Apply deep root-cause analysis and system-level insight to solve multi-threaded, high-concurrency performance and reliability issues across Linux platform internals. Cross-Functional Ownership & Leadership: Collaborate across product management, validation, and support teams to align technical roadmaps, establish architectural standards, and drive enterprise

pythonjavaaws
View job →
🔔

Get new lead cloud operations engineer jobs by email

Daily job updates · Unsubscribe anytime