Jobiba hiring network

Lead Cloud Operations Engineer Jobs

6,876 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current lead cloud operations engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

G
Godaddy
📍 United States• Full-time• From $154K/yr
1mo ago

Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time, others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. This position is not eligible to be performed in Alaska, Mississippi, North Dakota, or the Virgin Islands. GoDaddy is not currently considering candidates for this role in California, Seattle, or NYC. Join Our Team... We are seeking a highly skilled Senior Security Engineer to join our advanced Security Operations team. This role is focused on leading complex incident response and forensic investigations across Windows, macOS, Linux, and AWS environments while helping modernize security operations through automation and AI-driven capabilities. The ideal candidate is a hands-on security expert with deep AWS security expertise, strong threat detection and digital forensic skills, and experience leveraging AI and machine learning technologies to improve detection, response, and operational efficiency. You will play a key role in protecting critical assets, conducting high-impact investigations, mentoring team members, and driving the evolution of our security program against sophisticated and emerging threats. What You'll Get to Do... Lead high-priority incident response and forensic investigations, serving as the primary escalation point for advanced analysis, containment, recovery, root cause determination, and executive-level reporting. Drive threat detection and response across AWS, Windows, macOS, Linux, and endpoint security platforms, leveraging services such as GuardDuty, Security Hub, Detective, CloudTrail, IAM, VPC Flow Logs, and SentinelOne. Conduct malware analysis, host and cloud forensics, evidence collection, and threat hunting activi

pythonawsgit
View job →

Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Security Operations (SecOps) team protects Robinhood and its customers by detecting, investigating, and responding to security threats across our production systems, endpoints, and cloud environments. We combine detection engineering, incident response, and threat intelligence to identify emerging risks, strengthen our defenses, and reduce the impact of attacks before they reach our customers. As we continue to evolve our security platform, we're embracing AI and automation to help our teams move faster, uncover threats more effectively, and scale our defenses against increasingly sophisticated adversaries. As the Manager of Response, Automation, Intelligence and Detection Engineering (RAID) within SecOps, you will lead teams across North America, shaping the strategy, execution, and long term evolution of our defensive security capabilities. You'll be responsible for building high performing teams, maturing our detection and incident response programs, and ensuring we stay ahead of an increasingly sophisticated threat landscape. Working closely with Security, Engineering, Infrastructure, and Trust & Safety, you'll translate emerging threats into scalable defenses wh

vueawsai
View job →
O
OpenAI
📍 London• Full-time
1mo ago

About the Team OpenAI’s Network Security team designs and operates the secure, reliable connectivity behind our offices, labs, campuses, cloud environments, people, and devices. We combine strong network fundamentals with automation, observability, and close partnership across IT, Security, Research, Applied, and business teams. About the Role As a Network Engineer, you will design, operate, troubleshoot, and automate secure, reliable networks across offices, labs, cloud connectivity, and production services. You will balance strategic platform work—architecture, standards, roadmaps, lifecycle planning, and automation—with responsive operations such as incidents, escalations, break/fix, and time-sensitive delivery. We’re looking for broad network engineers who meet users where they are, lead with curiosity, own outcomes end-to-end, move with urgency grounded in security, and iterate with purpose. You will turn operational signals and recurring reactive work into durable systems and standards. In this role, you will: End-to-end ownership of secure enterprise routing, switching, wireless, WAN, network services, and cloud connectivity. A deliberate balance of strategic platform improvement and responsive troubleshooting, change safety, incident response, and operational delivery. Purposeful iteration through software, APIs, Infrastructure-as-Code, Git workflows, testing, and CI/CD that reduces recurring reactive work. You might thrive in this role if you have: End-to-end ownership of secure enterprise routing, switching, wireless, WAN, network services, and cloud connectivity. A deliberate balance of strategic platform improvement and responsive troubleshooting, change safety, incident response, and operational delivery. Purposeful iteration through software, APIs, Infrastructure-as-Code, Git workflows, testing, and CI/CD that reduces recurring reactive work. Compensation, Benefits and Perks This is a position with OpenAI UK Ltd., which controls the hiring and manageme

reactawsci/cd
View job →
O
1mo ago

About the Team Security is at the foundation of OpenAI's mission to ensure that artificial general intelligence benefits all of humanity. The Identity Infrastructure Engineering team sits at the core of this effort, designing and building the identity and access management solutions that protect model weights, customer data, and critical systems across multiple cloud environments. The team partners across OpenAI, including Applied Engineering, Research, IT, Security, Infrastructure, and Engineering, to provide secure and scalable platforms for identity, access management, permissioning, orchestration, and safe AI research. About the Role We’re looking for an engineering leader to lead Identity Infrastructure Engineering, the team building the systems that govern and scale access across OpenAI’s research, engineering, and internal platforms. This role sits at the center of cloud infrastructure, identity, software engineering, and security-critical operations. You’ll lead engineers building control planes, policy systems, workload and agent authorization patterns, infrastructure-as-code, and operational foundations that help OpenAI move quickly while keeping access reliable, auditable, least-privileged, and safe under failure. The ideal candidate has led teams responsible for large-scale, mission-critical infrastructure. They can go deep into code and architecture when needed, while giving engineers and technical leads the clarity and ownership to do their best work. They set technical direction, grow strong teams, make durable architecture decisions, and turn ambiguous 0-to-1 problems into platforms OpenAI can trust and build on for years. In this role, you will: Build and lead a high-performing Identity Infrastructure team, going deep enough technically to set direction while empowering the team to own delivery. Define the strategy for identity platform as the policy plane for access across people, agents, workloads, services, clouds, and internal systems. Scale Acc

awsgitrest
View job →
PE
1mo ago

Position: Engineering Manager - Database Job Location: Noida Role Overview We are seeking a Database Engineering Manager (Individual Contributor) with deep expertise in MySQL and strong working knowledge of MongoDB, PostgreSQL, and Cassandra. This role combines hands-on database administration and optimization with strategic ownership of database reliability, automation, and cloud adoption. The candidate will lead by example—driving technical excellence, influencing best practices, and partnering cross-functionally with DevOps, SRE, and product engineering teams to deliver highly available, secure, and scalable database platforms. Key Responsibilities 1. End-to-End Ownership of MySQL databases in production & staging—availability, performance, and reliability. 2. Architect, manage, and support MongoDB, PostgreSQL, and Cassandra clusters for scale and resilience. 3. Define and enforce backup, recovery, HA, and DR strategies across all critical database platforms. 4. Drive database performance engineering—tuning queries, optimizing schemas, indexing, and partitioning for high-volume workloads. 5. Own replication, clustering, and failover architectures ensuring business continuity. 6. Champion automation & AI-driven operations—design self-healing scripts, predictive scaling, and proactive monitoring solutions. Collaborate with Cloud/DevOps teams on AWS database services (RDS, Aurora, DynamoDB, EC2, S3) to optimize cost, security, and performance. 7. Establish monitoring dashboards & alerting mechanisms for slow queries, replication lag, deadlocks, and capacity planning. Ensure compliance & security standards—encryption, auditing, and regulatory requirements. 8. Lead incident management & on-call rotations, ensuring rapid response and minimal MTTR. 9. Act as a strategic technical partner, contributing to database roadmaps, automation strategy, and adoption of AI-driven DBA practices. Required Skills & Experience 1. 6–10 years of p

sqlpostgresqlmysql
View job →

Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. As part of Micron's Technology Engineering & Innovation (TE&I) organization, you will have the opportunity to shape the future of our global infrastructure platforms while enabling business growth, operational resilience, and digital transformation at scale. The Opportunity Micron is seeking a transformational Senior Director of Technology Engineering & Innovation (TE&I) to lead the strategy, engineering, operations, and modernization of our global infrastructure ecosystem. This role is responsible for defining and executing Micron's vision across enterprise networks, cloud platforms, data center strategy, database services, infrastructure engineering, automation, observability, and global infrastructure operations. As a key member of the TE&I leadership team, you will drive innovation, operational excellence, and strategic transformation while building a high-performing organization focused on business outcomes and exceptional customer experiences. The successful candidate will be equally comfortable developing multi-year technology strategies, leading large-scale infrastructure transformations, driving operational performance, developing talent, and fostering a culture of collaboration, accountability, and continuous improvement. What You Will Do Lead Global Infrastructure Strategy & Transformation Define and execute Micron's global infrastructure vision and strategy. Develop multi-year roadmaps for: Enterprise Network Services <l

airecruitment
View job →
E
12 days ago

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE The Everpure Engineering team builds the industry’s most innovative, high-performance, and highly available portfolio of storage products designed for the most demanding mission-critical applications. While we deliver hardware storage arrays, over 90% of our engineering staff are software engineers. Our customers are the core of our business, and they love FlashArray for its simplicity, continuous stream of upgrades, and ability to stay on the cutting edge without ever experiencing downtime. These innovations enable our customers to leverage the agility of the public cloud across both traditional IT and cloud-native applications. WHAT YOU'LL DO Hardware Design & Leadership: Plan, lead, implement, and debug hardware design and qualification projects from concept through production. Full Lifecycle Participation: Participate in all phases of product development, including architecture design, technical documentation, eDVT, failure analysis (FA), and custom hardware support. Validation & Automation: Define, execute, and automate hardware validation and qualification frameworks to enhance test coverage, efficiency, consistency, and product robustness. Cross-Functional Collaboration: Partner with Software, Operations, Manufacturing Test, and Support teams to seamlessly introduce new storage subsystems into production. Escalations & Root Cause Analysis: Serve as an escalation point for

pythonaisupply chain
View job →
W
WPP
📍 Chennai• Full-time
15 days ago

WPP is the trusted growth partner for the world’s leading brands. We unite cutting-edge media intelligence and data solutions, world-class creativity, next-generation production, transformative enterprise solutions and expert strategic counsel in a single company – powered by exceptional talent and our agentic marketing platform, WPP Open, to help our clients navigate change, capture opportunity and deliver transformational growth. We work with the world's most valuable brands and have global reach across 100+ markets, with deep local expertise. Our people are the key to our success. We're committed to fostering a culture of creativity, belonging and continuous learning, attracting and developing the brightest talent, and providing exciting career opportunities that help our people grow. For more information, visit WPP.com. Why we're hiring: The Automation Engineer is responsible for designing, developing, and maintaining security automation solutions that enhance detection, response, workflow efficiency, and operational consistency across Operational Security. Working under the Automation Lead, this role builds high-quality SOAR playbooks, integrations, scripts, AI-assisted workflows, and orchestration pipelines to reduce manual workloads and support the Autonomic Security Operations (ASO) model. What you'll be doing: Core Responsibilities Automation Engineering & Development Develop SOAR playbooks, workflows, and automations for alert triage, enrichment, containment, and remediation. Build scalable, reusable automation components, scripts, and integrations. Implement high-quality scripting using Python, PowerShell, and REST APIs. Ensure appropriate version control, QA, testing, and documentation of automation artefacts. Maintain reliability of automations by monitoring performance, exceptions, and system behaviour. Platform Integration & Tooling Engineering Integrate SOAR with SIEM, EDR, TIP, cloud-native secur

pythonazuregcp
View job →
W
WPP
📍 Lisbon• Full-time
15 days ago

WPP is the trusted growth partner for the world’s leading brands. We unite cutting-edge media intelligence and data solutions, world-class creativity, next-generation production, transformative enterprise solutions and expert strategic counsel in a single company – powered by exceptional talent and our agentic marketing platform, WPP Open, to help our clients navigate change, capture opportunity and deliver transformational growth. We work with the world's most valuable brands and have global reach across 100+ markets, with deep local expertise. Our people are the key to our success. We're committed to fostering a culture of creativity, belonging and continuous learning, attracting and developing the brightest talent, and providing exciting career opportunities that help our people grow. For more information, visit WPP.com. Why we're hiring: At WPP, technology is at the heart of everything we do, and it is the Technology Operations teams mission, as part of Enterprise Technology , to enable our stakeholders to collaborate, create and thrive. Enterprise Technology is undergoing a significant transformation to modernise ways of working, shift to cloud and micro-service-based architectures, drive automation, digitise colleague and client experiences and deliver insight from WPP’s petabytes of data. This role will carry out the effective and efficient everyday technology operations for WPP ET. A trusted pair of hands to deal with level 1 and 2 issues as they present to the IT Service Desk and a trusted resource for Infrastructure and Management personnel to assist with project work when needed. The role will report into the Enterprise Technology Operations Lead and work closely with other teams within Enterprise Technology. What you'll be doing: Deliver world class, on-site support services to WPP employees, agencies, and visiting clients, operating within predefined structur

PE
1mo ago

About Us: Paytm is India's leading mobile payments and financial services distribution company. Pioneer of the mobile QR payments revolution in India, Paytm builds technologies that help small businesses with payments and commerce. Paytm’s mission is to serve half a billion Indians and bring them to the mainstream economy with the help of technology. About the team: "Thriving on innovation, One97 delivers Customer Communication, Up-Selling, Mobile Content, Advertising based and Platform services to over 500 Million mobile consumers across the globe. Headquartered in New Delhi, and backed by marquee investors like Intel Capital, SAIF Partners, Silicon Valley Bank, SAP Ventures & Berkshire Hathaway, One97’s products and services are available on various models i.e. Capex, Opex & Revenue Share model as per the flexibility of Telco’s. Our Key Offerings are divided into 5 broad categories as follows: • Entertainment• Digital Platforms• CVM Solutions• Enterprise Services• Financial Platforms One97 has the widest and largest deployment of telecom applications on cloud platforms in India and has a myriad of VAS services that have helped operators augment their revenue even in complex markets like India, SAARC, Middle East, Africa and many more. About the role: Responsible for guiding teams towards continuous integration and continuous deployment; hence, he must have extensive knowledge of the automation tools and their application like Gradle, Git, Jenkins, Bamboo, Docker, Kubernetes, Puppet Enterprise, Nagios, Chef, and Ansible. The DevOps tackles our organization’s toughest technical problems and drives technical excellence at all levels, working with senior management to support the execution of the organization’s vision. The DevOps / Sr. Devops/ Lead Devops is a person who cares deeply about the technical side of operations and makes sure that everything is running smoothly for our applications. The ideal candidate understands technology deeply at both the server

pythonawsazure
View job →

About the Team OpenAI's Industrial Compute organization is building the world's most advanced AI infrastructure ecosystem. In partnership with leading cloud providers, hardware manufacturers, utilities, construction partners, and internal engineering organizations, we are delivering hyperscale AI campuses that power the next generation of frontier AI models. Infrastructure Delivery Operations sits at the center of this effort. Our team develops the operating model that connects infrastructure strategy, supply planning, manufacturing operations, and delivery into a single, integrated system that enables OpenAI to deploy AI infrastructure predictably at scale. We partner across Hardware Engineering, Network Engineering, Capacity Delivery, Hardware Operations, Security, Finance, Strategic Sourcing, and external infrastructure partners to create a single, integrated view of program health. Through governance, operational analytics, executive reporting, and scalable operating mechanisms, we enable leaders to proactively manage risk, optimize capacity, and deliver infrastructure predictably at Industrial Compute speed. About the Role We are seeking a Technical Program Manager, Infrastructure Delivery Operations to drive integrated strategy and delivery across OpenAI's rapidly expanding AI infrastructure portfolio. This role sits at the intersection of infrastructure strategy, New Product Introduction (NPI), supply planning, manufacturing operations, and infrastructure delivery. You will lead highly cross-functional programs spanning engineering, supply planning, manufacturing, logistics, construction, commissioning, and operations, ensuring technical and operational dependencies remain synchronized from planning through production readiness. Beyond driving program execution, you will leverage operational insights to improve capacity planning, infrastructure strategy, and deployment readiness. You will also help operationalize new technologies and suppliers by partnering w

REMOTEawsrestagile
View job →
V
Vanta
📍 United States• Full-time
1mo ago

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. We are seeking an experienced Manager of Security Operations, reporting to the Director of Security, to lead our exceptionally talented Security Engineering team. Vanta’s Security Operations team provides essential security operational services, including configuring, monitoring, and maintaining our security tools and infrastructure. You’ll be responsible for leading the Security Operations team and assisting with incident response, setting our detection and response strategy, and assisting with investigations. You’ll also work cross-functionally to ensure we maintain compliance and improve our secure maturity. What you’ll do as a Manager, Security Operations at Vanta: Lead and grow a team of the best security operations analysts in the world, with a view of security that is AI-first, human-centric, and trust-based. Help define the strategy for Vanta’s security operations team, and empower the team to implement robust security protocols and stay ahead of emerging threats. Leverage AI to improve efficiency of team processes, and improve the maturity of the overall security program. Lead and drive incident response from detection, remediation, to prevention Identify, develop, and implement new processes in our security operations program Identify new technologies and emerging threats for our organization, and plan a technology-driven approach to addressing new risks How to be successful in this role: Strong leadership experience in security and an ability to lead a global team from a foundation of transparency and trust. Strong security operations experience, with emphasis on implementing security controls in a SaaS and cloud env

restaigo
View job →
R
Roblox
📍 San Mateo• Full-time• From $345K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Principal Software Engineer leading Fleet Management, you will be the overall technical lead across three pods and the person who sets the technical direction for the fleet management layer of Roblox. This is a hands-on, deeply technical leadership role that owns all of Roblox's compute capacity end to end: from low-level provisioning and the data plane, up through the control planes that operate it, and all the way to the UI and internal-facing products that let teams self-serve capacity. Your org centralizes security, maintenance operations, and the uptime of every Roblox Kubernetes cluster, and governs the internal customer contracts that drive automation across the fleet spanning Roblox data centers and cloud providers. You will guide architecture, raise the engineering bar, and make sure compute capacity supply and demand stay in balance as the fleet grows. You will: Serve as the overall technical lead for three Fleet Management pods, setting and aligning the technical direction across low-level provisioning, the data plane, and the control plane and product surfaces above them. Architect the declarative, Kubernetes-style control planes that operate Roblox's compute fleet across o

sqlawskubernetes
View job →
P
Pinterest
📍 San Francisco• Remote
8 days ago

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . Pinterest is seeking a Sr. Manager to lead our Capacity Engineering team. The team ensures that Pinterest’s cloud infrastructure has the capacity it needs while operating reliably, efficiently and with clear financial accountability. You’ll lead the full portfolio across forecasting and supply, capacity-management systems, compute and GPU efficiency, infrastructure data and governance and capacity operations. What you’ll do: Lead the Capacity Engineering team and establish its 12–18 month functional and technical strategy, roadmap and success measures tied to Infrastructure and company goals. Develop CPU and GPU forecasts and supply plans that account for workload demand, delivery constraints, cost and reliability requirements. Guide the design and delivery of capacity requests, reservations, entitlements, allocation policy and infra

REMOTEkubernetesaifinance
View job →
W
11 days ago

WPP is the trusted growth partner for the world’s leading brands. We unite cutting-edge media intelligence and data solutions, world-class creativity, next-generation production, transformative enterprise solutions and expert strategic counsel in a single company – powered by exceptional talent and our agentic marketing platform, WPP Open, to help our clients navigate change, capture opportunity and deliver transformational growth. We work with the world's most valuable brands and have global reach across 100+ markets, with deep local expertise. Our people are the key to our success. We're committed to fostering a culture of creativity, belonging and continuous learning, attracting and developing the brightest talent, and providing exciting career opportunities that help our people grow. For more information, visit WPP.com. Why we're hiring: Responsible for leading the Cloud Automation Engineering function. Primary focus will be leading a team of other engineers in designing and implementing automation solutions to improve customer experience and increase productivity in our cloud estates. Responsible for maintaining and delivering automation solutions through infrastructure as code, ensuring security best practice, evangelising automation practice and tools, and supporting customer needs, both internal and external. What you'll be doing: Identify opportunities for improvement and automation of operations Design, build, test and implement use cases to drive automation adoption and improve operational efficiency Work closely with the IT Operations team to develop automated incident detection and response mechanisms. Implement proactive monitoring and alerting systems to quickly respond to and resolve critical issues, minimizing downtime and service disruptions Responsible for driving CSI initiatives to improve operations (processes/tools) working with various stakeholders Responsible for providing feedback at various leve

pythonawsazure
View job →
🔔

Get new lead cloud operations engineer jobs by email

Daily job updates · Unsubscribe anytime