Jobiba hiring network

Cloud Operations Engineer Jobs

2,288 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current cloud operations engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

M
1mo ago

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We are looking for strong engineers with experience and interest in designing, building, and maintaining the novel, high-performance systems that make up our serverless platform. Requirements: 5+ years of experience writing high-quality production code Experience building high-performance distributed systems at a large scale (the more battle scars, the better) Strong cloud skills Strong knowledge of low-level operating system foundations (Linux kernel, file systems, containers, etc.) Experience with performance engineering (tell us a story of when you shaved off a few milliseconds!) Ability to work in-person in our NYC or SF office. Prior experience with Rust is nice to have, but not required. Ability to participate in on-call rotation and respond to production incidents.

linuxrestai
View job →
M
1mo ago

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We are looking for strong engineers with experience in making ML systems performant at scale. If you are interested in contributing to open-source projects and Modal’s container runtime to push language and diffusion models towards higher throughput and lower latency, we’d love to hear from you! Requirements: 5+ years of experience writing high-quality, high-performance code. Experience working with torch, high-level ML frameworks, and inference engines (vLLM or TensorRT). Familiarity with Nvidia GPU architecture and CUDA. Experience with ML performance engineering (tell us a story about boosting GPU performance — debugging SM occupancy issues, rewriting an algorithm to be compute-bound, eliminating host overhead, etc). Nice-to-have: familiarity with low-level operating system foundations (Linux kernel, file systems, containers, etc).

linuxrestai
View job →
M
Modal
📍 Stockholm• Full-time
1mo ago

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We are looking for strong engineers with experience and interest in designing, building, and maintaining the novel, high-performance systems that make up our serverless platform. Requirements: 5+ years of experience writing high-quality production code Experience building high-performance distributed systems at a large scale (the more battle scars, the better) Strong cloud skills Strong knowledge of low-level operating system foundations (Linux kernel, file systems, containers, etc.) Experience with performance engineering (tell us a story of when you shaved off a few milliseconds!) Ability to work in-person in our Stockholm office. Prior experience with Rust is nice to have, but not required. Ability to participate in on-call rotation and respond to production incidents.

linuxrestai
View job →
E
Everpure
📍 Bengaluru• Full-time
17 days ago

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. Portworx by Everpure(Formerly Pure Storage) Everpure Acquired Portworx in October 2020, Creating the Industry's Most Complete Kubernetes Data Services Platform for Cloud Native Applications. This acquisition represents Pure’s largest to date and our deeper expansion into the fast-growing market for multi-cloud data services to support Kubernetes and containers. WHAT YOU'LL DO Lead the design, development, and deployment of Portworx's Kubernetes-native backup and restore solution, serving as a technical anchor for the team. Design and build cloud-native services that operate reliably at scale, with a strong emphasis on performance, quality, and efficiency in both design and implementation. Deeply leverage Kubernetes constructs (operators, CRDs, controllers, CSI) to build robust, production-grade data protection capabilities. Ensure security, resilience, stability, and high availability are first-class considerations in every design and implementation decision. Design and develop robust, well-defined APIs that integrate seamlessly with the frontend and Web UI. Drive architectural decisions and technical direction for the product, mentor engineers, and set the bar for engineering quality through design reviews and code reviews. Own end-to-end delivery of complex features — from design documents through implementation, testing at scale, and production readiness. We are primarily an in-office environment and therefore, you will

pythonsqlmysql
View job →
AC
17 days ago

Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. You will own end-to-end delivery for complex, cross-org modernization programs. You will create clarity from ambiguity, define operating mechanisms, drive dependency management, and ensure high-quality execution across Platform Engineering and Low Code Foundations. Required Experience & Skills 8–12+ years technical program management in engineering organizations Experience driving large, cross-team technical programs with measurable outcomes Excellent written and verbal communication; strong ability to influence without authority Demonstrated rigor in planning, risk management, and execution mechanisms Comfort partnering with senior engineering leaders and principal engineers Preferred Experience & Skills Experience with platform modernization (cloud-native, performance, reliability, migrations) Experience with AI-enabled programs or AI-first SDLC adoption Background in software engineering, systems engineering, or technical product delivery Tools and Resources Training and Development: During onboarding, we focus on equipping new hires with the skills and knowledge for success through department-specific training. Continuous learning is a central focus at Appian, with dedicated mentorship and the First-Friend program being widely utilized resources for new hires. Growth Opportunities: Appian provides a diverse array of growth and development opportunities, including our leadership program tailored for new and aspiring managers, a comprehensive library of specialized department training through Appian University, skills based training,

P
Pinterest
📍 San Francisco• Full-time• Remote• From $145.7K/yr
1mo ago

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . The Team: Pinterest Technical Program Managers are proactive owners with technical expertise. They are aligned to the highest level company priorities to support successful delivery of programs within their teams. The Platform T/PgM Team is responsible for overall program governance in Infrastructure, Infra Finance, Data Engineering and Security, as well as Compliance and managing our cloud budget. What you’ll do: As a Staff Technical Program Manager focusing on cross-engineering special projects, you will identify, scope, champion, and land key strategic projects important to advancing Pinterest’s platform and underpinning. Translate Strategy into Execution — Convert strategic platform priorities into structured, milestone-driven execution plans; operating at both breadth (big picture vision) and depth (dr

REMOTEawsrestai
View job →
S
1mo ago

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is an AI-powered observability platform built on the Snowflake AI Data Cloud and engineered for scale. We ingest and store logs, metrics, traces, and events on an open, scalable data lakehouse, using open formats like Apache Iceberg, at dramatically lower cost. A dynamic Context Graph and chat-based AI SRE provide rich context and automated workflows so teams can move from detection to root cause of production issue and resolution 10x faster. Leading engineering teams at companies like Capital One, Topgolf, and Dialpad rely on Observe to troubleshoot hundreds of terabytes of telemetry daily while maintaining reliability at enterprise scale. As part of Snowflake, Observe combines startup-style ownership and velocity with the global reach, operational excellence, and ecosystem of one of the world’s leading data platforms. The Team On the AI Backend team at Observe by Snowflake you'll be at the forefront of how AI is reshaping the way engineering teams operate; building the platform that powers intelligent, automated workflows across observability and beyond. Our team is small, and moves fast, with real ownership over hard problems that span agentic APIs, real-time pipelines, and AI quality. You'll work alongside talented engineers across ML, product, and

typescriptpythonai
View job →
AC
17 days ago

Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. At Appian, we’re passionate about technology—we love making it, and we love using it. Joining Appian Engineering offers you the opportunity to learn in an environment that prizes cross-functional collaboration and is deeply committed to your professional growth. We aim to revolutionize the way people work by building a platform so intuitive that our customers can truly thrive. In the Senior Manager of Software Development role, you will provide technical leadership for a specific product area while directly driving the professional development of the engineers on your team. Key Responsibilities: The ideal candidate will manage development teams, oversee architecture decisions, and ensure scalable, high-quality delivery aligned with business objectives. This leader is expected to guide software engineers through solving ambiguous and complex problems. They will champion an AI-first approach and bring deep experience in migrating and modernizing legacy applications using cloud-native, modular architectures and emerging technologies. We are looking for someone with an established background in delivering highly scalable systems and a strong operational track record while working across multiple teams. This role provides the autonomy and ownership needed to innovate and solve the industry's most difficult problems. Leadership & Team Management Passion for mentoring and developing engineering talent: Lead and mentor a team of developers, architects, and analysts. Set very high standards for the team and in hiring talent. Manage end-to-end

pythonjavaaws
View job →
G
Godaddy
📍 United States• Full-time• From $128K/yr
1mo ago

Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This position may be a hybrid or fully remote position, as decided by your manager. If designated as hybrid, you’ll divide your time between working remotely from your home and an office location, so you should live within commuting distance. If designated as remote, you’ll be working remotely from your home and may occasionally visit a GoDaddy office to meet with your team for events or meetings. Your hiring manager can share more about this role’s hybrid or remote designation. This position is not eligible to be performed in Alaska, Mississippi, North Dakota, or the Virgin Islands. GoDaddy is not currently considering candidates for this role in California, Seattle, or NYC. Join Our Team Join a team powering secure, scalable email services for millions of customers worldwide! As part of GoDaddy's Professional Email team, you'll solve complex challenges in distributed systems, cloud infrastructure, security, and AI while modernizing critical platforms that businesses rely on every day. If you enjoy owning impactful systems, working across a diverse technology stack, and building innovative solutions at scale, you'll feel right at home here. What you'll get to do... Design, build, and maintain highly available, scalable APIs and services used by millions of customers Deploy, manage, and optimize cloud infrastructure in AWS Architect and implement modern solutions that improve performance, reliability, and security Leverage AI technologies to enhance development workflows and create innovative customer experiences Monitor, troubleshoot, and resolve complex production issues using modern observability and monitoring tools Drive continuous improvement through automation, modernization, and operational excellence C

pythonawsci/cd
View job →
CI
Couchbase, Inc.
📍 Bengaluru• Full-time
17 days ago

Couchbase, the operational data platform for AI, empowers businesses to succeed by bringing data to life in new ways. Major market-leading companies rely on Couchbase for mission critical operational, analytical, mobile and AI workloads. Built to replace legacy infrastructure and fragmented data services, Couchbase empowers enterprises with a unified platform architected for performance, flexibility and global scale. With Couchbase, organizations bring their data to life, launching game‑changing customer experiences, exploring the limitless potential of AI, and seamlessly extending applications from the cloud to the edge and beyond. Couchbase’s AI‑ready technology and enterprise partnership model eliminate complexity and reduce total cost of ownership, enabling teams to stay agile, innovative and secure. Couchbase believes data should never slow you down, but act as the foundation for your next breakthrough. Discover why Couchbase is trusted to help the world’s biggest players scale, move fast and stay resilient, no matter what’s next on their roadmap. Visit couchbase.com and follow us on LinkedIn and X. Want to be part of our story? Apply today! AI Platform Engineering Location: Bangalore (Hybrid - in office at least 3 days/week) About the Role We are seeking an experienced and visionary technology leader to lead the development and scaling of our Operational AI platform capabilities. This role will own the strategy, architecture, delivery, and operational excellence of Couchbase AI Cloud Platform capabilities. This is a strategic leadership role at the intersection of distributed systems, cloud-native platforms, and AI. You will lead a large, multi-layered engineering organization responsible for delivering The Operational Data Platform for AI, while partnering closely with Product, Design, SRE, and Go-To-Market teams. Your leadership will directly influence company growth, customer adoption, platform reliability, and Couchbase’s competitive pos

pythonjavaaws
View job →
CI
10 days ago

Couchbase, the operational data platform for AI, empowers businesses to succeed by bringing data to life in new ways. Major market-leading companies rely on Couchbase for mission critical operational, analytical, mobile and AI workloads. Built to replace legacy infrastructure and fragmented data services, Couchbase empowers enterprises with a unified platform architected for performance, flexibility and global scale. With Couchbase, organizations bring their data to life, launching game‑changing customer experiences, exploring the limitless potential of AI, and seamlessly extending applications from the cloud to the edge and beyond. Couchbase’s AI‑ready technology and enterprise partnership model eliminate complexity and reduce total cost of ownership, enabling teams to stay agile, innovative and secure. Couchbase believes data should never slow you down, but act as the foundation for your next breakthrough. Discover why Couchbase is trusted to help the world’s biggest players scale, move fast and stay resilient, no matter what’s next on their roadmap. Visit couchbase.com and follow us on LinkedIn and X. Want to be part of our story? Apply today! As a Solutions Architect , you will be the trusted advisor helping our customers unlock the ultimate value of Couchbase. You will work directly with customer engineering teams, blending high-level architectural coaching with hands-on professional service delivery. If you love solving complex, real-world data challenges, designing cutting-edge AI-ready architectures, and guiding enterprise customers from design to production, this role is for you. What You’ll Do (Responsibilities): Lead Customer Engagements: Guide short- to medium-term on-site and remote client engagements. You will lead architecture and use case reviews, sizing, performance tuning, and deployment topology planning to get Couchbase up, running, and optimized. Solve Complex Technical Challenges: Work hand-in-hand with customers t

sqlawsazure
View job →
M
Mongodb
📍 Gurugram• Full-time
1mo ago

MongoDB Technical Services Engineers use their exceptional problem solving and customer service skills, along with their deep technical experience, to advise customers and to solve their complex MongoDB problems. Technical Service Engineers are experts in the entire MongoDB ecosystem - database server, drivers, cloud and infrastructure. This also includes services such as Atlas (database as a service), or Cloud Manager (which helps customers with automation, backup and monitoring of their MongoDB systems). Our engineers combine their MongoDB expertise with passion, initiative, teamwork and a great sense of humor to help our customers to be successful with MongoDB. We are looking to speak to candidates who are based in Gurugram for our hybrid working model. Cool things you’ll do You'll be working alongside our largest customers, solving their complex challenges - resolving questions on architecture, performance, recovery, security, and everything in between. You'll be an expert resource on best practices in running MongoDB at scale, whatever that scale may be. You'll be an advocate for customers' needs - interfacing with our product management and development teams on their behalf. And you'll contribute to internal projects, including software development of support tools for performance, benchmarking, and diagnostics. What you need We consider all candidates with an eye for those who are self taught, curious, and multi-faceted.. You should also have 3- 5 years of relevant experience, either technical or non-technical customer service experience Excellent communication skills, both written and verbal Good knowledge of operating systems (Linux preferred) Understanding of basic network protocols (DNS, TCP/IP, TLS/SSL) Fundamental computing concepts Desire and ability to rapidly learn a wide variety of new technical skills Strong teamwork: willingness and ability to get help from team members when required, and the good judgment to know when to seek help Ope

javascriptpythonjava
View job →
O
1mo ago

About the Team OpenAI’s Stargate and 3P Engineering teams are responsible for building and scaling the external infrastructure ecosystem that powers advanced AI systems. We work across hyperscalers, colocation providers, cloud partners, and strategic third-party operators to turn contracted capacity into production-ready compute. Our scope spans the full lifecycle of external deployments: commercial alignment, technical readiness, network integration, hardware enablement, operational readiness, and long-range scaling strategy. As OpenAI’s infrastructure footprint expands globally, we need leaders who can convert complex partner environments into reliable, high-velocity capacity for training and inference workloads. About the Role We are seeking a Technical Program Manager, Token-as-a-Service (TaaS) to lead delivery of external compute capacity that directly serves OpenAI model workloads. In this role, you will own complex cross-functional programs that transform third-party infrastructure into usable tokens at scale. You will partner across engineering, capacity planning, networking, hardware, finance, product, and external providers to ensure that deployed capacity translates into real production throughput. This role sits at the intersection of infrastructure execution, systems readiness, and business impact. Success requires strong technical fluency, elite program management, and the ability to drive accountability across internal teams and external partners. This is a high-visibility role with direct impact on OpenAI’s ability to scale model training and inference globally. This role is based in San Francisco, CA, with a hybrid work model of 3 days in office per week. Relocation assistance is available. Key Responsibilities Lead end-to-end delivery programs that convert external infrastructure capacity into production-ready token supply. Own readiness across compute, storage, networking, security, and operational dependencies for third-party environments. Build

awsrestai
View job →

Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Technical Program Manager 1. Overview This role provides Tier 1 and Tier 2 support for Jira Cloud users across Mastercard programs, ensuring seamless onboarding, issue resolution, and vendor coordination. You will work closely with engineering, product, and vendor teams to maintain operational excellence and drive adoption of Jira Cloud and integrated tools like Aha! and X-ray. 2. Role • Serve as the primary contact for Jira Cloud support , handling Remedy tickets, onboarding queries, and user escalations. • Facilitate onboarding of programs to Jira Cloud, including license validation, test case migration, and user provisioning. • Conduct demos and office hours to guide users through Jira workflows and integrations. • Partner with engineering teams to troubleshoot configuration issues, permission models, and data sync errors between Jira and Rally/Aha • Support CRQ planning and execution for Jira schema updates, field validations, and migration logic enhancements. • Coordinate with vendors to resolve plugin issues, validate architecture diagrams, and ensure compliance with Mastercard’s security standards. • Monitor and clean up orphaned work items, duplicate epics, and initiatives across programs using Domo reports and Confluence documentation. • Maintain Confluence pages for s

O
OMP
📍 Shanghai• Full-time
17 days ago

Your challenge As a product consultant, you use your expertise in one of our supply chain planning domains to support our customer project teams. In cooperation with our product experts, you implement project-specific solutions and best practices within your domain of expertise, with a focus on bringing new global, cross-industry product developments to the market. You’re responsible for: Growing expertise in one of our core product functions: network design, demand management, S&OP, operational planning, scheduling, cloud, data management & integration, or value enhancers. Providing our solution architects with direct support within your specific product domain of expertise. Providing implementation consultants with specialist product support in the field of project delivery, documentation, and training across all OMP’s core industries. Solving customer-specific challenges related to your product domain, guided by our product experts. Briefing product managers on potential product differentiators and gaps in our product offering based on interactions with customers, implementation consultants and solution architects. Designing and building prototypes for specific customer problems on an occasional basis. About you Essential talents and qualifications: A Master’s or PhD degree in Engineering, Business Administration, Applied Economics, Mathematics, Computer Science, or similar. Great analytical skills and problem-solving abilities. A desire to take ownership of your projects. Strong affinity with IT systems. Basic programming skills. The flexibility to travel. Comfortable under deadlines and pressure. Bonus points if you have: 1 to 3 years of relevant working experience in supply chain planning or business software implementation. Affinity with mathematical modeling and/or machine learning techniques. Prior experience in one of our core industries. But most of all, you are fun to work with and a great team player! So

sqlmachine learningai
View job →
🔔

Get new cloud operations engineer jobs by email

Daily job updates · Unsubscribe anytime

Explore verified demand

More cloud operations engineer opportunities

Browse all jobs →

Companies hiring

Employers are derived from current jobs in this exact search market.