Jobiba hiring network

Engineering Manager Platform Reliability Jobs

15 active opportunities · Updated for September 2026

Fresh results

15 shown

Explore current engineering manager platform reliability jobs. Use filters to narrow by work mode, employment type, experience and date posted.

A
1mo ago

We're looking for an Engineering Manager who combines strong technical judgment with people leadership to help build the systems and team practices that keep Asana resilient at scale. This role is a great fit for someone who enjoys turning broad reliability problems into clear priorities, and can balance foundational engineering work with iterative delivery. You'll help shape what Platform Reliability means at Asana while building a team that ships durable systems, strong operational practices, and high-trust partnerships. You will define the reliability roadmap for a rapidly growing global platform, transforming reliability into a core architectural advantage. You'll partner closely with platform engineering, infrastructure, and product teams in Warsaw, Reykjavik and San Francisco to protect Asana under real-world load, improve how traffic and failure modes are handled, and ensure reliability is designed in rather than added later. This is a role for someone who can lead through influence, coach engineers, and raise the bar for execution and collaboration. This role is based in our Warsaw office with an office-centric hybrid schedule. The standard in-office days are Monday, Tuesday, and Thursday. Most Asanas have the option to work from home on Wednesdays. Working from home on Fridays depends on the type of work you do, and your recruiter can share more about the in-office requirements. We offer a Contract of Employment (UoP) for our employees in Poland. What you'll achieve Build and lead a new Platform Reliability team, hiring and developing engineers while setting a clear standard for collaboration, ownership, technical excellence and growth. Partner with technical leaders to define the roadmap for core reliability systems such as load shedding, rate limiting, circuit breakers, traffic controls, and other platform guardrails. Establish and evangelize best-in-class operating practices for incident response, postmortems, and proactive risk management through SLOs a

awskubernetesrest
View job →
CV
Company via Lever
📍 Bangalore, Karnataka• Full-time
1mo ago

About the Team When 5% of Indian households shop with us, it’s important to build resilient systems to manage millions of orders every day. We’ve done this – with zero downtime!Sounds impossible? Well, that’s the kind of Engineering muscle that has helped Meesho become the e-commerce giant that it is today. We value speed over perfection, and see failures as opportunities to become better. We’ve taken steps to inculcate a strong ‘Founder’s Mindset’ across our engineering teams, making us grow and move fast. We place special emphasis on the continuous growth of each team member - and we do this with regular 1-1s and open communication. As Engineering Manager, you will be part of self-starters who thrive on teamwork and constructive feedback. We know how to party as hard as we work! If we aren’t building unparalleled tech solutions, you can find us debating the plot points of our favourite books and games – or even gossiping over chai. So, if a day filled with building impactful solutions with a fun team sounds appealing to you, join us. About the Role We are looking for an Engineering Manager to lead the Platform Team at Meesho.In this role, you will be responsible for building and growing a high-impact platform engineering team, owning the technical roadmap, and delivering a unified database platform used by hundreds of engineers across Meesho. This is not a pure operations role rather a platform engineering leadership role that blends system design, automation, developer experience, and reliability. You will work closely with Backend, Data, Infra, and SRE teams to ensure Meesho’s database ecosystem scales reliably while remaining easy to use and cost-efficient.

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Opportunity The Migration Services team builds the critical, data-driven services that seamlessly move customers across environments in real-time. We are looking for an Engineering Manager to lead, mentor, and grow this pivotal team.This is a unique opportunity to shape the team's technical direction and culture from the ground up, and you will be helping build and hire for this team. You are a people-focused leader who is also deeply technical, capable of guiding architectural decisions and willing to be hands-on when needed. This role requires extensive collaboration with numerous Product teams and a close partnership with our Customer Service organization, which utilizes the tooling your team builds to execute customer migrations. If you are passionate about building high-performing teams, driving technical strategy, and delivering high-impact projects, we want to hear from you. What You'll Achieve Build and lead a world-class team. You will recruit, mentor, and support a team of skilled engineers, fostering a strong, collaborative culture and keeping the team motivated and engaged. Own the roadmap and execution. You will partner with product and stakeholder teams to define the team's strategy and roadmap. You'll be responsible for prioritizing work and ensuring the team consistently delivers high-quality, reliable services. Drive engineering excellence. You will hold the team to a high standard of quality, reliability, and operational robustn

sqlpostgresqlmongodb
View job →
A
Asana
📍 Warsaw• Full-time• $348K – $504K/yr
1mo ago

We're looking for a Tech Lead to own technical direction for Asana's Platform Reliability Engineering team. This role is for someone who thrives as a technical leader without needing to take on people management. At Asana, Tech Leads own technical execution, architecture, and long-term engineering strategy. You'll lead complex projects, make key technical decisions, and set the standard for engineering quality. You'll guide ICs through technical design and project leadership, and partner with teams across the company. An Engineering Manager handles people leadership – you own the technical direction. Your team will build core platform systems like load shedding, rate limiting, circuit breakers, and traffic controls that protect Asana under real-world load. This is hands-on work – you'll spend significant time designing and building systems alongside the team, while partnering with other platform teams to make reliability something that's built in, not bolted on. Our tech stack includes: AWS, Kubernetes (EKS), MySQL (RDS), OpenSearch, DynamoDB, Redis, Terraform, Datadog, TypeScript, Scala, Go, and Python. (Yeah, we know this sounds like buzzword bingo – but we want this post to actually show up in your searches.) We're especially interested in people who think like backend engineers but obsess over failure modes, capacity planning, and building systems that degrade gracefully under pressure. This role is based in our Warsaw office with an office-centric hybrid schedule. The standard in-office days are Monday, Tuesday, and Thursday. Most Asanas have the option to work from home on Wednesdays. Working from home on Fridays depends on the type of work you do, and your recruiter can share more about the in-office requirements. We offer a Contract of Employment (UoP) for our employees in Poland. What You'll Achieve: Set technical direction for the team, making sure we're solving the right problems in the right way. Design and build core platform systems – load shedding, ci

typescriptpythonsql
View job →
NR
New Relic
📍 Atlanta• Full-time• From $186K/yr
1mo ago

We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity Are you ready to step into a pivotal leadership role where your engineering depth directly shapes the future of our core platform? As our new Engineering Manager, you will lead a talented, distributed team across US and EU time zones, acting as the critical manager bridging regional collaboration. Our Cloud Foundation team is the backbone of the New Relic platform. In this role, you won't just manage tasks; you will mentor and empower engineers, transitioning our operational framework from a reactive state to a culture of proactive ownership and engineering excellence. You will oversee critical global initiatives, including major regional expansions into FedRAMP High / IL4, India, and Australia. If you thrive on solving complex multi-cloud challenges at an exabyte scale while helping engineers grow in their careers, this is your opportunity to make a lasting impact. What you'll do Empower & Mentor: Lead and nurture a high-performing engineering team across the US and EU, facilitating career development, performance growth, and a collaborative team culture. Drive Strategic Ownership: Champion a shift from reactive delivery to proactive technical ownership, establishing best practices for platform reliability and cross-regional alignment. Lead Regional Expansions: Architect and execute key global infrastructure expansions across complex environments (including FedRAMP High / IL4, India, and Australia). Architect for Extreme Scale: Guide decisions around micr

reactawsazure
View job →
R
Roblox
📍 San Mateo• Full-time• From $295.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Observability team builds the infrastructure that empowers engineers to understand, operate, and improve the Roblox platform and ecosystem. Our team owns the end-to-end observability stack across telemetry, distributed tracing, logging, profiling, storage systems, and developer-facing visualization tools. We are looking for an Engineering Manager to lead the next generation of AI-powered observability platforms. In this role, you will help build intelligent systems that leverage AI to revolutionize CI/CD, testing, and DevOps workflows — enabling engineers to move faster, improve reliability, and operate large-scale distributed systems with greater efficiency and confidence. This is a highly impactful leadership role at the center of Roblox infrastructure. Your work will directly improve developer productivity, platform reliability, and operational excellence across the company. You will partner closely with infrastructure, product engineering, and AI platform teams to shape the future of developer tooling and autonomous operations at scale. You Have 3+ years of engineering management experience with a proven track record of hiring, mentoring, and growing high-performing teams. Strong ex

awsci/cdgit
View job →
P
Peloton
📍 New York• Full-time• From $173.5K/yr
1mo ago

ABOUT THE ROLE Peloton is looking for a talented Data Engineering Manager to join the Data Engineering team. In this role, you will lead the DataOps function, driving operational excellence across our data platforms while managing the successful delivery of data initiatives through a combination of internal and offshore engineering resources. You will work closely with business stakeholders, engineering teams, analytics partners, and platform owners to ensure our data ecosystem remains reliable, scalable, and well-governed. This role combines technical leadership, operational management, and stakeholder engagement, to help support Peloton's growing data and AI needs. This role will be hybrid, not remote. YOUR DAILY IMPACT AT PELOTON Lead the DataOps function within the Data Engineering team, driving operational maturity, platform reliability, and process improvements Manage and mentor offshore engineering resources, providing technical guidance, performance feedback, and delivery oversight Partner with stakeholders across multiple business functions to gather requirements, prioritize work, and ensure successful delivery of data solutions Own intake, prioritization, and execution processes for DataOps requests and operational support activities Drive adoption of data engineering standards, ETL best practices, documentation requirements, and operational procedures Serve as an operational owner for key data platforms, including Airflow, Airbyte, and Looker, owning platform governance activities such as user access reviews, compliance reviews, audit support, and operational controls Coordinate incident management, root cause analysis, monitoring, alerting, and operational readiness efforts across the data platform Help shape the long-term operating model for the Data Engineering team as Peloton continues to scale its data and AI initiatives YOU BRING TO PELOTON 5+ years of experience in data engineering, analytics engineering, software engineering, or related technical

sqlredisaws
View job →
N
Nvidia
📍 Santa Clara, United States
1mo ago

NVIDIA’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern deep learning — the next era of computing — with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as “the AI computing company.” We're looking to grow our company and establish teams with the most thoughtful people in the world. We are looking for an excellent engineering manager to own and deliver an end to end manageability stack for Data Center Systems. We are seeking an experienced manager who is deeply technical, hands-on, and has a wide system view. You will manage a team of experts, design & build OpenBMC based manageability software stack for NVIDIA’s next generation Data Center Compute Systems. We want to grow our teams with the smartest people in the world. If you're creative and autonomous, we want to hear from you! What you’ll be doing: Own and deliver OpenBMC based manageability stack for next generation Data Center Compute Systems. Own firmware delivered to data centers in terms of quality, reliability and telemetry performance. Manage and lead a distributed team of software engineers to deliver firmware stack with high quality. Work with data center architects and cloud customers for correct requirements and scope implementation to ensure speed of light product development. Work closely with cross functional teams to ensure scalable manageability architecture for all data centers products Drive efficiency, reliability and optimization in firmware architecture from a data center view point. Work closely with customers and internal teams to resolve issues at Speed of Light. What we need to see: BS, MS, or PhD in EE/CS or related field o

pythongitai
View job →
CV
Company via Lever
📍 Bangalore, Karnataka• Full-time
1mo ago

About the Team When 5% of Indian households shop with us, it’s important to build resilient systems to manage millions of orders every day. We’ve done this – with zero downtime!Sounds impossible? Well, that’s the kind of Engineering muscle that has helped Meesho become the e-commerce giant that it is today. We value speed over perfection, and see failures as opportunities to become better. We’ve taken steps to inculcate a strong ‘Founder’s Mindset’ across our engineering teams, making us grow and move fast. We place special emphasis on the continuous growth of each team member - and we do this with regular 1-1s and open communication. As Engineering Manager, you will be part of self-starters who thrive on teamwork and constructive feedback. We know how to party as hard as we work! If we aren’t building unparalleled tech solutions, you can find us debating the plot points of our favourite books and games – or even gossiping over chai. So, if a day filled with building impactful solutions with a fun team sounds appealing to you, join us. About the Role We are looking for an Engineering Manager to lead the Database Platform Team at Meesho.In this role, you will be responsible for building and growing a high-impact platform engineering team, owning the technical roadmap, and delivering a unified database platform used by hundreds of engineers across Meesho. This is not a pure operations role rather a platform engineering leadership role that blends system design, automation, developer experience, and reliability. You will work closely with Backend, Data, Infra, and SRE teams to ensure Meesho’s database ecosystem scales reliably while remaining easy to use and cost-efficient.

S
1mo ago

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We’re hiring a talented Software Engineering Manager to lead the Snowtrail infrastructure team at Snowflake. Snowtrail is the infrastructure that enables Snowflake to deliver dedicated coverage for customer-specific workloads. Its innovative approach allows Snowflake to precisely test and measure the impact of changes on individual customers, making it essential for ensuring the platform’s reliability, correctness, and performance. Through query replay, Snowtrail helps us catch regressions early. By leveraging machine learning models to intelligently sample queries and workloads, we continuously optimize for both cost and performance. Evolving Snowtrail to incorporate new engine features while improving scalability, efficiency, and reliability is central to our continued success OUR IDEAL MANAGER WILL HAVE : Strong passion and proven track record for shipping quality software in high code velocity environments 10+ years industry experience designing and building distributed data systems. Excellent problem solving skills, and strong CS fundamentals including data structures, algorithms, and distributed systems. Fluency in SQL, Java, C++, Python or Go. Ability to collaborate well across teams, build high-performing teams and mentor junior engineers. Excellent interpersonal co

pythonjavavue
View job →
DC
Diligent Corporation
📍 New York• Full-time• From $131K/yr
8 days ago

Help shape the technology that enables a global organisation to do its best work. As Senior Manager, Platform Engineering, you’ll lead the team responsible for Diligent’s Atlassian and Microsoft platforms while setting the architectural direction for the wider internal IT estate. You’ll combine people leadership, enterprise platform strategy and hands-on technical judgement to create secure, reliable and scalable experiences for employees worldwide. From modernising service management and automating joiner, mover and leaver processes to enabling AI safely through Microsoft Copilot and Atlassian Rovo, your work will reduce friction, strengthen governance and deliver measurable business impact. Working across IT, Security, HR, Finance, Legal, Compliance and business teams, you’ll turn complex requirements into well-governed platforms that are easy to use, resilient and ready for the future. Here’s a breakdown of what you’ll do (not all of it, just the important stuff) Lead, coach and grow a global team of platform engineers and systems administrators, building a high-performing and inclusive culture. Own the strategy, architecture, governance and roadmap for Atlassian Cloud, including Jira, Jira Service Management, Confluence, Atlassian Guard and Rovo. Set the direction for Diligent’s Microsoft 365 E5 estate, including Teams, SharePoint, Exchange Online, Intune, Defender, Purview, Power Platform and Copilot. Design scalable integration and automation patterns across identity, HRIS, ITSM and business systems using APIs, event-driven automation, Okta Workflows, Power Platform and scripting. Partner with IT Support to improve self-service, automate repetitive work and reduce ticket volume, escalation effort and time to resolution. Establish strong standards for security, access governance, AI adoption, reliability, compliance and business continuity across the internal technology estate. These are the essentials you’ll need to get an interview Significant experience in i

pythonawsgit
View job →
S
1mo ago

Squarespace is looking for a Backend Staff Software Engineer to lead the technical direction of our Communications Platform team. The Communications Platform is a critical engine powering customer engagement across Squarespace — providing internal teams with a scalable, reliable, and secure infrastructure for delivering communications via email, push notifications and in-product messaging. The team owns multiple interconnected production platforms that collectively send over 1.75+ million notifications per day, strategic initiatives including platform modernization, In-Product Placements, and multi-channel communication capabilities. In this role, you’ll serve as the senior technical voice for the team — driving architecture, shaping roadmap execution, and elevating the engineering quality of a small but high-impact team. This is a hands-on role with real technical depth, paired with broad cross-functional influence. This is a hybrid role based in our NYC office (3 days per week), reporting to the Engineering Manager of Communications Platform. You’ll Get To… Drive the architecture, design, and implementation of the Communications Platform’s platforms — including Email Delivery, In-Product Notifications and Push Notifications. Define and own the technical strategy for platform modernization, reliability hardening, and engineering standards across backend services. Provide hands-on technical leadership and mentorship to a team of backend and frontend engineers, helping them grow in system design, decision-making, and ownership. Write high-quality Java code and stay close to execution — leading design reviews, code reviews, and architectural decisions that raise the bar across the codebase. Proactively identify and address systemic risks in performance, reliability, security, and operability before they become incidents. Serve as a key technical partner on cross-team and company-wide architecture initiatives. Balance short-term delivery with long-term platform health,

javamongodbgcp
View job →
S
Squarespace
📍 New York City• Full-time• $185.5K – $299K/yr
1mo ago

Squarespace is looking for a Backend Staff Software Engineer to lead the technical direction of our Communications Platform team. The Communications Platform is a critical engine powering customer engagement across Squarespace — providing internal teams with a scalable, reliable, and secure infrastructure for delivering communications via email, push notifications and in-product messaging. The team owns multiple interconnected production platforms that collectively send over 1.75+ million notifications per day, strategic initiatives including platform modernization, In-Product Placements, and multi-channel communication capabilities. In this role, you’ll serve as the senior technical voice for the team — driving architecture, shaping roadmap execution, and elevating the engineering quality of a small but high-impact team. This is a hands-on role with real technical depth, paired with broad cross-functional influence. This is a hybrid role based in our NYC office (3 days per week), reporting to the Engineering Manager of Communications Platform. You’ll Get To… Drive the architecture, design, and implementation of the Communications Platform’s platforms — including Email Delivery, In-Product Notifications and Push Notifications. Define and own the technical strategy for platform modernization, reliability hardening, and engineering standards across backend services. Provide hands-on technical leadership and mentorship to a team of backend and frontend engineers, helping them grow in system design, decision-making, and ownership. Write high-quality Java code and stay close to execution — leading design reviews, code reviews, and architectural decisions that raise the bar across the codebase. Proactively identify and address systemic risks in performance, reliability, security, and operability before they become incidents. Serve as a key technical partner on cross-team and company-wide architecture initiatives. Balance short-term delivery with long-term platform health,

javamongodbgcp
View job →
S
1mo ago

Squarespace is looking for a Backend Staff Software Engineer to lead the technical direction of our Communications Platform team. The Communications Platform is a critical engine powering customer engagement across Squarespace — providing internal teams with a scalable, reliable, and secure infrastructure for delivering communications via email, push notifications and in-product messaging. The team owns multiple interconnected production platforms that collectively send over 1.75+ million notifications per day, strategic initiatives including platform modernization, In-Product Placements, and multi-channel communication capabilities. In this role, you’ll serve as the senior technical voice for the team — driving architecture, shaping roadmap execution, and elevating the engineering quality of a small but high-impact team. This is a hands-on role with real technical depth, paired with broad cross-functional influence. This is a hybrid role based in our NYC office (3 days per week), reporting to the Engineering Manager of Communications Platform. You’ll Get To… Drive the architecture, design, and implementation of the Communications Platform’s platforms — including Email Delivery, In-Product Notifications and Push Notifications. Define and own the technical strategy for platform modernization, reliability hardening, and engineering standards across backend services. Provide hands-on technical leadership and mentorship to a team of backend and frontend engineers, helping them grow in system design, decision-making, and ownership. Write high-quality Java code and stay close to execution — leading design reviews, code reviews, and architectural decisions that raise the bar across the codebase. Proactively identify and address systemic risks in performance, reliability, security, and operability before they become incidents. Serve as a key technical partner on cross-team and company-wide architecture initiatives. Balance short-term delivery with long-term platform health,

javamongodbgcp
View job →

Anyscale Platform Engineering Leader About Anyscale: At Anyscale , we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray , a popular open-source project that's creating an ecosystem of libraries for scalable machine learning. Companies like OpenAI , Uber , Spotify , Instacart , Cruise , and many more, have Ray in their tech stacks to accelerate the progress of AI applications out into the real world. With Anyscale, we’re building the best place to run Ray, so that any developer or data scientist can scale an ML application from their laptop to the cluster without needing to be a distributed systems expert. Proud to be backed by Andreessen Horowitz, NEA, and Addition with $250+ million raised to date. About the role: Anyscale is looking for an experienced Engineering leader to lead our Infrastructure, SRE and Enterprise Governance Engineering teams. Anyscale aims to provide the next generation of tools and infrastructure to make developing and running distributed AI applications in the cloud using Ray - the popular open source platform used by companies like Netflix, Uber, Instacart and others - seamless. In this position, you will guide the vision, technical direction of the team, and recruit, enable a high-performing engineering team that delivers critical values to developers and Anyscale customers by solving complex distributed systems challenges. You will oversee and drive the strategy and execution of components which includes cluster launcher, cloud providers (AWS/GCP/Azure/etc.), Kubernetes support, cluster autoscaling, control plane, data plane, reliability, billing stack, production database and related components. You will closely work with our customers and our field engineering team to solve their problems, understand their challenges and make sure they are successful. We'd love to hear from you if you have: Solid engineering management experience leading produ

awsazuregcp
View job →
🔔

Get new engineering manager platform reliability jobs by email

Daily job updates · Unsubscribe anytime