ROLE DESCRIPTION: We’re looking for a Senior Platform Backend Developer who can help us support the development organization to deliver value to customers in a reliable, efficient, and safe manner. You’ll be working in a focused team that owns one or more pieces of the production application environment and the developer experience, you will own and deliver in service of quarterly goals on the team. ABOUT THE TEAM: This role is within our Backend Platform team. The team primarily uses Go, Scala, and PHP and has expertise in technologies such as Kafka, various AWS services, and some infrastructure-as-code tools. Your primary focus will be on developing services and tools for our product development teams as well as modernizing our existing platform. Based out of British Columbia, you will report to the Senior Manager, Software Development, DevOps. WHAT YOU’LL DO: Design and build software - tools, libraries, automation, services, and glue scripts Responsible for the reliability, security, and integrity of our large, cloud-based platform Participate in a flexible on-call rotation Lead by owning project milestones, epics or features Practice continuous improvement, contributing to culture, process, and direction in your team and across our department Develop processes and automation to eliminate repetitive tasks Design and build our infrastructure platform Identify and implement new platform features Research and evaluate new technologies Refactor, rewrite or retire existing platform features Operate our developer experience and production application environments Diagnose and repair our distributed systems Perform maintenance, upgrades, and migrations Control or eliminate repetitive tasks, alert noise, and business-as-usual work Enable development teams Provide executable interfaces to our infrastructure platform Provide tools and best practices to support the entire software development lifecycle Collaborate with others across the orga
Jobiba hiring network
Reliability Engineer Jobs
2,049 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
We're looking for a Customer Experience Operations Manager who serves as the primary technical administrator and subject matter expert for the Customer Office tech stack and related platforms. Reporting to Senior Director, you'll manage and optimize system configurations, workflows, automations, integrations, and reporting to support customer engagement strategies and operational processes. In this role, you'll drive workflow design and AI-enabled automation across the customer lifecycle to enhance efficiency, streamline processes, and improve service delivery. This role is open to remote-applicants within the state of Maharashtra, India. WHAT YOU’LL DO: Customer Experience Systems Administration & Ownership Act as the primary technical administrator and subject matter expert for the Customer Office tech stack, including Gainsight, Kantata (Mavenlink), and related platforms. Own, manage, and optimize system configurations, including workflows, automation, business rules, integrations, and reporting. Ensure systems are accurately configured, integrated, and aligned to enable customer engagement strategies and operational processes. Provide advanced system support and troubleshooting, proactively resolving issues to minimize disruption and ensure continuity for Customer-facing teams. Monitor system performance, conduct regular audits, and implement enhancements to maintain scalability, reliability and best-in-class system operations Workflow Design, Automation & AI Enablement Design and implement end-to-end workflows across the customer lifecycle, ensuring seamless handoffs between teams and systems. Drive AI and automation initiatives, including predictive health scoring, intelligent triggers, and workflow automation to improve customer engagement and team efficiency. Identify and implement process improvements, leveraging automation and system enhancements to reduce manual work and optimize performance. Partner with cross-
PagerDuty (NYSE:PD) is a leader in Digital Operations Management. In an always-on world, organizations of all sizes trust PagerDuty to help them deliver a perfect digital experience to their customers, every time. Teams use PagerDuty to identify issues and opportunities in real time and bring together the right people to fix problems faster and prevent them in the future. Over 13,000 organizations (including 60 of Fortune 100) rely on PagerDuty to succeed with Digital Transformation, Cloud Migration, and DevOps Modernization. Notable customers include GE, Cisco, Genentech, Electronic Arts, Cox Automotive, Netflix, Shopify, Zoom, DoorDash, Lululemon and more. We are expanding rapidly as a platform for Digital Operations Management using AI/ML and Automation and growing our adoption by Development, IT, Customer Service, Security, and other teams across the organization. PagerDuty is growing, and we are looking for an experienced Salesforce Developer to help design and implement a world-class, global network. As a Salesforce Developer, you will help design, build, and support PagerDuty's growing enterprise application environment – ensuring availability, reliability, and security for our users and critical business data. About This Role Together with the other members of the enterprise applications team, you will have the opportunity to re-define how PagerDuty designs, builds, and maintains a growing suite of applications around the world. Responsibilities Develop and maintain complex Apex classes, triggers, Lightning Web Components (LWC), Visualforce pages, Flows and other declarative configuration to support business requirements and system integrations. Build robust integrations between Salesforce and other enterprise
About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. Available Locations Sweden About the Role You aren't just selling a product; you’re selling the security, performance, and reliability that major enterprises require to stay competitive. You will be a foundational part of our growth story in the region, bridging the gap between complex technical challenges and the business value that keeps our customers ahead of the curve. Why You’ll Love This Role This is a builder’s role. We’re looking for a hun
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As the Senior/Principal Product Manager for Engine Systems Foundations, you will drive the vision and strategy for the most foundational parts of the Roblox game engine and be hands-on with the execution and delivery of products that impact over 130 million players every day. This team is responsible for the core performance, reliability, and efficiency of the engine, and it owns key features like our memory allocation library, thread/work dispatch system, and the backing APIs that power our creator performance tooling. If you are a visionary product leader who thrives on deeply technical challenges to improve the speed and quality of a system, you’ll be a great fit! The role is based in San Mateo, CA (hybrid with Tues-Thurs onsite). You will: Define the long-term vision and strategy for Systems Foundations, ensuring we have plans in place to continually invest in the core pieces of a high-performance, realtime game engine. Take ownership of the engine-related content in the public Creator Analytics creators use to monitor the experiences on Roblox, ensuring we’re delivering actionable insights. Work with the Creator organization to define and drive end-to-end performance workfl
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Trust and Safety Operations team is focused (ok, maybe obsessed) on scaling Roblox's Operations organization and transforming our customer experience through our multi-year vision and strategy execution. The team provides support to Roblox's global players, developers, and advertisers. As a Lead, Product Support Ops on our Safety Ops Team, you’ll be mentor and lead the team that is focused on Escalations, Product Reliability, Process Improvement and Design and plays a critical role in managing and resolving escalations and high-impact incidents affecting our Roblox users and platform. Your team will be responsible for identifying trends, enhancing support processes, improving agent efficiency, and reducing operational costs while maintaining high reliability and service quality. You will be reporting to the Trust & Safety India Operations Manager and leading the Product Support Specialist team, providing coverage across billing, accounts, creator, and adjacent domains. This is a player-coach role, weighted toward leadership. You'll still go deep on the hardest problems and stay close enough to the work to be a genuine execution expert, but your primary leverage is the team, and you'
As an Enterprise Customer Success Manager, you will proactively drive new product attachment and effective strong relationships across our largest and most strategic customers. You’ll advocate for the customer internally and focus on a positive customer experience. Interactions are rooted in relationship-management, first and foremost, while also advocating for growth opportunities. Enterprise Customer Success Managers follow a well-defined methodology that helps them identify the customer's unique needs and clearly convey the value of the Datadog product. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Act as a strategic partner to customers, orchestrating cross-functional internal teams and engaging executive, technical, and business stakeholders to understand customer goals and translate them into a clear, deliverable Datadog value narrative. Proactively build and maintain executive relationships to deliver clear, outcome-driven value stories that connect Datadog technical use cases to measurable business results. Lead QBRs and strategic reviews as a forum to demonstrate impact, align on priorities, and define next-step initiatives. Analyze adoption and usage trends to quantify value delivered, extract insights from large datasets, identify gaps, and drive financially grounded commercial recommendations and strategic opportunities. Position Datadog as a critical observability platform that enables reliability, efficiency, and informed decision-making. Own and project manage the on-boarding process for new customers Collaborate cross-functionally with AEs, SEs, TAM, Product, Support, Enablement and other technical teams to ensure consistent value delivery and messaging. With demonstrated understanding of observability and security pla
The Enablement Delivery Specialist is a pivotal execution-focused role within the GTM Enablement organization, responsible for the end-to-end orchestration of world-class training experiences. Reporting to the Manager of Delivery, this role ensures that every training event, from foundational bootcamps to advanced skill-building workshops, is delivered with the logistical precision and professional excellence required to drive field performance. You will act as the "Producer" and "Emcee" for our regular and ad-hoc in-person and virtual events, serving as the connective tissue between strategic program design and field-facing execution. Based in Chicago, USA, you will thrive in a high-growth environment that values intellectual honesty, radical accountability, and a "team first" mindset. Success in this role means moving beyond simple event planning and facilitation to become a trusted partner who enhances the learner experience and provides critical insights for continuous enablement improvement. We are looking to speak to candidates who are based in Chicago for our hybrid working model. Responsibilities Training Event Orchestration and Logistics End-to-End Management: Own the full logistical lifecycle for in-person training events, including venue selection, meeting room configuration, catering, and on-site technology Operational Precision: Ensure flawless accuracy in scheduling, invitations, reminders, and learner communications across all assigned sessions Execution Reliability: Execute assigned tasks within scope and on time using project management tools, maintaining high delivery reliability Stakeholder Coordination: Partner with Workplace and TRO teams to orchestrate seamless delivery, owning handoffs and vendor coordination with "radical follow-through" Data Integrity: Maintain accurate program data (attendance, survey results) with zero critical errors, ensuring visibility for stakeholders Event Production and Emceeing Producer Role: Serve
The Enablement Delivery Specialist is a pivotal execution-focused role within the GTM Enablement organization, responsible for the end-to-end orchestration of world-class training experiences. Reporting to the Manager of Delivery, this role ensures that every training event, from foundational bootcamps to advanced skill-building workshops, is delivered with the logistical precision and professional excellence required to drive field performance. You will act as the "Producer" and "Emcee" for our regular and ad-hoc in-person and virtual events, serving as the connective tissue between strategic program design and field-facing execution. Based in our Dublin International Headquarters, you will thrive in a high-growth environment that values intellectual honesty, radical accountability, and a "team first" mindset. Success in this role means moving beyond simple event planning and facilitation to become a trusted partner who enhances the learner experience and provides critical insights for continuous enablement improvement. We are looking to speak to candidates who are based in Singapore for our hybrid working model. Responsibilities Training Event Orchestration and Logistics End-to-End Management: Own the full logistical lifecycle for in-person training events, including venue selection, meeting room configuration, catering, and on-site technology Operational Precision: Ensure flawless accuracy in scheduling, invitations, reminders, and learner communications across all assigned sessions Execution Reliability: Execute assigned tasks within scope and on time using project management tools, maintaining high delivery reliability Stakeholder Coordination: Partner with Workplace and TRO teams to orchestrate seamless delivery, owning handoffs and vendor coordination with "radical follow-through Data Integrity: Maintain accurate program data (attendance, survey results) with zero critical errors, ensuring visibility for stakeholders Event Production and Emc
We're transforming the grocery industry At Instacart, we invite the world to share love through food because we believe everyone should have access to the food they love and more time to enjoy it together. Where others see a simple need for grocery delivery, we see exciting complexity and endless opportunity to serve the varied needs of our community. We work to deliver an essential service that customers rely on to get their groceries and household goods, while also offering safe and flexible earnings opportunities to Instacart Personal Shoppers. Instacart has become a lifeline for millions of people, and we’re building the team to help push our shopping cart forward. If you’re ready to do the best work of your life, come join our table. Instacart is a Flex First team There’s no one-size fits all approach to how we do our best work. Our employees have the flexibility to choose where they do their best work—whether it’s from home, an office, or your favorite coffee shop—while staying connected and building community through regular in-person events. Learn more about our flexible approach to where we work. Why this role is on the menu The CX Shopper Vendor Management team at Instacart is dedicated to empowering Shoppers to deliver exceptional service. By providing co-pilot support and leveraging data-driven insights, the team enhances processes, drives innovation, and ensures a seamless experience for both Shoppers and customers. Operating within a fast-paced, cross-functional environment and working closely with a network of BPO partners, this role is what keeps high standards of quality, efficiency, and operational excellence front and center across the shopper support ecosystem. Without this owner in place, vendor performance gaps surface too slowly and the data informing decisions loses its reliability. A year from now, this person will have tightened the feedback loop between Instacart and its BPO partners, built scorecards the broader team trusts,
We're transforming the grocery industry At Instacart, we invite the world to share love through food because we believe everyone should have access to the food they love and more time to enjoy it together. Where others see a simple need for grocery delivery, we see exciting complexity and endless opportunity to serve the varied needs of our community. We work to deliver an essential service that customers rely on to get their groceries and household goods, while also offering safe and flexible earnings opportunities to Instacart Personal Shoppers. Instacart has become a lifeline for millions of people, and we’re building the team to help push our shopping cart forward. If you’re ready to do the best work of your life, come join our table. Instacart is a Flex First team There’s no one-size fits all approach to how we do our best work. Our employees have the flexibility to choose where they do their best work—whether it’s from home, an office, or your favorite coffee shop—while staying connected and building community through regular in-person events. Learn more about our flexible approach to where we work. Why this role is on the menu The CX Shopper Vendor Management team at Instacart is dedicated to empowering Shoppers to deliver exceptional service. By providing co-pilot support and leveraging data-driven insights, the team enhances processes, drives innovation, and ensures a seamless experience for both Shoppers and customers. Operating within a fast-paced, cross-functional environment and working closely with a network of BPO partners, this role is what keeps high standards of quality, efficiency, and operational excellence front and center across the shopper support ecosystem. Without this owner in place, vendor performance gaps surface too slowly and the data informing decisions loses its reliability. A year from now, this person will have tightened the feedback loop between Instacart and its BPO partners, built scorecards the broader team trusts,
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Lyft is hiring a Senior Financial Data Analyst to lead the Reporting sub-team within Finance Data & Insights. This sub-team is core to ensuring all key Run the Business (RTB) financial reports are complete, accurate, and operationally reliable on a day-to-day basis. The mandate is two-fold: build repeatable systems, playbooks, and documented workflows that reduce reactive work, and own the primary stakeholder relationships for RTB work across teams. This is a high-visibility role that combines hands-on ownership of critical reporting with team leadership and process-building. Responsibilities: Own all RTB and Compliance reporting (tax, airports, audit requests, etc.), ensuring accuracy, completeness, and operational reliability Build and maintain process documentation, checklists, and operational playbooks that make reporting workflows repeatable and resilient Receive and operationalize reports stabilized by Finance Data Products, integrating them into day-to-day RTB reporting operations Reduce fire-drill, ad-hoc response patterns by replacing them with documented, recurring workflows and automation Identify and drive automation opportunities across RTB reporting processes to improve efficiency, reduce manual effort, and further reduce ad-hoc work Own primary stakeholder relationships for RTB work, serving as the go-to point of contact across teams Lead large-scale, cross-functional initiatives with significant autonomy — from problem definition through execution — proactively communicating progress and risks to stakeholders and management Proactively identify risks, gaps, and opportunities for improvement within RTB reporting and drive initiatives to address them Experience: BA/BS with 5+ years of experience in finance, accounting, business, consulting, or analytics Advanced proficiency in
About Flexdrive At Flexdrive, we're at the forefront of revolutionizing transportation by building the operational backbone for autonomous vehicle (AV) fleets. As a leader in fleet management, we're leveraging our expertise to enter the AV space, forming strategic partnerships with cutting-edge technology providers. We're looking for dedicated team members to help us pioneer this new chapter, starting with our first AV depot in Nashville, Tennessee. The Opportunity Reporting to the Service Lead, Flexdrive is seeking a decisive and security-conscious Repair Coordinator to own the complete repair and parts lifecycle for our 24/7 AV operations — from the initial intake of a vehicle requiring service, through parts procurement and consumption, to the vehicle's final release back into the active fleet. This role is the central hub for service workflow, technician support, and inventory infrastructure: coordinating work assignments, comprehensive vehicle information, and correct parts availability to minimize vehicle downtime and uphold the high availability standards required for an autonomous fleet. The Repair Coordinator directly impacts the operational efficiency, inventory accuracy, and reliability of Flexdrive's AV deployment. This position offers unique exposure to cutting-edge AV technology and supply chain management in a highly regulated environment. The role requires someone who can balance operational efficiency with uncompromising security standards, provide 24/7 support coverage, and maintain perfect inventory accuracy. If you excel at detailed inventory management, shop-flow coordination, and thrive in mission-critical operations, we encourage you to apply. Work Schedule & Shift Availability Day Window Shift Pattern This role supports 24/7 AV depot operations and requires availability to work day window shifts, with hours scheduled between 6:30am - 9:30pm. Specific shift times and schedules will be determined based on operational needs and business requ
About the Team OpenAI's mission is to ensure that artificial intelligence benefits all of humanity. OpenAI for Government works with U.S. and allied government institutions to support the responsible adoption of AI across defense, intelligence, federal civilian, state and local, and international public-sector missions. We work at the intersection of technology, policy, operations, security, and delivery to help public servants use frontier AI in ways that are effective, trusted, aligned with democratic values, and grounded in real-world consequences. The team helps government agencies transform how they work through secure, compliant AI tools and mission-aligned deployments, including ChatGPT Enterprise, ChatGPT Gov, APIs, Codex, and emerging frontier capabilities. We partner with government leaders, operators, technologists, and policy stakeholders to translate cutting-edge AI into measurable mission impact while meeting government requirements for safety, reliability, security, compliance, and trust. Cyber is a critical government mission area. OpenAI's government cyber work brings together OpenAI for Government, Product, Research, Security, Product Policy, Legal, Global Affairs, and Communications to help trusted public-sector defenders responsibly use AI to protect government networks, critical infrastructure, and national-security systems. About the Role We are seeking a senior government cyber leader to serve as Head of Government Cyber Integration for OpenAI for Government. This role will integrate OpenAI's government cyber work across strategy, testing and evaluation, trusted access, deployment, policy, security, and external engagement. This is a matrix leadership role, not a replacement for line management. Product teams still own product roadmaps. Research and Safety still own model capability measurement and risk mitigation. Security still owns OpenAI's security posture and customer security requirements. Product Policy, Legal, and Global Affairs still
About the Team OpenAI is evaluating multiple infrastructure pathways, including powered land, colo/BTS, and NeoCloud opportunities. The Site Readiness & Development team provides the diligence layer needed to compare opportunities, identify risk, and support credible deployment decisions across those pathways. About the Role The NeoCloud & Colo Due Diligence Lead will evaluate third-party infrastructure opportunities where OpenAI is considering deployment through NeoCloud, colo, or BTS structures. This role will focus on facility and deployment readiness, including MEP readiness, rack strategy, developer capability, facility design, power deliverability, schedule credibility, and operating assumptions. Unlike the land diligence team, this role is centered on technical and operational readiness of third-party infrastructure rather than greenfield site master planning, civil development, and entitlement strategy. This is an individual contributor lead role and does not have direct reports initially. The role determines whether each opportunity is fit-for-use and fit-for-service against OpenAI facility, rack, power, network, reliability, and operational standards; identifies material deficiencies and tracks remediation with developers/operators; and evaluates commissioning, validation, AHJ/code, and deployment interfaces such as structured cabling, network readiness, and high-density rack support where relevant. Key Responsibilities Lead diligence on NeoCloud, colo, and BTS opportunities across technical and operational readiness dimensions. Assess each opportunity against OpenAI facility, rack, power, network, reliability, and operational standards to determine deployment fit. Validate MEP readiness, rack deployment strategy, facility design assumptions, power deliverability, and schedule credibility. Identify material deficiencies and work with developers/operators to define remediation plans, owners, timing, and residual risk. Review reliability, availabilit
Get new reliability engineer jobs by email
Daily job updates · Unsubscribe anytime