About ElevenLabs ElevenLabs is an AI research and product company transforming how we interact with technology. We launched in January 2023 with the first human-like AI voice model. Today, we serve millions of users and thousands of businesses - from fast-growing startups to large enterprises like Deutsche Telekom and Meta. Our investors are some of the world's most prominent, including Andreessen Horowitz, ICONIQ Growth and Sequoia. We've raised $781M in funding and our last valuation was $11B - multiples of 11, always. We have expanded from voice into three main platforms: ElevenAgents enables businesses to deliver seamless and intelligent customer experiences, with the integrations, testing, monitoring, and reliability necessary to deploy voice and chat agents at scale. ElevenCreative empowers creators and marketers to generate and edit speech, music, image, and video across 70+ languages. ElevenAPI gives developers access to our leading AI audio foundational models. Everything we do is the result of the creativity and commitment of our team - builders doing the best work of their lives. We are researchers, engineers, and operators. IOI medalists and ex-founders. If you want to work hard and create lasting positive impact, we want to hear from you. How we work High-velocity: Rapid experimentation, lean autonomous teams, and minimal bureaucracy. Impact not job titles: We don’t have job titles. Instead, it’s about the impact you have. No task is above or beneath you. AI first: We use AI to move faster with higher-quality results. We do this across the whole company—from engineering to growth to operations. Excellence everywhere: Everything we do should match the quality of our AI models. Global team: We prioritize your talent, not your location. What we offer Innovative culture: You’ll be part of a generational opportunity to define the trajectory of AI, surrounded by a team pushing the boundaries of what’s possible. Growth paths: Joining ElevenLabs means joining a
Jobiba hiring network
Reliability Engineer Jobs
2,028 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
About ElevenLabs ElevenLabs is an AI research and product company transforming how we interact with technology. We launched in January 2023 with the first human-like AI voice model. Today, we serve millions of users and thousands of businesses - from fast-growing startups to large enterprises like Deutsche Telekom and Meta. Our investors are some of the world's most prominent, including Andreessen Horowitz, ICONIQ Growth and Sequoia. We've raised $781M in funding and our last valuation was $11B - multiples of 11, always. We have expanded from voice into three main platforms: ElevenAgents enables businesses to deliver seamless and intelligent customer experiences, with the integrations, testing, monitoring, and reliability necessary to deploy voice and chat agents at scale. ElevenCreative empowers creators and marketers to generate and edit speech, music, image, and video across 70+ languages. ElevenAPI gives developers access to our leading AI audio foundational models. Everything we do is the result of the creativity and commitment of our team - builders doing the best work of their lives. We are researchers, engineers, and operators. IOI medalists and ex-founders. If you want to work hard and create lasting positive impact, we want to hear from you. How we work High-velocity: Rapid experimentation, lean autonomous teams, and minimal bureaucracy. Impact not job titles: We don’t have job titles. Instead, it’s about the impact you have. No task is above or beneath you. AI first: We use AI to move faster with higher-quality results. We do this across the whole company—from engineering to growth to operations. Excellence everywhere: Everything we do should match the quality of our AI models. Global team: We prioritize your talent, not your location. What we offer Innovative culture: You’ll be part of a generational opportunity to define the trajectory of AI, surrounded by a team pushing the boundaries of what’s possible. Growth paths: Joining ElevenLabs means joining a
A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role Apollo is Palantir’s autonomous software management and deployment platform. It enables seamless, continuous delivery of mission-critical software (Foundry, Gotham, AIP) across a vast range of environments: on-prem, public cloud, disconnected (air-gapped) networks, and highly regulated settings (including IL-5 and FedRAMP). As a Software Engineer on the Apollo team, you’ll build and operate a large-scale distributed system to allow the remote operation and maintenance of Kubernetes clusters. Our mission is to extract the entire state of a cluster into a portable, high-performance artifact within minutes, enabling full and almost instant cluster reconstruction from the ground up—all while pushing the limits of speed, reliability, and scale. You’ll design and implement backup and restore solutions for Kubernetes, leveraging proprietary compression infrastructure tailored to Palantir’s unique deployment models. You’ll also build and optimize our container artifact store, which is based on the OCI (Open Container Initiative) distribution spec—the industry standard for storing and distributing container images and artifacts. You’ll own the backbone of every environment Apollo supports, from hyperscalers to Army trucks. If you’re excited by challenges at the intersection of container technologies like OCI and docker, storage, and distributed systems, you’ll find opportunities here to dive deep into storage formats and low-level optimizations, where milliseconds matter. As we increasingly automate cluster creation and management on diverse hardware, you’ll play a key role in scaling Palantir’s presence at the edge and solving tough distributed systems proble
A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role Apollo is Palantir’s autonomous software management and deployment platform. It enables seamless, continuous delivery of mission-critical software (Foundry, Gotham, AIP) across a vast range of environments: on-prem, public cloud, disconnected (air-gapped) networks, and highly regulated settings (including IL-5 and FedRAMP). As a Software Engineer on the Apollo team, you’ll build and operate a large-scale distributed system to allow the remote operation and maintenance of Kubernetes clusters. Our mission is to extract the entire state of a cluster into a portable, high-performance artifact within minutes, enabling full and almost instant cluster reconstruction from the ground up—all while pushing the limits of speed, reliability, and scale. You’ll design and implement backup and restore solutions for Kubernetes, leveraging proprietary compression infrastructure tailored to Palantir’s unique deployment models. You’ll also build and optimize our container artifact store, which is based on the OCI (Open Container Initiative) distribution spec—the industry standard for storing and distributing container images and artifacts. You’ll own the backbone of every environment Apollo supports, from hyperscalers to Army trucks. If you’re excited by challenges at the intersection of container technologies like OCI and docker, storage, and distributed systems, you’ll find opportunities here to dive deep into storage formats and low-level optimizations, where milliseconds matter. As we increasingly automate cluster creation and management on diverse hardware, you’ll play a key role in scaling Palantir’s presence at the edge and solving tough distributed systems proble
A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role As a Site Reliability Operations Analyst you are the engine behind Palantir deployments. You are responsible for crafting, implementing and executing processes to streamline workflows and reduce friction. You track and stabilize projects, remove roadblocks, and anticipate customer needs to free up our engineers to focus their time and attention on the technical problems they are best equipped to solve. This position requires a combination of project management, process optimization, and execution skills. You are a person who loves fixing problems and always embraces the best idea, even when it is not your own.
A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role As a Site Reliability Operations Analyst you are the engine behind Palantir deployments. You are responsible for crafting, implementing and executing processes to streamline workflows and reduce friction. You track and stabilize projects, remove roadblocks, and anticipate customer needs to free up our engineers to focus their time and attention on the technical problems they are best equipped to solve. This position requires a combination of project management, process optimization, and execution skills. You are a person who loves fixing problems and always embraces the best idea, even when it is not your own.
A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role As a Site Reliability Operations Analyst you are the engine behind Palantir deployments. You are responsible for crafting, implementing and executing processes to streamline workflows and reduce friction. You track and stabilize projects, remove roadblocks, and anticipate customer needs to free up our engineers to focus their time and attention on the technical problems they are best equipped to solve. This position requires a combination of project management, process optimization, and execution skills. You are a person who loves fixing problems and always embraces the best idea, even when it is not your own.
GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An overview of this role As a Staff Backend Engineer (AI) in the Verify stage at GitLab, you'll help shape and scale the core infrastructure behind GitLab CI. You'll play a central role in how we integrate AI into CI/CD workflows. Your work will impact performance, reliability, and usability for people running millions of CI jobs, from small teams to the largest enterprises. AI is a top priority in the year ahead. In this role, you'll go beyond using AI tools and help define how we design, build, and iterate on AI-assisted and agentic CI experiences. You'll set standards for what good looks like across our AI agent portfolio, inc
GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An Overview of this role The Integration Engineer for RPA UiPath will be responsible for designing, developing, and implementing enterprise-level RPA solutions using UiPath to automate manual and repetitive processes. They will collaborate with business stakeholders to gather requirements, analyze processes, and propose automation solutions that align with the company's goals and objectives. The successful candidate will ensure timely delivery and quality of RPA automations, provide production support, and continuously improve and optimize existing solutions to maintain efficiency and reliability. What You’ll Do Desi
GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. Senior Backend Engineer - Platform Insights An overview of this role As a Senior Backend Engineer in Platform Insights, you'll build and evolve the core backend systems that power GitLab's data platform. In this role, you'll focus primarily on Go-based development for the Data Insights Platform and Siphon, with an emphasis on core system design, production readiness, and deployment across GitLab environments. You'll own high-throughput, multi-component backend systems and help shape architecture, reliability, deployment, and operational maturity across GitLab's SaaS, Dedicated, and Self-Managed environments. What you'll do Design
Location Details: Colombia, remote At GoDaddy, the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) , and some work entirely remotely. This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team Join our Growth team, where you'll build intelligent agents that serve hundreds to thousands of customers in production. We're at the forefront of applying AI to solve real business problems at scale. You'll architect, deploy, and operate AI-powered systems in live environments, tackling challenges in reliability, scalability, and performance as we grow our AI footprint across GoDaddy. What you'll get to do... Design, build, and deploy production-ready AI agents using Node.js, integrating LLMs (e.g., OpenAI, Anthropic) into scalable backend services, and delivering AI-powered experiences through full-stack applications with React frontends. Architect and manage scalable cloud infrastructure on AWS or Azure to support AI workloads for thousands of users, including database design and end-to-end system ownership. Develop and optimize APIs that orchestrate AI agents, handle asynchronous processing, and manage complex workflows with a focus on performance, reliability, and observability in production environments. Work across the full stack and collaborate with cross-functional teams to define scalable, available, and maintainable technical solutions while reducing technical debt and strengthening engineering foundations. Mentor junior engineers on full-stack and AI best practices, and actively participate in on-call rotations, incident response, and post-mortems to ensure high operational standards. Your experience should include... 5+ years of experience building and scaling full-stack applications, with a strong focus on bac
Who We Are At Justworks, you’ll enjoy a welcoming and casual environment, great benefits, wellness program offerings, company retreats, and the ability to interact with and learn from leaders in the startup community. We work hard and care about our most prized asset - our people. We’re helping businesses get off the ground by enabling them to focus on running their business. We solve HR issues. We’re data-driven and never stop iterating. If you’d like to work in a supportive, entrepreneurial environment, are interested in building something meaningful and having fun while doing it, we’d love to hear from you. We're united by shared goals and shared motivations at Justworks. These are best summed up in our company values, which are reflected in our product and in our team. Our Values If this sounds like you, you’ll fit right in. Who You Are You are self-driven and like to work with others to remove roadblocks. You are curious, and love to explore and learn new technologies. You have demonstrated the ability to build, deploy and maintain large-scale, complex applications. You care more about solutions and impacting the customer experience than using a particular tool or framework. Your Success Profile What You Will Work On Independently own larged-scoped projects from initial discovery through launch with a focus on outcomes and customer impact Collaborate with a cross functional team to find solutions to customer challenges Improve the performance, reliability, and scalability of our existing systems Promote engineering excellence through technical leadership, knowledge-sharing and mentorship Add capabilities to our high-volume, fault-tolerant processing infrastructure Enhance our testing, monitoring and continuous deployment infrastructure Keep extremely sensitive data compartmentalized and secure How You Will Do Your Work As a Software Engineer, how results are achieved is paramount for your success and ultimately result in our success as an organization. In this
About Mixpanel Mixpanel is the leading product intelligence and analytics platform, trusted by more than 29,000 companies to help understand how people use the products they build. By combining powerful analytics with AI that knows your business, Mixpanel helps teams see what’s working, diagnose what’s not, and decide what to build next. Learn more at mixpanel.com . About the DevInfra Team The DevInfra team is Mixpanel’s platform engineering team. Our mission is to be a force multiplier for Mixpanel engineering as a whole. We achieve this by partnering with our engineering teams to build a world-class software development life cycle together. About the Role We are looking for an engineer with a strong track record of solving tooling pain points, automating away toil, and helping teams deliver software faster, more safely, and more reliably. Someone who is energized by supporting their fellow engineers. An engineer who has good taste and leverages AI to great effect without generating slop. Frequent forays into uncharted territory requires our team to be highly adaptable and always ready to learn. We serve as trailblazers for the rest of Mixpanel engineering. Responsibilities We partner with a wide variety of teams to build out their software development lifecycle to be optimized for speed, safety, and reliability. We support teams writing front-end UI code as well as teams maintaining our highly stateful storage systems deep in our stack. We are responsible for deploying AI tools like Claude Code to our entire engineering team. We are building agents into our platform to do things like augment on-call response and automatically one-shot bugs that come through a team’s triage queue. We support systems that ingest more than 1 Trillion user-generated events every month while balancing low end-to-end query latency. Mixpanel queries typically scan more than 1 Quadrillion events over the span of a month. More details in this blog post . Mixpanel runs entirely on Google Cl
At Playlist, life's richest moments happen when people step away from screens to move, connect, explore, and play. We're building the definitive platform for intentional living, connecting people with inspiring experiences in fitness, wellness, and beyond. With popular brands like Mindbody and ClassPass, Playlist empowers businesses and individuals, making it effortless for aspirations to become actions. Join us in reshaping technology's role to foster meaningful, real-world connections. Mindbody equips wellness entrepreneurs with technology to support thriving businesses and create exceptional experiences. Innovation and curiosity drive our culture, connecting businesses and individuals through cutting-edge solutions. Join us if you're passionate about enhancing wellness through technology. The Role You'll Play: At Playlist, we're reimagining how technology can foster meaningful, real-world connections. As a Senior Software Engineer on the SmartDesk team, you'll be at the forefront of building an AI-powered front desk assistant that transforms how wellness businesses communicate and operate. Crafting end-to-end AI-powered experiences that seamlessly connect wellness businesses with their clients Designing and implementing robust web and messaging services that power conversational and workflow automation Integrating large language models (LLMs) and generative AI frameworks with a laser focus on safety, reliability, and user trust Collaborating closely with product, design, and applied AI teams to transform complex challenges into intuitive solutions Architecting scalable systems across frontend, backend, and integration layers Mentoring teammates through thoughtful code reviews and
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your Opportunity As a Senior Software Engineer within the Container Fabric (CF) organization, you will be a key driver in evolving New Relic’s global internal platform. We are looking for an operations-heavy engineer with 5–8 years of relevant experience who can leverage open-source and custom tooling to orchestrate and maintain large-scale Kubernetes environments. You will play a "Captain" role—leading critical deliverables and mentoring junior engineers while maintaining the reliability of our global fleet. What You'll Do Architectural Leadership: Drive the design and implementation of internal tools, specifically focusing on Kubernetes Operators and Controllers to automate resource management. Platform Orchestration: Lead complex, large-scale infrastructure shifts. Operational Excellence: Take ownership of incident response, author comprehensive retrospectives, and implement systemic hardening to prevent recurrence using advanced overcommit strategies. This Role Requires Experience: 5–8 years in a DevOps, Site Reliability, or Infrastructure Engineering role. Kubernetes Mastery: Deep internals knowledge of Kubernetes and hands-on experience writing custom operators. Tooling Proficiency: Strong experience building production-grade tools and services, specifically for infrastructure automation. Operations-Heavy Mindset: A proven track record of Day 1/Day 2 operations for a large-scale Kubernetes fleet, handling high-severity incidents, and improving SLA compliance through auto
Get new reliability engineer jobs by email
Daily job updates · Unsubscribe anytime