Jobiba hiring network

Software Reliability Engineer Jobs

6,428 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current software reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

M
Mongodb
📍 San Francisco• Full-time• From $126K/yr
1mo ago

Join the Atlas Search team to design and develop the next generation of Semantic and Vector Search infrastructure. Atlas Search is a growing cloud service that allows users to execute complex search and vector search queries using the MongoDB Query Language. Our users can focus on relevance and data retrieval instead of the machinery needed to search data at scale. Our team is building a cloud-based distributed system responsible for the core components of search including data ingestion, performance, query language, query execution, for both relevance-based search and vector search. Our product is being adopted quickly and there are many interesting projects. This is a technical role where you will be responsible for the infrastructure and features enabling our at-scale cloud service powering vector and semantic search. We are looking to speak to candidates who are based in the San Francisco Bay Area for our hybrid working model. What You’ll Do Lead complex projects across the MongoDB ecosystem, for instance, development of a new Search deployment framework within the MongoDB managed cloud Set project level strategy, architect features, and lead projects to successful execution Identify, design, and implement features enhancing our reliability, performance, security and efficiency Perform code reviews with peers and make recommendations on how to improve our software development processes Influence and grow team members through active mentoring and leading by example What We Look For 5+ years experience in data management/search systems, ideally with a strong distributed systems and infrastructure background Experienced in the development and maintenance of concurrent, stateful services Eager to shape the technological direction of a complex system and have the ability to lead initiatives through collaboration with others Experienced in writing features, debugging and optimizing multithreaded applications written in Java Familiarity with LLM

javamongodbaws
View job →

We're looking for a Senior Engineer with a strong background in computer science fundamentals, systems design, experience in the Java ecosystem, streaming systems, and data-intensive applications to join our engineering team. In this role, you will be instrumental in designing, building, and optimizing the underlying data structures, algorithms, and database interactions that power our generative AI platform, code generation and migration tools. This involves crafting sophisticated orchestration layers, robust integration points, and high-performance data systems that seamlessly connect and leverage advanced AI capabilities for code generation and building a sophisticated data migration suite using a modern technology stack, which includes Java, Spring Boot, Kafka, Debezium, and React.You will work on critical components that ensure the scalability, efficiency, and reliability of our services, collaborating closely with AI researchers, product management and other engineers to design and implement cutting-edge products that solve complex customer challenges.. We are looking to speak to candidates who are based in Sydney for our hybrid working model. The ideal candidate for this role will have 6+ years of engineering experience in backend systems, distributed systems, or core platform development. Proficiency in one or several of Java, Rust, C/C++, and/or Python, with a strong understanding of systems-level programming, memory management, and performance tuning. Extensive experience with streaming data platforms such as Apache Kafka and Change Data Capture (CDC) tools like Debezium Extensive experience with relational data modeling and hands-on experience with at least one SQL database (Postgres, MySQL, etc) Exposure to client-side technologies such as JavaScript and React is a plus Good understanding of algorithms, data structures and their time and space complexity Curiosity, a positive attitude, and a drive to continue learning Excellent verbal and wri

javascriptpythonjava
View job →
M
Mongodb
📍 Sydney• Full-time
1mo ago

We’re looking for a Software Engineer 3 to help bring Voyage’s embedding models - used for semantic search, retrieval, and AI-native experiences; to the platforms and environments where customers already run their workloads, beyond first-party MongoDB Atlas. You’ll join the broader Search and AI Platform organization and collaborate closely with the engineers building Voyage’s first-party inference. Together, we’re extending that platform across cloud marketplaces, third-party inference providers, and self-managed deployments so customers get the same Voyage models, behaving consistently, wherever they choose to run them. As a Software Engineer 3, you'll focus on building the systems, tooling, and deployment workflows that power third-party model delivery. You'll own key components of how Voyage models are packaged, validated, and deployed, work across teams to ensure tight integration with the core inference platform, and contribute to delivery surfaces designed for reliability, observability, and ease of use. We are looking to speak to candidates who are based in Sydney for our hybrid working model. What you'll do Port and tune the model server that runs Voyage embedding and reranking models: improving inference performance, consistency, and runtime behavior across environments Productionize new Voyage models for delivery beyond first-party Atlas, owning the packaging, configuration, and deployment workflows that get them running on AWS, Azure, GCP and more Design correctness, correlation, and performance validation that proves third-party deployments match first-party behavior Build operability into every surface: structured logging, metrics, diagnostics, and health checks with tools like Prometheus and OpenTelemetry Debug problems that span model servers, containers, deployment configuration, and partner cloud environments Work alongside Voyage's model-serving teams, and partner with GTM, SAs, TSEs, and strategic customers on the hardest external deployments Who

pythonmongodbaws
View job →
O
Okta
📍 Washington• Full-time• From $136K/yr
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Classified Senior Software Engineer in Test Opportunity We are seeking a highly specialized and experienced individual for a critical Software Engineer in Test role within a restricted environment. This position is vital for the validation and certification of all platform releases, ensuring customer-critical flows remain intact and reliable. This role requires a unique blend of testing acumen, operational experience, and system-level understanding. The successful candidate will serve as the sole Quality assurance resource in this environment, requiring them to wear multiple hats across the software lifecycle—from understanding the underlying product infrastructure and release versioning to certifying the final product. A passion for ensuring the uptime and reliability of large-scale, mission-critical software is paramount. What you’ll be doing Analyze/refine requirements with Development and Engineering Leads, focusing on test planning and scope to certify releases. Participate effectively in a scrum team environment. Manage deliverables working under minimal supervision. Partner with Development and Release Teams to build and execute on product infrastructure growth and high availability. Understand Okta’s top customers usage patterns to define critical test coverage. Review new changes going into test and staging environments. Ensure new changes are compatible with existing customer deployments and usage. Develop, maintain, and debug end-to-end

javaawsrest
View job →
O
Okta
📍 India• Full-time
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Get to know Okta Okta is The World’s Identity Company. We free everyone to safely use any technology anywhere, on any device or app. Our Workforce and Customer Identity Clouds enable secure yet flexible access, authentication, and automation that transforms how people move through the digital world, putting Identity at the heart of business security and growth. At Okta, we celebrate a variety of perspectives and experiences. We are not looking for someone who checks every single box we’re looking for lifelong learners and people who can make us better with their unique experiences. Join our team! We’re building a world where Identity belongs to you. About Technology Data and Intelligence at Okta At Okta, the Technology Data and Intelligence (TDI) team drives internal efficiency through secure, scalable, and innovative systems. TDI partners with teams across the company to build and support the infrastructure, automation, and enterprise applications that keep operations running smoothly. Focused on enabling productivity and aligning technology with business goals, TDI plays a vital role in both day-to-day operations and long-term strategic growth. The Staff Software Engineer Opportunity We are looking for a Staff Software Engineer to join our growing team in TDI and help scale our internal business solutions with a sharp focus on security, reliability, scalability, and intelligent automation. You will be responsible for designing and developing customization

javascriptpythonjava
View job →
O
Okta
📍 San Francisco• Full-time• From $194K/yr
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The AI-Core Team The team is building a scalable Agentic AI platform for agents that run real engineering work, agents that write code, triage incidents, investigate alerts, remediate vulnerabilities, upgrade libraries, patch flaky tests, automate runbooks, and other workloads across Okta's infra. Owning this platform means owning the hard cross cutting problems: agent identity, delegated access, secure isolated execution, orchestration, and governance, at the scale and reliability bar that Infra demands. The work is novel and high ownership, and the candidate should be drawn to problems the industry hasn't solved yet. What You'll Work On Design and implement backend APIs and services that make up the agentic platform Build the agent identity and machine-to-machine authentication system, including credential management and delegated access flows Build the agent knowledge base and memory layer so agents retain context within and across sessions Buil

typescriptpythonnode.js
View job →
O
Okta
📍 India• Full-time
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Product Auth0 is a developer-friendly identity platform that simplifies authentication and authorization for applications. Designed by developers for developers, we make access to applications safe, secure, and seamless for the more than 100 million daily logins around the world. Our modern approach to identity enables this Tier-Ø global service to deliver convenience, privacy, and security so customers can focus on innovation. Know more about our product at https://auth0.com/ . The Role We are seeking a founding Engineering Manager to lead and bootstrap our newest team: Core Frontier . This team sits at the vital intersection of deep product innovation and the actual customer experience. Your mission is to ensure that the sophisticated features developed across the Core Identity organization (such as Native to Web, Cross-App Access and Custom Token Exchange) are translated into a seamless and intuitive journey for our users. You will act as the champion for a complete and coherent product experience, bridging the gap between complex internal logic and the polished final result our customers interact with every day. As the inaugural manager for Core Frontier, you will lead the hiring and onboarding of a high-caliber team in Bengaluru while establishing the operational rhythms that drive success. You will partner with global engineering leaders to maintain our high standards for security and reliability, ensuring that every release meets the rigor

node.jsawsrest
View job →

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. At Okta, we’re building the future of secure, enterprise-grade AI Agents . We’re looking for a Principal Engineer to join our global AI Engineering team. In this role, you will be instrumental in designing and building the intelligent, user-facing experiences and the end-to-end AI solutions that power them. This is a senior individual contributor role for a hands-on engineer who can set technical direction, mentor others, and partner closely with our cross-geo counterparts as part of one global team. What you'll do : Drive the architecture and design of AI solutions, leading cross-functional initiatives across Product, Design, and Data Science. Design, build, and refine the core backend services that power our AI solutions, including LLM orchestration, RAG pipelines, and generative AI features. Build high-performance Agentic Experiences (AX) for web and mobile, engineered for streaming responses and low latency. Champion observability and operational excellence to ensure our AI services meet enterprise-grade standards for reliability and performance. Develop robust backend services to power our AI solutions, including LLM orchestration, RAG pipelines, and generative AI features Enable the successful delivery of key AI projects through technical leadership and hands-on execution. Mentor engineers, raising the bar on technical craftsmanship and solution quality across the full stack. Collaborate with cross-geo peers to ensure globally aligned designs, shared

typescriptpythonreact
View job →
O
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Staff SET Opportunity We are looking for an experienced Staff Software Engineer in Test to join our Identity Management Engineering (IDM) team serving the Privileged Access Team (PAM). This team is passionate about delivering large-scale, mission-critical software in a fast-paced Agile environment. In this role you'll be working with a team of highly-skilled and talented engineers, responsible for delivering sophisticated backend solutions that help Okta reliably operate at large scale and be highly available. As part of the team, you’ll be ensuring projects are completed with the highest quality and reliability using automation at every level for fast, robust and secure releases. What you’ll be doing Review requirements and design specs to develop relative test plans and test cases Automate API tests, end-to-end tests, reliability/scale tests Work with engineering management to scope and plan engineering efforts Communicate and document QE plans for scrum teams to review Review application code, identify bug and other areas of weakness, architect tools for future coverage Automate all critical features to maintain zero-debt cadence Release features with solid quality Respond to production issues/alerts and customer issues during on-call rotation Be a strong customer advocate with a strong quality DNA What you’ll bring to the role 5+ years of QE experience preferably in an enterprise SaaS company 3+ years experience in quality engineering for ente

pythonjavaaws
View job →
O
Okta
📍 San Francisco• Full-time• From $194K/yr
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Okta is The World’s Identity Company. We free everyone to safely use any technology - anywhere, on any device or app. Our Identity as a Service solution enables secure yet flexible access, authentication, and automation that transforms how people move through the digital world, putting Identity at the heart of business security and growth. At Okta, we celebrate a variety of perspectives and experiences. We are not looking for someone who checks every single box - we’re looking for lifelong learners and people who can make us better with their unique experiences. The Engineering Opportunity Okta's Workforce Identity Cloud (WIC) secures billions of global logins every day. The Resilience Team within WIC maximizes service availability by building foundational, high-scale systems. Our engineers solve complex distributed-systems challenges - from database proxies to traffic shaping - to ensure the platform remains trustworthy and reliable for every customer. We are seeking a Staff Back-End Software Engineer to lead technical initiatives and own the reliability of our expanding Identity as a Service platform. In this role, you will develop microservices, build innovative frameworks, and optimize our cloud-native environment. You will champion reliability during design phases and collaborate across teams to address technical gaps. We value iterative development, automated testing, and individual ownership. Join our creative, fast-paced team to work on impac

javaawsazure
View job →
S
Squarespace
📍 New York City• Full-time• $185.5K – $299K/yr
1mo ago

Squarespace is looking for a Backend Staff Software Engineer to lead the technical direction of our Communications Platform team. The Communications Platform is a critical engine powering customer engagement across Squarespace — providing internal teams with a scalable, reliable, and secure infrastructure for delivering communications via email, push notifications and in-product messaging. The team owns multiple interconnected production platforms that collectively send over 1.75+ million notifications per day, strategic initiatives including platform modernization, In-Product Placements, and multi-channel communication capabilities. In this role, you’ll serve as the senior technical voice for the team — driving architecture, shaping roadmap execution, and elevating the engineering quality of a small but high-impact team. This is a hands-on role with real technical depth, paired with broad cross-functional influence. This is a hybrid role based in our NYC office (3 days per week), reporting to the Engineering Manager of Communications Platform. You’ll Get To… Drive the architecture, design, and implementation of the Communications Platform’s platforms — including Email Delivery, In-Product Notifications and Push Notifications. Define and own the technical strategy for platform modernization, reliability hardening, and engineering standards across backend services. Provide hands-on technical leadership and mentorship to a team of backend and frontend engineers, helping them grow in system design, decision-making, and ownership. Write high-quality Java code and stay close to execution — leading design reviews, code reviews, and architectural decisions that raise the bar across the codebase. Proactively identify and address systemic risks in performance, reliability, security, and operability before they become incidents. Serve as a key technical partner on cross-team and company-wide architecture initiatives. Balance short-term delivery with long-term platform health,

javamongodbgcp
View job →
F
Figma
📍 Ca New York• Full-time• From $185K/yr
1mo ago

Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI. From idea to product, Figma empowers teams to streamline workflows, move faster, and work together in real time from anywhere in the world. If you're excited to shape the future of design and collaboration, join us! The Build Systems team within Figma’s Developer Experience organization owns Figma’s build and CI infrastructure, enabling engineers to ship changes to production quickly and safely. We build and operate core platforms across our polyglot monorepo, including build systems, artifact repositories, merge queues, test frameworks, and CI pipelines. We’re looking for an experienced technical leader to help shape these platforms, uplevel the team, and deliver high-impact platforms that accelerate engineering velocity. The ideal candidate has deep experience with large monolithic codebases, builds durable and scalable systems, and is motivated by solving high-leverage problems that amplify productivity across the engineering organization. This is a full time role that can be held from one of our US hubs or remotely in the United States. What you'll do at Figma: Drive technical roadmap and strategy for the Build Systems team Partner with cross-functional teams and leadership to identify developer pain points and design elegant, scalable solutions Lead complex, multi-quarter initiatives reducing build/test times and improving CI reliability, all while balancing technical excellence with pragmatic delivery Design, build, and maintain modern developer tools including scalable build systems, distributed CI platforms, and test frameworks that serve thousands of engineers Architect and implement large-scale infrastructure on AWS that powers our entire build pipeline to ensure reliability, performance, and cost

typescriptpythonaws
View job →
F
Figma
📍 Ca New York• Full-time• From $153K/yr
1mo ago

Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI. From idea to product, Figma empowers teams to streamline workflows, move faster, and work together in real time from anywhere in the world. If you're excited to shape the future of design and collaboration, join us! The Data Platform team at Figma builds and operates the foundational systems that power analytics, AI/ML, and data-driven decision-making across the company. We serve a diverse set of stakeholders, including AI researchers, machine learning engineers, data scientists, product engineers, and business teams that rely on data for insights and strategy. Our team owns and scales critical platforms such as the Snowflake data warehouse, ML Datalake, orchestration and pipeline infrastructure, and large-scale data ingestion and processing systems, managing all data flowing into and out of these platforms. Despite being a small team, we take on high-scale, high-impact challenges. In the coming years, we're focused on building the data infrastructure layer for Figma's AI-powered products, driving cost and performance optimizations across our data stack, scaling our ingestion and reverse ETL capabilities for new product use cases, and strengthening data quality, reliability, and compliance at every layer. If you're passionate about building scalable, high-performance data platforms that empower teams across Figma, we'd love to hear from you! This is a full-time role that can be held from one of our US hubs or remotely in the United States. What you'll do at Figma: Design and build large-scale distributed data systems that power analytics, AI/ML, and business intelligence across Figma. Develop batch and streaming solutions to ensure data is reliable, efficient, and scalable across the company. Manage and evo

pythonsqlaws
View job →
F
Figma
📍 Ca New York• Full-time• From $153K/yr
1mo ago

Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI. From idea to product, Figma empowers teams to streamline workflows, move faster, and work together in real time from anywhere in the world. If you're excited to shape the future of design and collaboration, join us! Application Platform is part of Product Platform within Figma’s Infrastructure organization. We build shared foundations that help engineers ship backend product work quickly, safely, and reliably. The team owns a large Ruby application that powers Figma’s REST APIs, asynchronous jobs, and workflow orchestration. Our systems sit on critical production paths and shape the day-to-day experience of backend engineers across Figma. We’re looking for an experienced backend engineer with meaningful production Ruby experience who enjoys building for other engineers. You don’t need to be a Ruby language specialist. You should be comfortable making informed tradeoffs in a substantial shared codebase and turning recurring problems into durable systems, tools, and paved paths that improve engineering velocity across the company. This is a full time role that can be held from one of our US hubs or remotely in the United States. What you’ll do at Figma: Design and evolve shared Ruby frameworks for REST APIs, asynchronous jobs, workflow orchestration, and data access Improve the availability, reliability, scalability, and performance of backend systems that support critical product functionality Modernize a large Ruby codebase through safe architectural changes, including moving REST APIs from Sinatra toward Rails and introducing stronger endpoint abstractions and type safety Make backend development faster by improving hot reloading, local workflows, test infrastructure, CI reliability, debugging tools,

awskubernetesrest
View job →
P
1mo ago

ABOUT THE ROLE Mid level Software Engineer will be a member of the agile team which is responsible for design and development of scalable microservices for Peloton's core features across all platforms (Bike, Tread, Strength, and Digital). The Content AI Team is responsible for training and hosting various models in the domains of Natural Language Processing (NLP) and Automatic Speech Recognition (ASR). The candidate will be responsible for developing, testing, deploying, and monitoring microservices that specifically power features like search, voice, and subtitles, which are crucial for platform expansion and international growth. In addition to technical delivery, good communication skills are essential for providing updates to engineering leads and other stakeholders, such as Technical Program Managers (TPM). YOUR DAILY IMPACT AT PELOTON Develop and maintain business-critical APIs and services with a focus on high availability, low latency, security and scalability under guidance from senior engineers Effectively provide updates to the team leaders and participate in sprint planning and team meetings Write understandable, well-tested code with an eye towards maintainability and scalability Contribute to building reusable code, libraries, and patterns for use across teams Implement solutions to scale services while meeting business and product requirements, often under supervision Utilize production monitoring/profiling/tracing and load testing tools to discover bottlenecks and apply techniques such as data modeling, query optimization, and caching to address them Understand and implement industry best practices such as feature toggles, CI/CD, test automation, logging, and monitoring in order to ensure confidence in our release process Participate in on-call rotations and assist with incident response efforts to maintain system reliability YOU BRING TO PELOTON MS or BS degree in fields such as Computer Science, Engineering or Mathematics 3+ years of software devel

pythonsqlpostgresql
View job →
🔔

Get new software reliability engineer jobs by email

Daily job updates · Unsubscribe anytime