Jobiba hiring network

Software Reliability Engineer Jobs

6,326 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current software reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

I
1mo ago

We're transforming the grocery industry At Instacart, we invite the world to share love through food because we believe everyone should have access to the food they love and more time to enjoy it together. Where others see a simple need for grocery delivery, we see exciting complexity and endless opportunity to serve the varied needs of our community. We work to deliver an essential service that customers rely on to get their groceries and household goods, while also offering safe and flexible earnings opportunities to Instacart Personal Shoppers. Instacart has become a lifeline for millions of people, and we’re building the team to help push our shopping cart forward. If you’re ready to do the best work of your life, come join our table. Instacart is a Flex First team There’s no one-size fits all approach to how we do our best work. Our employees have the flexibility to choose where they do their best work—whether it’s from home, an office, or your favorite coffee shop—while staying connected and building community through regular in-person events. Learn more about our flexible approach to where we work. Senior Software Engineer, Core Experience Job #SECE39 Overview Maplebear Inc. D/B/A Instacart in San Francisco, CA seeks Senior Software Engineer, Core Experience (Multiple Openings). Employee may be stationed anywhere in the continental United States but will report directly to Instacart headquarters in San Francisco, CA. About the Job Design, build, and implement software solutions to optimize and scale the Instacart platform and related technologies as part of the Core Experience team. Duties include: Developing highly scalable software using innovative computer science and software engineering principles while ensuring proper protocols are in place to rapidly roll out upgrades and new features; Analyzing and optimizing expensive SQL queries; Enabling sophisticated, data-driven features that adapt to customer shopping behavio

REMOTEpythonjavasql
View job →

We are the GPU Communications Libraries and Networking team at NVIDIA. We deliver communication libraries like NCCL, NVSHMEM, UCX for Deep Learning and HPC. DL and HPC applications have a huge compute demand already and run on scales which go up to tens of thousands of GPUs. The GPUs are connected with high-speed interconnects (eg. NVLink, PCIe) within a node and with high-speed networking (eg. Infiniband, Ethernet) across the nodes. Communication performance between the GPUs has a direct impact on the end-to-end application performance; and the stakes are even higher at huge scales! We are looking for a technical leader to manage our NVSHMEM and UCX libraries. This is an outstanding opportunity to push the limits on the state-of-the-art and deliver platforms the world has never seen before. Are you ready for to contribute to the development of innovative technologies and help realize NVIDIA's vision? What you will be doing: Lead, mentor, and grow your library engineering team and be responsible for the planning and execution of projects as well as the quality, and performance of your libraries. This is a technical leadership role so you will participate in feature design and implementation. Interact with internal and external partners and researchers to understand their use cases and requirements. Collaborate with engineering teams, program and product management, and partners to define the product roadmap. Continuously review and identify improvement opportunities in established processes, infrastructure, and practices to ensure the teams are executing in the most efficient and transparent manner. What we need to see: 10+ overall years of experience in the software industry with specialization in HPC networking or system software. 4+ years of management experience. BS, MS, or Ph.D. in C

V
Vanta
📍 Toronto• Full-time• Remote
1mo ago

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. As a Staff Software Engineer on the Automation - Foundations team, you will lead the re-platforming of the data layer underneath Vanta’s entire compliance product. This is a migration that spans multiple teams, has to preserve every public API contract along the way, and cannot lose a single customer’s evidence while it happens. Foundations is Vanta’s data platform team. We ingest security and compliance data from across our customer’s environments, currently tens of thousands of resources per second, with single-customer bursts running into the millions. We store the data, catalog it, make it queryable, and turn it into evidence that has to survive a real SOC 2 or FedRAMP audit. We are in the middle of moving our platform from a Mongo-centric architecture to a schema-aware, Postgres-backed architecture, on a Kafka and S3 pipeline that decouples data fetching from processing. Both pipelines run in parallel today, the hard problems here are correctness under migration, eventual consistency, and multi-tenancy, in a domain where “mostly right” is not an acceptable failure mode. Visit our Vanta Engineering Blog to learn more about what our team is working on. What you'll do as a Staff Software Engineer at Vanta: Lead the migration of Vanta’s resource data model from a Mongo-centric solution to a schema-aware Postgres-backed solution and running both generations in parallel without breaking a customer integration. Drive solutions across teams that you do not own but are dependent on the platform built by your team. Design for correctness under eventual consistency with idempotent session handling, conditional writes that survive out

REMOTEtypescriptmongodbredis
View job →
V
Vanta
📍 United States• Full-time• Remote
1mo ago

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. As a Staff Software Engineer on the Automation - Foundations team, you will lead the re-platforming of the data layer underneath Vanta’s entire compliance product. This is a migration that spans multiple teams, has to preserve every public API contract along the way, and cannot lose a single customer’s evidence while it happens. Foundations is Vanta’s data platform team. We ingest security and compliance data from across our customer’s environments, currently tens of thousands of resources per second, with single-customer bursts running into the millions. We store the data, catalog it, make it queryable, and turn it into evidence that has to survive a real SOC 2 or FedRAMP audit. We are in the middle of moving our platform from a Mongo-centric architecture to a schema-aware, Postgres-backed architecture, on a Kafka and S3 pipeline that decouples data fetching from processing. Both pipelines run in parallel today, the hard problems here are correctness under migration, eventual consistency, and multi-tenancy, in a domain where “mostly right” is not an acceptable failure mode. Visit our Vanta Engineering Blog to learn more about what our team is working on. What you'll do as a Staff Software Engineer at Vanta: Lead the migration of Vanta’s resource data model from a Mongo-centric solution to a schema-aware Postgres-backed solution and running both generations in parallel without breaking a customer integration. Drive solutions across teams that you do not own but are dependent on the platform built by your team. Design for correctness under eventual consistency with idempotent session handling, conditional writes that survive out

REMOTEtypescriptmongodbredis
View job →
O
OpenAI
📍 San Francisco• Full-time• Remote
1mo ago

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role You will build the low-level device runtime that turns compiled programs into efficient, functional and performant execution on OpenAI’s custom AI accelerator. This software will schedule kernel launches, manage device memory and address spaces, coordinate synchronization, and expose reliable abstractions to higher-level runtimes and frameworks. You will work at the boundary of software and hardware, partnering with compiler, kernel, architecture, verification, and silicon teams to define interfaces and validate behavior. You will also use and improve event-based, cycle-accurate simulation to develop runtime capabilities before silicon is available, diagnose performance and correctness issues, and guide hardware-software co-design. In this role, you will: Design and implement the low-level device runtime for OpenAI custom silicon. Build kernel-launch scheduling, command submission, queueing, dependency tracking, and completion handling. Manage device memory spaces, allocation, virtual-to-physical mappings, data movement, and lifetime across concurrent workloads. Implement synchronization primitives, events, barriers, streams, and ordering guarantees that are correct and efficient. Define clean interfaces between the runtime, drivers, firmware, compiler-generated code, kernels, and higher-level execution systems. Use event-based, cycle-accurate simulators to develop, validate, debug, and performance-tune runtime behavior before and after silicon availability. Di

REMOTEawsrestai
View job →
R
Roblox
📍 San Mateo• Full-time• From $243.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Backend Engineer on the User Communities team, you’ll own the backend systems that power social engagement features (Announcements, Forums, Polls, Roles & Permissions systems) for millions of players and creators on Roblox. Your day-to-day will focus on building and scaling the infrastructure that makes both in-experience and on-platform communication and player engagement fast, safe, and reliable. That means driving architectural decisions, reasoning through trade-offs, influencing stakeholders, championing user-first safety standards, and defining developer-facing APIs so creators can build richer experiences. You'll work closely with frontend and backend engineers, product, design, and data scientists across teams, and you'll have real influence over our technical and product direction. If you're an experienced engineer who gets excited about large-scale distributed systems and wants to shape the way millions of people connect inside virtual worlds, we'd love to talk. You Will: Shape the backend engineering culture across the Communities team. Your influence extends beyond your immediate pod to improve how we build, review, and ship software across the entire org. Act as

javaawsgit
View job →
R
Roblox
📍 San Mateo• Full-time• From $243.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Software Engineer within the Creator Organization, you will develop innovative full-stack solutions that define the future of Roblox’s Content Understanding Platform. This platform processes billions of pieces of content—spanning 3D models, audio files, text, video, and entire experiences—extracting structured information about their meaning, context, and relationships. The Content Understanding Team develops cutting-edge AI models, advanced computer vision systems, and highly scalable backend platforms to power search, discovery, and moderation across Roblox. Your contributions will enable seamless asset discovery, automate moderation at scale, and drive transformative generative AI tools that reshape how millions of creators and users engage with Roblox. You Will: Solve full-stack challenges to improve how AI, creators, and users describe and get along with content, including images, 3D models, audio, text, and video. Craft and build scalable pipelines for training, evaluating, and deploying machine learning models to support content annotation and discovery. Develop robust backend systems to power real-time search, discovery, and powerful generative AI features. Blend innovat

pythonjavareact
View job →
L
Lyft
📍 Toronto• Full-time• From C$108K/yr
1mo ago

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Marketplace teams are at the heart of our products and decision-making, owning everything from rider pricing to driver earnings, incentives, and efficient matching. We’re looking for passionate, driven engineers to build systems that empower our riders and drivers to have the best transportation experience possible through prediction, adaptivity, and personalization. We’re looking for someone who is excited about working in a fast-paced, innovative, and impactful environment to create reliable solutions to distributed computing, ML, and data problems. The Pricing team is a centerpiece of Lyft’s Marketplace org, determining prices for all rideshare products and supporting new initiatives. Rider Engagement develops rider-facing engagement levers and optimizes user pricing experience to drive both short term and long term business outcomes. We work with Product & Science to solve and implement complex pricing requirements, balancing the needs of riders, drivers, and the business goals. As an owner of one of the most critical flows in the company, you will work on a wide array of challenges such as latency-sensitive concurrency problems, large scale distributed systems, and experimentation. If you’re interested in playing a large part in demand / supply management and improving the Lyft customer experience, this could be a great fit for you. Responsibilities: Drive high-impact projects and innovate new solutions to provide the best user experience. Work closely with cross-functional teams and partner teams to develop solutions based on technology and business needs, and advance team’s goals and priorities Independently lead features from idea to positive execution and launch Unblock, support and communicate with internal partners to achieve results Write well-crafted, well-tested, readable, maintaina

pythonawsrest
View job →
S
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies - from the world’s largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team Stripe Terminal helps Stripe users extend their online presence into the physical world. The Terminal team’s mission is to make it as easy for businesses to accept in-person payments as the Stripe API has done for online payments. With Terminal, businesses can unlock in-person payments use cases that are right for their business model, whether it’s creating a flagship retail experience, extending their website to a pop-up store, or enabling a mobile point of sale at their next event. What you’ll do We are looking for a Software Quality Assurance Engineer to join the Stripe Terminal Device Software Quality Assurance team. In this role, you will own the quality assurance process and results for firmware and device operating system components across our payment terminal devices. You'll work closely with software development teams to define quality standards, develop test strategies, and drive software quality improvements through technical expertise and process enhancements. Responsibilities Own software quality for Terminal device software components, including firmware and the device operating system, and ensure software quality targets are met for customers. Collaborate with developers during software planning, change review, and release meetings to understand new features, specifications, expected behaviors, and the impact and risks of changes. Create spot tests before branch cut. Deeply understand customer use cases, Produ

aiexcelhr
View job →
P
1mo ago

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . About tvScientific tvScientific is the first and only CTV advertising platform purpose-built for performance marketers. We leverage massive data and cutting-edge science to automate and optimize TV advertising to drive business outcomes. Our solution combines media buying, optimization, measurement, and attribution in one, efficient platform. Our platform is built by industry leaders with a long history in programmatic advertising, digital media, and ad verification who have now purpose-built a CTV performance platform advertisers can trust to grow their business. As a Data Engineer at tvScientific, you will be a key player in implementing the robust data infrastructure to power our data-heavy company. You will collaborate with our cross-functional teams to evolve our core data pipelines, design for efficiency as we scale, and store data

REMOTEsqlawsmachine learning
View job →

NVIDIA is the world leader in GPU Computing. We are passionate about markets including gaming, automotive, professional vision, HPC, datacenters and networking in addition to our traditional OEM business. NVIDIA is also well positioned as the ‘AI Computing Company’, and NVIDIA GPUs are the brains powering modern Deep Learning software frameworks, accelerated analytics, modern data centers, and driving autonomous vehicles. We have some of the most experienced and dedicated people in the world working for us. If you are dedicated, forward-thinking, and if working with hard-working technical people across countries sounds exciting, this job is for you. We are now looking for a Software QA Test Development Engineer, you will collaborate with multi-functional groups. SWQA test developer engineer at NVIDIA is responsible for test planning, execution, and reporting, you will also write scripts to automate testing, design and develop tools for QA team, or develop integration tests for validation, so QA engineer can improve productivity or optimize test plan. As a SWQA test developer, you must identify weak spots and constantly design better and creative test plans to break software and identify potential issues. You will have a huge impact on the quality of NVIDIA's products. What You’ll Be Doing Analyze requirements and design test matrices covering functionality, performance, and edge cases. Develop test plans and cases; build and maintain automated test suites (API / UI / CLI / E2E) in Python. Leverage AI-powered tools and agentic workflows to accelerate test generation, triage, and root cause analysis. Manage the full bug lifecycle — filing, reproduction, and driving multi-functional collaboration to resolution. Reproduce and verify customer-reported issues to ensure quality before release.

pythondockerkubernetes
View job →
🔔

Get new software reliability engineer jobs by email

Daily job updates · Unsubscribe anytime