Jobiba hiring network

Senior Software Reliability Engineer Jobs

7,292 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current senior software reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

N
Nuro
📍 Mountain View• Full-time• From $160.4K/yr
1mo ago

Who We Are Nuro is a self-driving technology company on a mission to make autonomy accessible to all. Founded in 2016, Nuro is building the world’s most scalable driver, combining cutting-edge AI with automotive-grade hardware. Nuro licenses its core technology, the Nuro Driver™, to support a wide range of applications, from robotaxis and commercial fleets to personally owned vehicles. With technology proven over years of self-driving deployments, Nuro gives the automakers and mobility platforms a clear path to AVs at commercial scale, empowering a safer, richer, and more connected future. About the Team A rider taps "request ride" and within seconds, an autonomous vehicle is matched, dispatched, and on its way. The On-Road Experience team owns the systems that make that moment happen. The On-road Experience team builds the real-time distributed system behind Nuro's ride-hailing product. This is foundational infrastructure where correctness and performance have consequences on a real road. About the Role This is the perfect role if you love architecting performant, reliable distributed systems, processing real-time vehicle data, and building the foundational APIs that enable seamless product experiences. The position demands technical excellence, a deep understanding of system design, and a knack for solving tough problems. Come join us in defining the future! What you'll own Backend services for ride planning, vehicle matching, and real-time trip execution Partner API integrations (Uber and others) routing ride requests to Nuro vehicles Telemetry pipelines surfacing live vehicle state to riders and operations About You 4+ years of backend engineering experience with strong Go proficiency Experience designing and operating distributed systems where reliability isn't optional End-to-end product ownership — from requirements through launch and iteration Solid hands-on experience with GCP (CloudSQL, BigQuery, Redis, or equivalent) A track record of technical lead

sqlredisgcp
View job →
N
Nuro
📍 Mountain View• Full-time• From $160.4K/yr
1mo ago

Who We Are Nuro believes self-driving vehicles are the most immediate and profound opportunity for AI to drive positive change in the physical world. Safer streets, more time for what matters, and easier access to the world around us, that’s why we’re building a universal autonomy platform: self-driving for all roads and all rides. Founded in 2016, Nuro is a physical AI company developing Level 4 autonomous driving technology for a wide range of vehicles, use cases, and markets. Powered by the Nuro Driver™, our universal autonomy platform enables the global mobility ecosystem to deploy autonomy at scale, from robotaxis and logistics fleets to personal vehicles. With years of real-world deployment experience and a flexible, partner-led business model, Nuro is working toward a future where millions of autonomous vehicles powered by our technology help make everyday life safer, easier, and more connected. Nuro has raised over $2B in capital from Uber, NVIDIA, Google, Softbank, Fidelity, T. Rowe Price, and other leading investors. About the Team The Devices Platform team's mandate is to lay the foundation of Nuro's onboard software for our sensor and compute platform, including device drivers, inter-device protocols and pipelines, and device runtime APIs. Sensors and compute hardware are the eyes, ears, and brains of our self-driving robots. We are creating the hardware-agnostic platform to be used by the perception and autonomy SW stack, and to realize the full potential of our sensor and compute HW in reliability, quality, and performance. The projects we work on are high impact and high visibility within Nuro. This team is also responsible for working with internal stakeholders and external suppliers to define, evaluate, integrate the next generation HW platform for Nuro's products and to build the necessary tooling to assist continuous testing and validation. About the Work Design and develop sensor and compute systems for robotics Architect and/or deploy Nuro

linuxmachine learningai
View job →
N
Nuro
📍 Mountain View• Full-time• From $193.9K/yr
1mo ago

Who We Are Nuro believes self-driving vehicles are the most immediate and profound opportunity for AI to drive positive change in the physical world. Safer streets, more time for what matters, and easier access to the world around us, that’s why we’re building a universal autonomy platform: self-driving for all roads and all rides. Founded in 2016, Nuro is a physical AI company developing Level 4 autonomous driving technology for a wide range of vehicles, use cases, and markets. Powered by the Nuro Driver™, our universal autonomy platform enables the global mobility ecosystem to deploy autonomy at scale, from robotaxis and logistics fleets to personal vehicles. With years of real-world deployment experience and a flexible, partner-led business model, Nuro is working toward a future where millions of autonomous vehicles powered by our technology help make everyday life safer, easier, and more connected. Nuro has raised over $2B in capital from Uber, NVIDIA, Google, Softbank, Fidelity, T. Rowe Price, and other leading investors About the Role Nuro takes a machine-learning-first approach to autonomous driving, and the ML Infrastructure team builds and operates the infrastructure that makes that possible. We own the systems that train the models at the core of the Nuro Driver™ - from distributed GPU training and closed-loop reinforcement learning, to the workflows, orchestration, observability, and cost management that keep the fleet running efficiently. Our work sits directly on the critical path of autonomy development. When a training run stalls, when a pipeline silently regresses, or when GPU utilization slips, it shows up in how fast the rest of the company can ship. We care as much about reliability and operational maturity as we do about raw scale. About the Work Contribute to Nuro’s training infrastructure, spanning multi-generation accelerators, and multi-cluster scheduling and orchestration. Design and operate large-scale data pipelines - batch and strea

pythongcpkubernetes
View job →
S
1mo ago

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. Corporate Systems Engineering builds and operates the software platforms, integrations, and automations that power Smartsheet’s core business functions across Finance, Sales/GTM, and People & Culture. Our team owns mission-critical systems and workflows that enable how the company hires, sells, bills, pays, reports, and scales. We operate at the intersection of software engineering, enterprise platforms, and business-critical data, treating internal systems with the same rigor, reliability, and product mindset as customer-facing software. The Automation team builds human-to-system and system-to-system automations that reduce manual effort and friction across the business. We combine cloud-native services, agentic AI, and workflow orchestration to enable employees to interact with enterprise systems through intelligent, secure, and auditable automation. As a Senior Software Engineer I (Automation), you will lead the design, build, and operation of systems and workflows that directly support business execution at scale. You will own complex technical initiatives, partner with Product Managers and stakeholders on technical roadmaps, and mentor junior engineers. This full-time position reports to the Sr. Director, Development and can be located in our Bellevue, WA office, or you may work remotely from anywhere in the US where Smartsheet is a registered employer. You Will: Architect AI Agents: Take a leading role in designing Agentic Workflows using AWS Step Functions and Bedrock Agents that reason

javascripttypescriptpython
View job →
S
1mo ago

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. Corporate Systems Engineering builds and operates the software platforms, integrations, and automations that power Smartsheet’s core business functions across Finance, Sales/GTM, and People & Culture. Our team owns mission-critical systems and workflows that enable how the company hires, sells, bills, pays, reports, and scales. We operate at the intersection of software engineering, enterprise platforms, and business-critical data, treating internal systems with the same rigor, reliability, and product mindset as customer-facing software. The Finance Systems team engineers and operates the platforms that support financial operations, including ERP, procurement, billing, and compliance. We work across configuration, extensibility, and integration to ensure systems are scalable, auditable, and resilient, treating code, configurations, and controls with the same rigor as software. As a Senior Software Engineer I (Finance Systems), you will lead the design, build, and operation of systems and workflows that directly support business execution at scale. You will own complex technical initiatives, partner with Product Managers and stakeholders on technical roadmaps, and mentor junior engineers. You will report into a Manager, Enterprise Systems, and can be based in our Bellevue, WA office, or you may work remotely from anywhere in the US where Smartsheet is a registered employer. You Will: Systems Architecture & Optimization: Engineer and lead the end-to-end lifecycle—analysis, prioritization, and tec

REMOTEvueawsrest
View job →
R
Roblox
📍 San Mateo• Full-time• From $243.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Software Engineer on the Orchestration pod within Engine Productivity, you'll design and run the platform that executes large-scale end-to-end and integration tests, running the real, shipping client on real hardware, across Roblox's data centers, cloud, and our own device labs, so our engineering teams can ship the engine, clients, Studio, and more with speed and confidence. Every Roblox engine, client, and Studio change, along with the experiences built on top of them, should ship with confidence, and the Orchestration team is the layer that makes that possible. We build large-scale distributed services that turn thousands of test suites into a reliable, push-button pipeline: fanning work out across fleets of machines and real devices, moving artifacts to where they're needed, managing single- and multi-client test state, and giving test owners and maintainers a system to validate their own runs. It looks a lot like building a specialized cloud platform, with capacity-aware scheduling, isolation and sandboxing, and smart retry and backoff, plus the classic distributed systems problems (fairness, efficiency, failure handling, and reliability) at Roblox scale. You Will: Design a

pythonawsgit
View job →
R
Roblox
📍 San Mateo• Full-time• From $196.8K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Consumer Apps Consoles team owns the console app foundation that makes Roblox feel fast, fluid, and reliable for millions of players and is dedicated to delivering superior performance, reliability, and user experience on PlayStation and Xbox. As a Senior Software Engineer on the Consoles team, you'll be a driving technical force — leading complex client work end-to-end, raising the quality bar across our codebase, and partnering closely with product, design, and platform teams to ship experiences that are polished and performant at scale. You Will: Shape & improve the PlayStation and Xbox experience that millions of players see every day Translate ambiguous product ideas into well-scoped, high-quality engineering plans in close partnership with product, design, and data science. Raise the engineering bar through rigorous code review, proactive mentorship of junior engineers, and the standards you set in your own code. Drive platform-level investments that benefit the broader engineering organization. Work with our Xbox and PlayStation counterparts to ensure feature parity and platform-specific excellence. You Have: 5+ years of experience building and shipping client software, with

reactawsgit
View job →
R
Roblox
📍 San Mateo• Full-time• From $243.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Software Engineer on our Release Engineering team, you will build the systems that hundreds of engineers across Roblox use every day to safely ship code to tens of millions of concurrent players. Your work will empower teams to ship bold, high-impact changes quickly and confidently on every device Roblox runs on. If you enjoy building developer-facing infrastructure where reliability and blast radius directly impact end-users, you will be right at home on our growing team. You Will: Design and develop backend services and automation that power our release and experimentation process across desktop, mobile, console, VR, and servers Work in C++ engine code to improve telemetry reporting, enhance feature rollout and automatic abort capabilities, and extend release functionality Build progressive client and server rollout, regression detection, and automated rollback systems that keep our weekly multi-platform releases safe at scale Work directly with engineering customers to turn pain points into durable, flexible, and safe-by-default tooling You Have: 5+ years building backend services with C#, Python, TypeScript, or similar Familiar with and comfortable working with C++ Familiar

typescriptpythonaws
View job →
R
Roblox
📍 San Mateo• Full-time• From $243.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Roblox's data infrastructure processes petabytes of data daily, powering analytics, ML, and product decisions. As a Senior Software Engineer in our Data Infra org, you will design, build, and scale the distributed data infrastructure platforms that power Roblox. You will own and drive the next-generation architecture of our core platforms, which span Kafka, Flink, Spark, Trino, Druid, Airflow and Data Catalog. This role combines high ambiguity and ownership to push the boundaries of what our infrastructure can handle at massive scale, giving you the unique opportunity to steer the evolution of the data landscape. You Will: Own and Scale Core Platform Components: Take responsibility for the design, architecture, and implementation of 1–2 key data platform frameworks within our stack Collaborate and Align: Partner with infra, data science, and product engineering teams to ensure your target platform's capabilities are directly guided by platform governance and product requirements. Optimize Performance at Scale: Dive deep into engine internals, query planning, state management, memory optimization, serialization efficiency to maximize throughput and reliability under heavy load. Drive Infrast

javaawsgcp
View job →
T
Twilio
📍 - Ireland• Full-time• Remote
1mo ago

Who we are At Twilio, we’re shaping the future of communications, all from the comfort of our homes. We deliver innovative solutions to hundreds of thousands of businesses and empower millions of developers worldwide to craft personalized customer experiences. Our dedication to remote-first work , and strong culture of connection and global inclusion means that no matter your location, you’re part of a vibrant team with diverse experiences making a global impact each day. As we continue to revolutionize how the world interacts, we’re acquiring new skills and experiences that make work feel truly rewarding. Your career at Twilio is in your hands. . Hiring and how we work We use Artificial Intelligence (AI) to help make our hiring process efficient. That said, every hiring decision is made by real Twilions! Also, while we are a remote-first company, you may be asked to report in person on an ad-hoc basis for team gatherings, functional off-sites or customer meetings. . See yourself at Twilio Join the team as Twilio’s next Senior Software Engineer. About the job This position is needed to design, build, and optimize the core signalling infrastructure that powers real-time video communications for our customers. You will play a key role in ensuring high performance, reliability, and scalability of our video platform, enabling seamless and secure video experiences. Responsibilities In this role, you’ll: Design, implement, and maintain video signalling protocols and server components for real-time video calls (e.g., WebRTC, SIP, RTCP/RTP) in a highly scalable distributed system. Collaborate with cross-functional distributed teams and various stakeholders to deliver high-performance, low-latency media experiences. Ensure secure transmission and compliance with industry best practices (e.g., end-to-end encryption, privacy standards). Contribute to architectural decisions and code reviews, mentoring junior engineers as needed. Stay current with

REMOTEjavaawsazure
View job →
D
Datadog
📍 Spain; Paris, France• Full-time
1mo ago

We are looking for a Senior Software Engineer to help us take REDAPL, our Referential Data Platform, to the next level. REDAPL is Datadog's main platform for tracking our customers' infrastructure resources and relationships. The platform enables products where customers can understand, keep track of, and gain insights into their infrastructure related to performance, cost, security, and more. Many Datadog products use REDAPL today such Cloud Security Posture Management, Resource Catalog, Cloud Cost Management, and Service Catalog and others - REDAPL ingests more than 4.5mil updates/second. As a Senior Engineer, you will drive, lead and collaborate on projects both inside and outside the platform. You can expect to contribute to key technical decisions relating to our data ingestion, processing, and query pipelines. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Build a query engine that supports efficient relationship traversals for our most demanding workloads. Contribute to design and drive high-priority, high-visibility projects to increase the platform's value, resilience, and scalability across multiple teams. Lead and guide other engineers through architectural platform decisions Identify potential system risks and trends in reliability and design solutions to address them Provide input on prioritizing engineering-led initiatives in short- and long-term planning and roadmaps Collaborate with internal product teams to understand their requirements and how we plan for their product growth as they integrate and depend on REDAPL Who You Are: You have a BS/MS/PhD in a Computer Science, Engineering or related scientific field or equivalent experience You have worked extensively with multiple types of data stores You have contributed to in

aigorust
View job →
D
Datadog
📍 France; Sophia Antipolis, France• Full-time
1mo ago

This role will join Datadog’s Data Visualization organization, a team responsible for the visualization experiences that power dashboards, notebooks, investigations, and product workflows used across the platform. The team is a highly product-oriented organization, building AI-native experiences that help customers understand, investigate, and interact with complex operational data. As a Staff Software Engineer, you will provide technical leadership in applying AI technologies to customer-facing product experiences, helping shape how users interact with Datadog through agents, conversational interfaces, and intelligent investigation workflows. You will partner across engineering and product teams to develop reliable, scalable, and trustworthy AI-powered experiences while helping establish AI engineering expertise within the broader organization. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead the design and delivery of AI-powered product experiences across Datadog’s visualization and investigation surfaces. Develop systems that combine deterministic product capabilities with LLM-powered experiences to deliver trustworthy and explainable customer outcomes. Drive innovation in context engineering, prompt engineering, evaluation frameworks, and AI application reliability. Partner with product and engineering teams to improve investigation workflows and help customers discover insights more efficiently. Build experiences that enable Datadog capabilities to operate within third-party AI platforms, agents, and conversational environments. Provide technical leadership and mentorship while helping establish AI engineering best practices across the Data Visualization organization and broader Graphing group. Who You Are: You have extensive softw

aigorust
View job →
D
Datadog
📍 Remote; United Kingdom, Remote• Full-time• Remote
1mo ago

Please note that the job is only available from the locations outlined. We are looking for a Senior Software Engineer to help us take REDAPL, our Referential Data Platform, to the next level. REDAPL is Datadog's main platform for tracking our customers' infrastructure resources and relationships. The platform enables products where customers can understand, keep track of, and gain insights into their infrastructure related to performance, cost, security, and more. Many Datadog products use REDAPL today such Cloud Security Posture Management, Resource Catalog, Cloud Cost Management, and Service Catalog and others - REDAPL ingests more than 4.5mil updates/second. As a Senior Engineer, you will drive, lead and collaborate on projects both inside and outside the platform. You can expect to contribute to key technical decisions relating to our data ingestion, processing, and query pipelines. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Build a query engine that supports efficient relationship traversals for our most demanding workloads. Contribute to design and drive high-priority, high-visibility projects to increase the platform's value, resilience, and scalability across multiple teams. Lead and guide other engineers through architectural platform decisions Identify potential system risks and trends in reliability and design solutions to address them Provide input on prioritizing engineering-led initiatives in short- and long-term planning and roadmaps Collaborate with internal product teams to understand their requirements and how we plan for their product growth as they integrate and depend on REDAPL Who You Are: You have a BS/MS/PhD in a Computer Science, Engineering or related scientific field or equivalent experience You

REMOTEaigorust
View job →

We are looking for a Senior Software Engineer to help us take REDAPL, our Referential Data Platform, to the next level. REDAPL is Datadog's main platform for tracking our customers' infrastructure resources and relationships. The platform enables products where customers can understand, keep track of, and gain insights into their infrastructure related to performance, cost, security, and more. Many Datadog products use REDAPL today such Cloud Security Posture Management, Resource Catalog, Cloud Cost Management, and Service Catalog and others - REDAPL ingests more than 4.5mil updates/second. As a Senior Engineer, you will drive, lead and collaborate on projects both inside and outside the platform. You can expect to contribute to key technical decisions relating to our data ingestion, processing, and query pipelines. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Build a query engine that supports efficient relationship traversals for our most demanding workloads. Contribute to design and drive high-priority, high-visibility projects to increase the platform's value, resilience, and scalability across multiple teams. Lead and guide other engineers through architectural platform decisions Identify potential system risks and trends in reliability and design solutions to address them Provide input on prioritizing engineering-led initiatives in short- and long-term planning and roadmaps Collaborate with internal product teams to understand their requirements and how we plan for their product growth as they integrate and depend on REDAPL Who You Are: You have a BS/MS/PhD in a Computer Science, Engineering or related scientific field or equivalent experience You have worked extensively with multiple types of data stores. You have contribute

aigorust
View job →
M
Mongodb
📍 San Francisco• Full-time• From $126K/yr
1mo ago

Join the Atlas Search team to design and develop the next generation of Semantic and Vector Search infrastructure. Atlas Search is a growing cloud service that allows users to execute complex search and vector search queries using the MongoDB Query Language. Our users can focus on relevance and data retrieval instead of the machinery needed to search data at scale. Our team is building a cloud-based distributed system responsible for the core components of search including data ingestion, performance, query language, query execution, for both relevance-based search and vector search. Our product is being adopted quickly and there are many interesting projects. This is a technical role where you will be responsible for the infrastructure and features enabling our at-scale cloud service powering vector and semantic search. We are looking to speak to candidates who are based in the San Francisco Bay Area for our hybrid working model. What You’ll Do Lead complex projects across the MongoDB ecosystem, for instance, development of a new Search deployment framework within the MongoDB managed cloud Set project level strategy, architect features, and lead projects to successful execution Identify, design, and implement features enhancing our reliability, performance, security and efficiency Perform code reviews with peers and make recommendations on how to improve our software development processes Influence and grow team members through active mentoring and leading by example What We Look For 5+ years experience in data management/search systems, ideally with a strong distributed systems and infrastructure background Experienced in the development and maintenance of concurrent, stateful services Eager to shape the technological direction of a complex system and have the ability to lead initiatives through collaboration with others Experienced in writing features, debugging and optimizing multithreaded applications written in Java Familiarity with LLM

javamongodbaws
View job →
🔔

Get new senior software reliability engineer jobs by email

Daily job updates · Unsubscribe anytime