#TeamNextdoor Nextdoor is where you connect to the neighborhoods that matter to you so you can belong. Our purpose is to cultivate a kinder world where everyone has a neighborhood they can rely on. Neighbors around the world turn to Nextdoor daily to receive trusted information, give and get help, get things done, and build real-world connections with those nearby — neighbors, businesses, and public services. Today, neighbors rely on Nextdoor in more than 300,000 neighborhoods across 11 countries. Meet Your Future Neighbors As a Software Engineer at Nextdoor, you’ll join a focused team of developers, product managers, and designers who are passionate about using technology to cultivate a kinder world where everyone has a neighbor they can rely on. We are a small team of engineers that wear multiple hats and work across different languages and services to deliver value to our members. We care about moving fast and delivering impact, without compromising on quality and reliability. You will have the opportunity to learn from your co-workers and teach them. As a team, we will make each other better and build great software. What You’ll Bring To The Team If you didn't see an opportunity listed that looked right for you, we would still love to hear from you and consider you for other opportunities! At Nextdoor, we empower our employees to build stronger local communities. To create a platform where all feel welcome, we want our workforce to reflect the diversity of the neighbors we seek to serve. We encourage everyone interested in our purpose to apply. We do not discriminate on the basis of race, gender, religion, sexual orientation, age, or any other trait that unfairly targets a group of people. In accordance with the San Francisco Fair Chance Ordinance, we always consider qualified applicants with arrest and conviction records. #LI-DNI
Jobiba hiring network
Software Reliability Engineer Jobs
6,326 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current software reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Join our Modern Virtualization team to engineer the next generation of enterprise storage for cloud-native and virtualized environments. You will architect and deploy best-in-industry Kubernetes, Nutanix, and OpenStack integrations, bringing enterprise-grade storage mobility and automation directly to our customers’ mission-critical workloads. You’ll collaborate closely with global engineering leads to design robust, production-ready code that defines the future of our storage ecosystem and ensures seamless platform interoperability. WHAT YOU'LL DO Feature Architecture & Development: Architect, develop, and deliver scalable features across Kubernetes (OpenShift), OpenStack, and Nutanix stacks for FlashArray and FlashBlade, ensuring seamless, enterprise-grade storage performance for virtualized platforms. End-to-End Ownership: Own the full lifecycle of your features, from initial design and implementation to writing rigorous test automation, guaranteeing high reliability and stability for mission-critical production workloads. Cross-Functional Collaboration: Partner with engineering teams across geographies to drive design reviews, troubleshoot complex distributed system challenges, and continuously elevate our engineering practices, code quality, and test coverage. WHAT YOU'LL BRING Core Engineering Proficiency: Deep hands-on expertise in building production-grade software using Go, Java, or Python, with a pro
Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. You will own end-to-end delivery for complex, cross-org modernization programs. You will create clarity from ambiguity, define operating mechanisms, drive dependency management, and ensure high-quality execution across Platform Engineering and Low Code Foundations. Required Experience & Skills 8–12+ years technical program management in engineering organizations Experience driving large, cross-team technical programs with measurable outcomes Excellent written and verbal communication; strong ability to influence without authority Demonstrated rigor in planning, risk management, and execution mechanisms Comfort partnering with senior engineering leaders and principal engineers Preferred Experience & Skills Experience with platform modernization (cloud-native, performance, reliability, migrations) Experience with AI-enabled programs or AI-first SDLC adoption Background in software engineering, systems engineering, or technical product delivery Tools and Resources Training and Development: During onboarding, we focus on equipping new hires with the skills and knowledge for success through department-specific training. Continuous learning is a central focus at Appian, with dedicated mentorship and the First-Friend program being widely utilized resources for new hires. Growth Opportunities: Appian provides a diverse array of growth and development opportunities, including our leadership program tailored for new and aspiring managers, a comprehensive library of specialized department training through Appian University, skills based training,
About the Team OpenAI’s Hardware organization develops silicon and system-level solutions designed for the unique demands of advanced AI workloads. The team builds next-generation AI-native silicon and systems while working closely with software, research, and manufacturing partners to co-design hardware tightly integrated with AI models. In addition to delivering systems for OpenAI’s supercomputing infrastructure, the team develops the tools, methodologies, and strategic partnerships needed to accelerate hardware innovation. About the Role We’re seeking an experienced Hardware Strategic Sourcing Manager to own sourcing strategy and supplier partnerships for fiber and optical interconnect components across OpenAI’s next-generation AI infrastructure. Reporting to the Head of Partnerships & Strategic Sourcing, you will lead sourcing across fiber cable assemblies, internal optical harnesses, fiber shuffles, optical backplane assemblies, connectorized and standalone passive optical assemblies, fiber-array units (FAUs), fiber-to-chip and coupling interfaces, detachable connectors, optical routing, and assigned optical packaging, assembly, and test services. You will work closely with electrical engineering, optical engineering, systems engineering, mechanical and packaging engineering, quality, rack integration, data-center deployment,manufacturing, supply chain, finance, legal, and program management teams to translate demanding bandwidth, signal integrity, reliability, and scale requirements into resilient supplier partnerships and scalable commercial strategies. Your work will directly support the performance, reliability, manufacturability, and scale of the high-speed optical connectivity required for OpenAI’s next-generation AI systems. In this role, you will: Develop and execute a comprehensive sourcing strategy for fiber and optical interconnect components supporting high-bandwidth AI systems and infrastructure. Own sourcing across optical fiber cable assembli
About the Team OpenAI Consumer Devices is building the next generation of products that bring powerful AI into people’s everyday lives. Guided by OpenAI’s mission to ensure AGI benefits all of humanity, our team combines world-class researchers, engineers, designers, and operators who care deeply about creating useful, intuitive, and responsible technology. You’ll have the opportunity to work alongside exceptional people on ambitious, zero-to-one challenges at the intersection of hardware, software, and AI. This is a chance to help define an entirely new category of products—and shape how people experience AI in the future. The Systems Integration team is critical in this mission, turning complex hardware-software development into reliable product signals. We build the shared infrastructure, tooling, and lab environments that let teams test quickly, understand failures, and ship with confidence. About the Role As a Systems Integration Manager , you will lead the team responsible for device validation infrastructure, test automation, developer tooling, and lab operations. This is a player-coach leadership role: you’ll set technical and operational direction, build and develop a team of engineers and lab operations professionals, and stay close to the architecture and hardest systems problems. You will partner closely with device software, OS, firmware, hardware, reliability, QA, and release infrastructure teams to define validation strategy, improve release readiness, and ensure our test environments and quality signals scale with the product. Because this is a new category of devices, you’ll have the rare opportunity to build the validation foundation early—shaping the systems, standards, and operating model that will support products from prototype through launch. We’re looking for a leader who combines strong technical judgment with people leadership, operational rigor, and experience building reliable systems for complex hardware-software products. This role is b
Role Purpose: At Jumio, the Software Engineer II (QA) will focus on ensuring the quality and performance of highly scalable web and backend applications. In this role, you will design and implement automated tests for web (Playwright/Selenium) and API-based solutions, leveraging your knowledge of Java or JavaScript. Collaborating closely with development and product teams, you will ensure Jumio's products meet the highest standards of quality and reliability. You’ll have an opportunity to learn and grow in a fast-paced environment while exploring new tools and methodologies. A problem-solving mindset, willingness to innovate, and attention to detail will make you successful in this role. T-Shaped Engineering Expectation: As part of Jumio’s engineering culture, you will adopt a T-shaped engineering approach. In addition to developing expertise in test automation and quality engineering, you will collaborate across the development lifecycle, including understanding software architecture, contributing to design discussions, and ensuring robust and scalable test solutions. Role Value: This role is critical to ensuring the reliability, scalability, and security of Jumio’s products. By building and maintaining automated testing frameworks, you will enable faster releases and higher confidence in the quality of our software. Example Responsibilities: Develop, maintain, and execute automated test scripts for web applications using Playwright, Selenium, or similar automation frameworks. Create and execute API test suites using tools such as Postman, REST Assured, or equivalent testing frameworks. Design and execute functional, regression, integration, and exploratory test cases based on business and technical requirements. Validate application functionality, backend services, APIs, and data flows across different environments. Identify, document, track, and verify defects, working closely with developers to ensure timely resolution. Execute automated test suites as part of C
Graphcore Senior Principal AI SoC Validation (Bring-up lead) Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. We are opening a new AI Engineering Campus in Bengaluru which will play a central role in Graphcore's work building the future of AI computing. We are developing the next generation of AI compute, a large-scale system-on-chip (SoC) designed to power future high-performance AI systems. As the SoC Validation Lead, you will be responsible for enabling pre-production software to run reliably on new silicon quickly and efficiently, before showing that the silicon meets the highest standards of quality, reliability and functionality, ready for production deployment. You will lead a team delivering post-silicon validation across the full AI SoC, working across silicon, firmware, and platform levels. The role requires a deep technical understanding, strong hands-on debug experience, and the ability to collaborate effectively with hardware, software, and systems engineering teams. Key responsibilities Define and lead post-silicon validation strategy Develop and refine the overall post-silicon validation approach for our AI SoCs, ensuring reliable and timely delivery of validated silicon, architectural correctness, feature robustness, and at-scale system reliability. Drive cross-domain debug and issue resolution Lead investigation and resolution of complex issues spanning silicon, firmware, operating systems, and platform interactions. Ensure that fixes are effective and sustainable. Promote collaboration and shared understanding Work closely with
We are a team of engineers that translate our real-world experience to help our user communities solve problems. With a focus on AI-accelerated workflows and next-generation developer ecosystems, you will have the opportunity to meet fast-moving teams where they are, helping them lay strong engineering foundations and broadening your impact to the developer community at large. This is a unique opportunity to use both your engineering depth and authentic storytelling skills to shape how the next wave of builders approach software health, scalability, and observability. What You’ll Do: Help developers hone their craft in an AI-accelerated world by exploring how AI coding assistants and rapid-prototyping tools change software workflows, and guiding teams on how to balance rapid prototyping with established engineering practices around performance, code health, and system reliability. Build in public by creating authentic, peer-to-peer technical content, sharing your engineering insights directly where modern developers collaborate in person and online. Drive a constructive, bidirectional feedback loop between fast-moving developer communities and our internal teams, translating real-world developer experiences into actionable product insights to shape our roadmap while advocating for user needs from the inside. Design and ship high-quality open-source boilerplate templates, quickstarts, and tools that make integrating observability seamless across AI native workflows, next-gen platforms (Vercel, Supabase, etc.), and cloud providers (AWS, GCP, Azure). Who You Are: A builder with approximately 10+ years of software engineering experience, ideally having shipped applications from scratch or worked within an early-stage startup environment where you've worn multiple hats. An AI-fluent developer who natively leverages AI coding assistants and rapid-prototyping tools as a core part of your regular, day-to-day development workflow. Multi-stack familiar, comfortable writing an
About the Team OpenAI’s Industrial Compute team is building and productizing infrastructure capabilities that help organizations deploy and operate advanced AI systems at scale. The team works across AI hardware, systems engineering, physical infrastructure, and customer delivery to turn emerging technologies into reliable, repeatable infrastructure solutions. Our work sits at the intersection of technical strategy, product development, engineering, and deployment. We partner closely with customers and internal engineering teams to solve complex infrastructure challenges spanning compute, power, cooling, controls, and facility efficiency. About the Role We are seeking a senior, hands-on Data Center Infrastructure Architect to develop and optimize the physical infrastructure required for large-scale AI deployments. This is a broad technical role spanning data center architecture, electrical and mechanical systems, high-density compute, controls, telemetry, and digital modeling. You will use simulation, operational data, and digital-twin approaches to evaluate infrastructure designs, identify system-level constraints, and improve efficiency, reliability, cost, and speed of deployment. The ideal candidate can move fluidly between first-principles analysis, facility and equipment design, computational modeling, engineering review, and real-world implementation. You should be comfortable working across disciplines rather than operating solely within electrical, mechanical, or software boundaries. Key Responsibilities Define system-level architectures for high-density AI data centers across power, cooling, IT equipment, controls, and facility infrastructure. Develop digital twins and other computational models that represent the behavior of data center systems under changing workloads, environmental conditions, equipment configurations, and failure scenarios. Use design and operational data to identify constraints, improve PUE and related efficiency metrics, and optimize
Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. This position is based at our office in Chennai, India. Appian was built on a culture of in-person collaboration, which we believe is a key driver of our mission to be the best. You will be the product manager working closely with the team whose mission is to strengthen and optimize site infrastructure by delivering essential upgrades, resource efficiency, and scalable solutions. You will be responsible for the direction and roadmap of a component of the Appian Cloud data plane that ensures reliable, high-performance operations for all Appian Cloud customer sites. This role is specifically focused on the cloud-native persistence and messaging layer. You will oversee the backend sub-systems—including technologies like S3 and Redis—that power the platform's internal data plane and core services. This component of the software is not directly user-facing but has strong implications on the scalability and reliability requirements our customers expect. What you will be doing: Prioritize and Define: Work on an agile team to prioritize, define, and ensure the success of infrastructure and managed services features for a high-level strategic roadmap. Stakeholder Collaboration: Prioritize what we should build and when by collaborating with stakeholders on product vision and strategy, while taking customer feedback into account. Technical Discussions: Define how infrastructure features will work through close collaboration with engineers in design sessions an
About Us: Sauce Labs is the world’s largest full-lifecycle, test automation platform, and the company behind Selenium. Trusted by 80% of the world’s top ten largest financial institutions and over 300,000 enterprise users, Sauce Labs provides the only AI platform capable of turning business intent into autonomous testing and quality assurance. With a proprietary dataset of 8.7 billion test runs, Sauce Labs empowers the Fortune 2000 to bridge the gap between AI-driven code generation and enterprise-grade software quality. Learn more at saucelabs.com . The Role: We are seeking an innovative and experienced AI Architect to join our engineering leadership team. This is a strategic role that will be instrumental in designing and building the next generation of AI-powered features for our continuous testing platform. You will be responsible for architecting scalable and robust AI solutions that transform how our customers gain insights from their test data and production environments, and how they create tests. Responsibilities: Define AI Architecture: Lead the design and architecture of cutting-edge AI/ML solutions for new product offerings, ensuring scalability, performance, quality and reliability within a cloud-native environment. AI-Powered Insights (Test & Production): Architect AI systems to derive actionable insights from vast quantities of test run logs and analytics data. This includes identifying patterns, anomalies, and performance trends. Production Error Reporting Integration: Design AI solutions that integrate with our existing error reporting product to analyze production issues for mobile and web applications, providing deeper understanding and predictive capabilities. Unified Data Intelligence: Develop architectures for combining insights from both test runs and production data, creating a holistic view of application quality and user experience. Automated Failure Analysis & Remediation: Architect AI models and systems t
NVIDIA pioneers computer graphics, gaming, AI, and accelerated computing. We are looking for a Senior Solution Architect with full-stack software engineering experience to join our team and play an important role in developing Sales AI applications. This position offers the opportunity to design, build, and evolve solutions that bring generative AI and intelligent workflows into everyday sales experiences. You will work across the application stack and collaborate with Product, AI and machine learning, Data, Security, Solution Architecture, and Engineering teams to deliver secure, reliable, and scalable solutions used globally. What you’ll be doing: Collaborate with application teams to design, develop, and maintain scalable full-stack solutions for enterprise sales workflows. Guide technical solutions across front-end, back-end, APIs, data services, integrations, and cloud infrastructure. Translate product requirements and business needs into secure, maintainable solutions and intuitive user experiences. Integrate generative AI models, AI services, APIs, retrieval systems, and agentic workflows into production applications. Design application architectures that support performance, availability, observability, security, scalability, and long-term maintainability. Lead technical design discussions, compare implementation approaches, make informed architecture decisions, and evaluate emerging technologies. Improve engineering practices for testing, code quality, continuous integration and delivery, monitoring, documentation, and production readiness. Investigate complex issues and develop solutions that improve reliability and user experience. Mentor engineers, share technical knowledge, and contribute to engineering standards and collaborative team practices. What we need to see: <
Own the architecture of Myntra’s new product platforms to drive business results Drive and own the architecture and design of some of the most advanced & complex software systems / products in the industry to create company wide impact Help build, mentor and coach a team of very talented Engineers, Architects, Quality engineers, System Operation Engineers and DevOps engineers in architectural and design best practices Experience in distributed systems, cloud service development, deployment and delivery Accountable for the design, for the ease of evolution, quality of the systems, performance, scaling, and availability characteristics and limitations of the systems Envision and develop the long-term architectural direction, with emphasis on platforms/ reusable components while adopting an agile delivery process. Establish structures and processes that ensure a high level of quality and reliability and extensibility of deliverables Drive the creation of next generation extensible web, mobile and fashion commerce platforms, security protocols, customisation and tools to support continuous scaling, internationalisation and platform extensions Drive code and design reviews of components / systems / products in scope and drives the architectural governance for them Set directional paths for the teams/department for adoption of new technology stacks for solving business problems Represent multiple technology domains and Myntra in external technical forums Work with product management, business stakeholders and other engineering leaders to help define mid-term, long-term roadmaps and shape business directions Initiate and deliver leadership training within the engineering organisation, including training new managers, and drive the growth of leaders to create a strong leadership bench. Qualifications & Experience 8+ years of experience in software product development Must have a degree in Computer Science o
NVIDIA Networking is a leading provider of innovative end-to-end InfiniBand and Ethernet connectivity solutions for servers and storage. Our portfolio includes adapter cards, switches, cables, and software designed to optimize Data Center performance with industry-leading bandwidth and scalability. We serve diverse sectors such as high-performance computing, enterprise, cloud computing, and Web 2.0. Our mission is to stay ahead of the market by delivering groundbreaking products and services. Our Ethernet solutions are tailored for industries like Media & Entertainment and any domain that benefits from advanced DataStream and TCP/IP acceleration. What You’ll Be Doing: Lead a team of 8+ mechanical design engineers. Define priorities, create project plans, and allocate resources for mechanical programs in coordination with Product Managers. Drive all electro-mechanical, automated JIG and thermal design aspects, of production test setups, ensure readiness of test and assembly infrastructure for high-volume manufacturing. Develop multiple early design concepts in fast-paced product development cycles. Lead task forces to investigate and resolve production issues, reliability concerns, and conduct failure analysis. Perform risk assessments and implement mitigation strategies during product design. What We Need to See: B.Sc. in Mechanical Engineering or higher. 10+ overall years of relevant experience including 4+ years of experience managing teams of engineers in R&D environment. Proven expertise in developing, testing, and manufacturing of complex automated connection systems with precise moving parts, pneumatic and electro-mechanical systems design. Strong hands-on experience with 3D CAD tools (Creo preferred) static and dynamic mechanical simulation Solid background in designing components and sub systems an
As a Senior Platform Product Manager focused on AI SDLC Trusted Throughput, you will define and drive the product strategy for enabling safe, reliable software delivery at AI-native scale across Datadog’s Internal Developer Platform. As AI accelerates development velocity and system complexity, you will help evolve SDLC systems from human-supervised workflows to platforms with built-in safety, observability, and correctness guarantees. You will partner closely with engineering, security, and developer platform teams to improve deployment reliability, operational visibility, and governance while enabling both engineers and AI agents to move quickly with confidence. This role offers the opportunity to shape foundational developer infrastructure and influence how AI-powered software delivery operates across Datadog. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Own the product strategy, roadmap, and execution for AI-native SDLC throughput and reliability initiatives across Datadog’s Internal Developer Platform Define and drive platform outcomes aligned to DORA metrics, balancing deployment velocity with reliability, change failure reduction, and operational safety Partner with engineering, infrastructure, security, and developer experience teams to build automated validation, auditability, and risk-scoring capabilities into deployment workflows Deliver actionable SDLC observability and diagnostic capabilities that connect executive-level metrics to operational signals across the software delivery lifecycle Drive systems that monitor and validate AI-generated or AI-attributed changes to ensure correctness, compliance, and trustworthy automation Serve as a cross-functional product leader across SDLC Foundations, Security Engineering, and compl
Get new software reliability engineer jobs by email
Daily job updates · Unsubscribe anytime