Jobs in United States

Software Reliability Engineer in United States

2,071 active opportunities · Updated October 2026

Explore current software reliability engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

N
📍 Santa Clara, United States
✓ High-confidence listingCompany trend -8%
Quick readStrong listing-quality and freshness signals

NVIDIA’s Silicon Co-Design Group is the team that gets every GPU, SoC, and CPU silicon program from first power-on to high-volume production. We are hiring a Senior Manager to lead our Test, Manufacturability, Reliability & Quality (TMRQ) organization. This is not a coordination role . Your work decides if a product can be built at scale and trusted in the field. These include production test development (SLT, BLT), control run flow, system reliability stress (HTOL), platform- and board-level manufacturing issue closure, and field diagnostic test development. You lead a team of individual contributors and a first-line manager at the layer where silicon, platform, and software collide with manufacturing reality. Decisions you make show up in yield curves, production ramp , and customer escapes. You are the leader the program turns to when a build is stuck, a control run is fallout-heavy, or a field return points back at silicon . The exceptional hire also uses AI deliberately — with demonstrated workflow impact and the judgment to know where it compresses real work and where it introduces risk. What you will be doing: Keep programs moving. Own the technical execution and velocity of SLT, BLT, Board/Chip/Rack CR, and system reliability stress (HTOL) across every GPU, SoC, and CPU silicon program. Close the hardest multi-functional failures. Resolve Vmin and binning escapes, performance shortfalls, and power anomalies by driving root-cause across design, methodology, DFT, ATE, package, software/firmware, and manufacturing — and own the WARs and productized fixes through to confirmation. Give leadership the clarity to act. Convert raw integration signals — CR fallout, BLT/SLT yield, SHTOL/CHTOL data, RMA trends, customer escalations — into decision-ready options that enable executive leadership to act with confidence on POR, QS/PS gates, a

N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -8%

NVIDIA builds the silicon behind AI, accelerated computing, and graphics. Every watt of performance and every degree of thermal headroom traces back to decisions made in power, performance, and thermal architecture. We are the Silicon Co-Design Group (SCG). We identify, own, and drive system-level co-design ideas. We start with initial concepts and advance to product differentiation across NVIDIA's roadmap. We are hiring a Principal System Power Management and Performance Architect who operates at the ambiguous boundary where workload behavior, silicon capabilities, firmware policies, and platform constraints collide, and who turns that ambiguity into architecture that survives across multiple silicon generations. SCG scope spans architecture, design, software, operations, platforms, and productization. This role shapes system, platform, and data center features and behavior, and partners with teams across NVIDIA. What You'll Be Doing: The work here is rarely well-defined when it arrives. You will be given problems that appear to be performance gaps or power anomalies and encouraged to build a framework for solving them, not just tackle a single instance. Define the multi-generation roadmap for system-level power and performance features, grounded in prototyping, use-case analysis, and cost/benefit trade-offs across segments. You will decide what problems are worth solving and why. Own the architecture and integration strategy for HSIO power management, DVFS, P-states, and low-power features. Your decisions improve product performance, power, and reliability across product lines — not just the current program. Lead system-level boot and IST architecture defining how power and clock domains initialize, sequence, and recover across complex multi-IP systems where the interaction space is large and the failure modes matter. Drive power management strategy at data

G
📍 United States· Full-time· Remote
✓ High-confidence listingCompany trend -100%

From $126.4K/yr

Quick readStrong listing-quality and freshness signals

GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An overview of this role As a Staff Human Resources Information Systems Analyst, People Technology - Workday, you won't just manage our people systems—you'll own them. You'll drive how our people technology evolves, partner deeply with stakeholders to solve root problems, and build scalable solutions that reduce friction and improve reliability, usability, and insight across the People technology landscape. This is a strong fit if you think like a product owner, refuse to accept requests at face value, and can move with speed and precision within public company compliance requirements, including SOX ITGCs. In this role, you'll le

GitRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -83.9%

$342K – $445K/yr

Quick readStrong listing-quality and freshness signals

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role We are seeking a Technical Lead to lead deployment and operations for OpenAI’s Silicon & Systems team. This person will become the Directly-Responsible Individual responsible for bringing OpenAI’s custom silicon and associated systems into data center environments, ensuring successful deployment, bring-up, validation, operational readiness, and ongoing reliability at scale. This role sits at the intersection of silicon, systems, infrastructure, data center operations, and software. You will lead a team focused on taking new hardware platforms from lab validation into production data center deployment. You will be responsible for building the operational processes, technical workflows, tooling, and cross-functional alignment required to deploy and operate custom AI hardware reliably in OpenAI’s supercomputing infrastructure. The ideal candidate is both a strong leader and a deeply technical operator. You should be comfortable staying close to the technical details of hardware bring-up, fleet deployment, debugging, system validation, data center integration, and production operations. This role requires strong execution, excellent cross-functional judgment, and the ability to drive clarity in ambiguous, fast-moving environments. In this role, you will: Lead a team responsible for deployment and operations of OpenAI’s custom silicon and systems in data center environments Own the path from hardware bring-up and validation through production deployment, operati

AWSRestAIGo
M
📍 O Fallon, Missouri, United States
✓ High-confidence listingCompany trend +212.5%
Quick readStrong listing-quality and freshness signals

Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Technical Program Manager Overview As a Senior Technical Program Manager at Mastercard, you’ll bring your expertise with conceptualizing, coordinating and driving technology projects, and coaching Agile practices in high-performing, self-organizing development teams. We’re building a global B2B, microservices-based platform to help businesses of all sizes streamline how they manage payments when buying or selling products & services. As a global business, the projects you lead for Mastercard will deliver software operating at massive scale requiring a focus on performance, security, and reliability. This role will support our Network Solutions team within Payment Networks, assisting the development team with building out software solutions for Mastercard. Role: • Dive as deep as you want into the tech stack, the integration patterns, the organizational capabilities, and the company wide assets that can be leveraged to provide technical solutions to customer problems. • Contribute to the strategies, design choices, and even the cloud infrastructure necessary to build comprehensive and achievable execution plans to deliver high-profile new features and capabilities for our customers. • Drive the execution of an initiative that may span multiple teams and integrations, reporting meani

AIRecruitment
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -8%

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology and amazing people. Today, we're harnessing the boundless possibilities of AI to build the next era of computing. An era in which our GPU acts as the brain of computers, robots, and self-driving cars that can understand the world. Accomplishing unprecedented goals calls for imagination, inventiveness, and exceptional talent from around the world. As a NVIDIAN, you'll be immersed in a diverse, encouraging environment where everyone is inspired to do their best work. Join our team and discover how you can build a lasting impact on the world. NVIDIA designs the silicon behind AI, accelerated computing, and graphics. Power and thermal architecture decisions sit behind every watt of performance and every degree of thermal headroom! We are the Silicon Co-Design Group (SCG). We are hiring a Principal Architect to scout the research and industry landscape, identify emerging system-level co-design ideas in power and thermal, and drive them across teams into product differentiation across NVIDIA's roadmap. This role shapes system, platform, and datacenter feature/behavior, and partners with teams across Nvidia. SCG scope spans across architecture, design, software, operations, platforms, and productization. What you'll be doing: Architect next-generation system, platform, and datacenter-level power and thermal co-design solutions. Scan internal research, academia, standards bodies, and silicon, memory, packaging, and platform partners for what is emerging. Build the product differentiation case for each candidate idea, performance, power, reliability, schedule, cost, and brainstorm what is worth pursuing. Lead end-to-end co-design from concept to product. Drive alignment across architecture, VLSI, softw

Machine LearningAI
T
📍 Tg, Dallas Office, United States
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

The Engineering Manager, Memberships leads the team responsible for building and operating Topgolf’s Memberships systems, from guest-facing membership experiences through the backend services that power them. This role owns the people, process, and delivery of the Memberships engineering team, while staying technically credible across a full stack built on Go, Vue.js, and PostgreSQL. Requirements Lead, grow, and manage a team of full-stack engineers building Memberships systems, including hiring, performance management, career development, and mentorship Set clear goals and expectations for the team, run effective 1:1s, and build a culture of ownership, accountability, and continuous improvement Balance workload and staffing across Memberships initiatives, escalating resourcing gaps and continuity risks early Guide architecture and design decisions across Memberships systems, drawing on full-stack experience spanning Go, Vue.js, PostgreSQL, and API design Set engineering standards and best practices for code quality, testing, and release processes, and stay close to the codebase through code reviews and hands-on problem solving on critical issues Own delivery of the Memberships roadmap end to end, from technical planning through implementation, QA, release, and post-launch monitoring, ensuring systems are observable, testable, secure, and built to scale with guest demand Partner with product, design, QA, and platform engineering to translate guest needs and business priorities into a clear, prioritized Memberships roadmap Represent the Memberships team in cross-functional planning and architecture discussions, communicating progress, risks, and tradeoffs to engineering leadership and business stakeholders Critical Skills Strong architectural judgment and the ability to balance technical debt, delivery speed, and long-term maintainability <

PythonVuePostgreSQLAWS
C
📍 United States· Remote
✓ High-confidence listingCompany trend +340.2%
Quick readStrong listing-quality and freshness signals

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Job Summary Entry‑level developer role focused on building and supporting modern full‑stack applications. Ideal for recent graduates or early‑career professionals with experience in AWS cloud development (including Python and PySpark) and SQL Server. Experience Agile practices and exposure to AI‑assisted development is also needed. Key Responsibilities Develop, maintain, and support applications and services deployed in AWS cloud environments. Build and optimize SQL Server database components, including queries and stored procedures. Translate business and technical requirements into scalable, maintainable solutions. Write clean, secure code adhering to development standards and best practices. Leverage AI assistants for code generation, refactoring, testing, and documentation with appropriate validation. Participate in Agile practices and events (planning, stand‑ups, reviews, retrospectives). Troubleshoot and resolve application and data issues. Contribute to code reviews and continuous quality improvement. Collaborate across engineering, QA, and business teams. Support application deployment and operations in AWS. Required Qualifications Bachelor’s degree in Computer Science, Information Technology, Software Engineering, or a related field. 1–5 years of professiona

PythonSQLAWSGit
C
📍 United States· Remote
✓ High-confidence listingCompany trend +340.2%
Quick readStrong listing-quality and freshness signals

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary The Software Development Engineer is responsible for designing, developing, testing, and maintaining software applications and services under the guidance of senior engineers and technical leads. This role collaborates with cross-functional teams to deliver high-quality, scalable, and secure solutions while continuously building technical expertise and contributing to team objectives. The Software Development Engineer participates in all phases of the software development lifecycle, including requirements analysis, coding, testing, deployment, and production support. Required Qualifications 0–2 years of experience in software development, application development, or a related technical field. Basic knowledge of one or more programming languages such as Java, Python, C#, JavaScript, or similar. Understanding of software development principles, data structures, algorithms, and object-oriented programming concepts. Familiarity with relational databases, APIs, and web technologies. Experience working with source control systems such as Git. Strong analytical, problem-solving, and debugging skills. Effective written and verbal communication skills. Ability to work collaboratively in a team-oriented environment. Preferred Qualifications Internship or academic pr

JavaScriptPythonJavaAWS
L
📍 Bethesda, United States
✓ High-confidence listingCompany trend +500%
Quick readStrong listing-quality and freshness signals

Leidos has an exciting opportunity for a Software Engineer (SME) in our Intel Security Sector's Analysis Solutions Business Area . Our talented team is at the forefront in Security Engineering, Computer Network Operations (CNO), Mission Software, Analytical Methods and Modeling, Signals Intelligence (SIGINT), and Cryptographic Key Management. At Leidos , we offer competitive benefits , including Paid Time Off, 11 paid Holidays, 401K with a 6% company match and immediate vesting, Flexible Schedules, Discounted Stock Purchase Plans, Technical Upskilling, Education and Training Support, Parental Paid Leave, and much more. Join us and make a difference in National Security! Job Summary As a Software Engineer on this program, you will have the opportunity to build strong systems, software, and cloud environments while providing operations and maintenance for critical systems. This role will provide technical expertise in the design, development, implementation and testing of customer tools and applications. Based in a DevOps framework, this role participates in and/or directs major deliverables of projects through all aspects of the software development lifecycle including scope and work estimation, architecture and design, coding and unit testing. Primary Responsibilities: Participates in and/or directs software programming initiatives using Java, JavaScript, Python, SpringBoot, and Hibernate. Develops and directs software system validation and testing methods using Junit and Katalon and uses integrated custom developed software solutions to leverage automated deployment technologies Develop, prototype and deploy solutions within Commercial Cloud Solutions leveraging infrastructure platform services Coordinate closely with team members, Product Owners and Scrum Masters to ensure User Story alignment and implementation to customer use cases

JavaScriptPythonJavaAWS
N
📍 Santa Clara, United States
✓ High-confidence listingCompany trend -8%
Quick readStrong listing-quality and freshness signals

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. NVIDIA is looking for an outstanding Quality Assurance Engineer to join our Digital Marketing Operations Team to help drive the delivery of global web initiatives and operations. This role entails being responsible for the project management and coordination of software testing for all web and mobile releases. The ideal candidate is someone who thrives on solving complex problems and has hands-on experience in quality assurance and automation. What you'll be doing: Work together with development teams to triage issues, conduct root cause analysis, and confirm fixes. Define new tests and improve existing test plans to ensure the highest quality standards. Develop and update automation testing scripts using Java, TestNG, Selenium for Frontend, REST/Services, Database, and AEM components. Collaborate with teams located in different regions, joining reviews and meetings to keep alignment and cohesion. Contribute to the creation and improvement of sophisticated, stable automation frameworks that boost efficiency, re-usability, and flexibility. Improve NVIDIA's continuous integration framework by merging build, compiling, test, and validating processes with scheduling, automated notifications, and

JavaScriptJavaReactGit
N
📍 San Francisco, California, United States· Internship· Remote
✓ High-confidence listingCompany trend -88.6%
Quick readStrong listing-quality and freshness signals

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. Notion is an in person company, and currently requires its employees to come to the office for three Anchor Days (Mondays, Tuesdays, and Thursdays). This internship will take place from Jan 25 - Apr 16 and you will need to be able to work out of our NY or SF office during this time. About the Role: Mobile devices have reshaped personal computing—the power of a desktop now fits in our pockets. By giving everyone the building blocks to create their own tools, we'll solve more problems wherever we go. A delightful, intuitive mobile experience: that's our guiding light, and we need Android and iOS engineers to make it a reality. During your 12-week internship, you will be paired with a mentor that will help guide you as you work closely with our team to build and ship impactful projects. These projects will drive valuable impact to our customers and engineers. Take a look at what our past intern cohort worked on with our Blog post here and here and TikTok post. What You'll Achieve: Write clean, secure, tested, and documented code. You'll collaborate with the team to strategize, mold, and develop innovative features for our Android and iO

TypeScriptPythonJavaReact
Y
📍 New York, NY, United States
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Yext (NYSE: YEXT) is the enterprise agentic marketing platform. Built on the world's most comprehensive structured data platform for local businesses, Yext gives brands and their partners the visibility intelligence to win every moment of discovery – across AI and traditional search. Yext's API-first architecture connects structured data to APIs, MCP servers, and generative interfaces, so partners and developers can build purpose-built experiences on the same infrastructure powering Yext's own products. Thousands of brands and digital marketing partners in financial services, healthcare, retail, hospitality, and food rely on Yext to manage, measure, and optimize visibility at scale. For more information, visit yext.com . At Yext, Product Engineering builds and evolves the technology behind our products and services. We’re looking for software engineers who want to solve meaningful technical problems, contribute to systems at scale, and help shape what we build next. We work in an agile environment with two-week sprints and regular demos that keep teams aligned and give engineers clear visibility into the impact of their work. From day one, you’ll contribute directly to the codebase and collaborate with experienced engineers from a wide range of leading universities and technology companies. We are looking for an engineer to join Team Watson , which owns and develops the systems that power Yext Search and Yext Chat. The team builds the indexing, retrieval, and serving technology that enables brands to deliver fast, relevant answers across their websites and digital experiences. Yext Search handles more than 50 million requests each month, serving users around the world in over a dozen languages. Watson also brings these search and retrieval capabilities to Yext Chat, helping conversational experiences generate useful answers grounded in trusted customer content. Because Watson’s systems serve real-time, customer-facing experiences at a global scale, engineers o

PythonJavaAIC++
C-
📍 New York, NY, United States
✓ High-confidence listing

$220K – $290K/yr

Quick readStrong listing-quality and freshness signals

CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. We're looking for Senior Fullstack Software Engineers to help build the next generation of CLEAR's mission and vision. Beyond verifying identity, we're creating a secure, networked digital identity that enables seamless experiences across travel, enterprise, healthcare, financial services, and beyond. As a Senior Software Engineer III, you'll own complex technical problems from design through deployment, partnering closely with Product, Design, Security, and Operations to deliver reliable, scalable solutions. We're looking for engineers with a strong builder mindset who thrive in ambiguity, take ownership, and enjoy turning ideas into production systems. Level and team matching (open roles across the three pillars that make up Technology at CLEAR: Core Identity, CLEAR1 , and CLEAR Travel ) will occur towards the end of our interview process. Tech stack overview: A brief highlight of our tech stack: Java / Kafka / Postgres AWS cloud What you'll do: Advance our capabilities across a wide array of industries and domains and gain hands-on experience with privacy, security, data modeling and a

JavaPostgreSQLAWSDocker
C
📍 Work At Home South Carolina, United States
✓ High-confidence listingCompany trend +340.2%
Quick readStrong listing-quality and freshness signals

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary Join the CVS Mail Dispensing Applications Technology team as a Software Development Engineer, where you will play a key role in delivering innovative healthcare technology solutions that support automated mail-order pharmacy operations. This role is responsible for designing, developing, implementing, and supporting software applications that enable critical dispensing processes across multiple integrated systems, including prescription intake, inventory management, filling, scanning, sorting, packing, postage, and reporting. You will work in a collaborative, diverse, and fast-paced environment focused on operational excellence, continuous improvement, and high-quality software delivery. The ideal candidate will bring strong technical expertise, a proven ability to lead complex application development initiatives, and experience partnering with both internal and external teams to drive business outcomes. Responsibilities include: Develop and implement technical solutions for Dispensing Applications Technologies, ensuring adherence to established methodologies, standards, and guidelines. Analyze, design, code, test, debug, and document software applications supporting automated mail-order pharmacy operations. Lead the development and implementation of large-scale and complex application and process redesign/improvement initiatives across

AngularSQLAzureDocker
🔔

Get new software reliability engineer jobs in United States by email

Daily job updates · Unsubscribe anytime