Senior Software QA Test Development Engineer - Diagnostics — US, CA, Santa Clara. Apply via Workday.
Jobs in United States
Test Production Shift Supervisor in United States
724 active opportunities · Updated October 2026
Showing
15 jobs
Explore current test production shift supervisor jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
Senior System Software Test Engineer, Networking — US, CA, Santa Clara. Apply via Workday.
Data Center Power Test Architect — US, CA, Santa Clara. Apply via Workday.
We are seeking a highly skilled and hard-working Senior Test Developer / test engineer to join our multifaceted Enterprise Software QA team. This role offers an outstanding opportunity to leave your mark on the design, construction, optimization and testing of large-scale infrastructure for various foundational NVIDIA unified cloud services and data center offerings. If you are a dedicated engineer with strong expertise in cloud infrastructure and distributed systems and want to apply your skills with AI tools, this role could fit you perfectly. You will thrive in an exciting, innovative environment. What you'll be doing: Work with development teams on test plans for all layers of SW stack for cloud infrastructure, execution, reviews, failure analysis and assessing overall quality and risk. Work with customer PMs on software issues including technical feedback from OEMs and CSPs. Develop key benchmarks to track execution and deploy process improvements to improve efficiency Leverage AI skills to expedite the test scope, test plan, execution and automation workflows. Lead NVIDIA Cloud and Data Center bring up activities which will involve validation, reporting, working with engineering to debug issues, providing design input at times, adding coverage in different areas. Design, develop and maintain CI/CD pipelines for continuous testing in cloud environments when needed. Perform performance, scalability, and reliability testing of cloud services. Implement and maintain test environments in cloud platforms such as AWS, Azure, or Google Cloud. Supervise the infrastructure to alert on significant events, ensuring the highest level of system performance and reliability. Work with various different partner teams to ensure availability of clusters to test on and take the lead in resolve all issues. Working with tea
At NVIDIA, we are at the forefront of technological innovation, pushing the boundaries of AI and accelerated computing. Our team in Santa Clara, CA is looking for a Senior Software Engineer in Test to join us in this exciting journey. This is a ground breaking opportunity to work with powerful technology, collaborate with a world-class team, and make a significant impact in the industry. If you are passionate about AI and quality assurance, and thrive in a dynamic environment, this role is perfect for you! What you'll be doing: Accomplishing test cases to validate NVIDIA enterprise offerings, such as NIM, NeMo, and BioNeMo. Crafting, implementing, and maintaining automated test cases and supporting automation infrastructure. Collaborating with development teams to triage issues, perform root cause analysis, verify fixes, define additional tests, and improve test plans. Investigating and bringing to bear AI capabilities to accelerate the Quality Assurance (QA) process. What we need to see: MS or PhD degree in computer science or relevant field, or equivalent experience. At least 5+ years of professional experience in software testing. Proficiency in oral and written English. Comfort working with Linux OS. Strong skills in shell and Python programming. Strong knowledge of QA principles and background in software testing. Experience using AI development tools for crafting test plans, developing test cases, and automating test cases. Excellent problem-solving abilities. Strong interpersonal skills, quick learning ability, proactive approach, innovation, and dedication. Self-motivation and a passion for learning new hardcore technology. Knowledge in LLM and AI models is a plus. <
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The success and effectiveness of our engineering teams is the most critical factor in helping customers succeed with Snowflake. As a Staff Product Manager for Developer Experience, you will lead the mission to build an industry-leading, AI-forward developer experience for the engineers at Snowflake. This is a high-impact, high-ownership role where you will act as the single-threaded owner (DRI) for our internal developer platform and product validation strategy. Your primary focus will be redefining how the Snowflake product is developed and tested - shaping an AI-forward and opinionated developer experience and a modern, intelligent test pyramid that balances speed, reliability, and coverage. You will ensure our internal tools enable engineers to move with maximum velocity while maintaining the world-class quality our customers expect. AS A STAFF PRODUCT MANAGER, YOU WILL : Drive AI-Native Validation: Lead the transformation of our testing suites by integrating AI to generate test cases, identify regression risks, and optimize test selection, significantly accelerating the time-to-impact. Redefine Developer Workflow: Build an AI-forward developer experience that removes friction from the "inner loop" of development, ensuring engineers can validate and iterate on their idea
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. The HBM Design Technology Package Co-Optimization (DTPCO) organization is seeking a Staff Engineer! Lead the design and development of test vehicles that enable technology learning and risk reduction for future HBM products! This role serves as a key technical contributor within DTPCO, partnering closely with Architecture, Design, Layout, Product Engineering, Reliability, and Advanced Packaging teams to translate emerging product requirements and technology challenges into actionable test vehicle solutions. You will apply expertise in semiconductor design, physical implementation, silicon characterization, and advanced packaging technologies to define test structures, develop validation strategies, and generate data that drives product decisions across future HBM generations. Unlike a traditional product role centered on ownership of product blocks, this position emphasizes enabling learning in development and engineering. It does so by developing representative test vehicles and characterization structures. Key Responsibilities Serve as a technical lead for HBM test vehicle design and structure development within DTPCO. Partner with HBM product, technology and reliability team. Define and implement test structures targeting key learning areas such as package interaction and reliability Drive test vehicle content definition, structure specifications, measurement requirements, and validation objectives. Analyze silicon characterization and qualification results to identify optimization opportunit
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. The Firmware & Product Test (FPT) team plays a critical role in delivering high-quality enterprise SSD solutions by ensuring firmware functionality, reliability, and compliance. We work across simulation, FPGA, and hardware environments to validate modern storage technologies, build scalable automation, and drive continuous improvement in validation methodologies. Our team values technical excellence, collaboration, and innovation, including the use of AI-enabled tools to enhance engineering productivity and quality. As a Principal Test Development Engineer, you will serve as a technical leader for firmware validation, defining verification strategies, advancing automation frameworks, and driving complex failure analysis efforts. This role offers the opportunity to influence product quality across multiple SSD programs while mentoring engineers and partnering closely with firmware architects to improve testability and validation effectiveness. Responsibilities: Lead verification strategy, test planning, automation, and coverage closure for NVMe front-end firmware features across multiple product lines Architect and enhance scalable Python-based test automation frameworks, CI/CD integration, regression infrastructure, and reporting capabilities Drive root-cause analysis and failure triage using firmware traces, protocol analyzers, system logs, and structured debug methodologies Define validation standards, review test code, mentor engineers, and promote standard methodologies in automation and qua
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Our Senior Software Engineers independently drive complex technical work, shape the systems and technical decisions within their teams, and enable other engineers to deliver high-quality, scalable solutions. Vanta's product monitors the security posture of thousands of companies, pulling tens of millions of API calls of data per day, pushing information from hundreds of thousands of laptop agents, and running tests against that data continuously to identify potential security threats. Our infrastructure and tooling need to stay ahead of exponential growth in our customer base. As a Senior Software Engineer at Vanta, you'll drive complex projects across our technical stack, contribute to the technical direction of your team, and mentor other engineers. Your past experience will be leveraged to enable and accelerate Vanta's growth. Visit our Vanta Engineering Blog to learn more about what our team is working on! Tests are at the heart of how Vanta continuously monitors security and compliance for our customers. The Test Core team builds the runtime platform that powers these checks. We own how Tests are scheduled and executed, how their results are persisted and exposed, and the systems that keep this runtime reliable as Vanta grows. In this role, you'll work on some of the core systems behind Vanta's Tests platform. You'll tackle problems around the reliability, correctness, and performance of Test execution, evolve the systems and abstractions that allow the platform to scale, and make it easier for other engineering teams to build on the Tests runtime. Many of these problems span multiple systems and teams and require a deep u
Principal Electro-Optical Infrared Design, Assembly, Integration, and Test Engineer Company: The Boeing Company Boeing Defense, Space and Security (BDS) is seeking a Principal Electro-Optical Infrared Design, Assembly, Integration, and Test Engineer (Level 6) for a satellite development program in Huntington Beach or El Segundo, California. This position will lead a team of engineers, analysts and staff within the Space and Missile Systems (SMS) and Boeing Technology and Innovation (BTI) Mission Systems organizations in the development, assembly, integration, and test of a constellation of satellites. The SMS and BTI organizations develop and capture technology for designing disruptive Mission Systems solutions. Focused on visible and infrared spectrum Electro-Optics Infrared (EO/IR) sensors operating in Space and Air domains, our growing team is leaping ahead of our competition with an exceptional mix of mission architectures, sensor designs, and algorithms for advanced image and data processing. Position Responsibilities: Drives technical excellence throughout all Optical Sensor development activities Troubleshoots sensor design and integration issues with the telescope and sensor focal plane Provides technical leadership to SMS and BTI sensor teams Serves as the primary technical interface with the EO/IR suppliers Provides oversight and approval of technical approaches, products and processes. Works closely with program manager, program chief engineer, and other Boeing teams Ensure mission success and prompt resolution of anomalies Provides mentoring and coaching to Optical Sensor Engineers and Technicians Performs Technical Lead Engi
Experienced Electro-Optical Infrared Design, Assembly, Integration, and Test Engineer Company: The Boeing Company Boeing Defense, Space and Security (BDS) is looking for an Experienced Electro-Optical Infrared Design, Assembly, Integration, and Test Engineer (Level 3) to join our team in Huntington Beach or El Segundo, California. This position will be joining a team of engineers, analysts and staff within the Space and Missile Systems (SMS) and Boeing Technology and Innovation (BTI) Mission Systems organizations in the development, assembly, integration, and test of a constellation of satellites. The SMS and BTI organizations develop and capture technology for designing disruptive Mission Systems solutions. Focused on visible and infrared spectrum Electro-Optics Infrared (EO/IR) sensors operating in Space and Air domains, our growing team is leaping ahead of our competition with an exceptional mix of mission architectures, sensor designs, and algorithms for advanced image and data processing. Position Responsibilities: Develops and validates requirements for various communication, sensor, electronic warfare and other electromagnetic systems and components Develops and validates electromagnetic requirements for electrical\electronic systems, mechanical systems, interconnects and structures Develops architectures to integrate systems and components into higher level systems and platforms Performs trade studies, modeling, simulation and other forms of analysis to predict component, interconnects and system performance and to optimize design around established requirements Defines and conducts tests to validate performance of designs to requirements Manages appropriate aspe
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Responsibilities and Duties We are seeking a highly skilled System Tests & Diagnostics Engineer to develop, extend, and integrate specialized silicon validation and diagnostics tools for next-generation AI SoCs. Unlike traditional validation roles focused on executing test plans, this position is responsible for developing the diagnostic software and stress tools that expose hardware failures, characterize silicon behavior, and improve platform observability throughout bring-up and validation. You will work closely with Arm engineers to understand and extend existing diagnostics technologies while developing Graphcore-specific capabilities for future AI hardware. Role Summary You will work with existing Arm-developed diagnostics technologies and extend them to support Graphcore's next-generation AI silicon. You will be responsible for developing system-level diagnostics and stress tools that integrate with an existing framework to detect data integrity, computational correctness, performance, and reliability issues across CPUs, AI accelerators, memory, storage, PCIe, firmware, BMC, and other platform components. Examples include silent data corruption (SDC) tests, power transient stress tools, and platform diagnostics, with opportunities to develop new diagnostics as future hardware capabilities evolve. This role requires close collaboration with hardware architects, firmware enginee
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Responsibilities and Duties We are seeking a highly skilled System Tests & Diagnostics Engineer to develop, extend, and integrate specialized silicon validation and diagnostics tools for next-generation AI SoCs. Unlike traditional validation roles focused on executing test plans, this position is responsible for developing the diagnostic software and stress tools that expose hardware failures, characterize silicon behavior, and improve platform observability throughout bring-up and validation. You will work closely with Arm engineers to understand and extend existing diagnostics technologies while developing Graphcore-specific capabilities for future AI hardware. Role Summary You will work with existing Arm-developed diagnostics technologies and extend them to support Graphcore's next-generation AI silicon. You will be responsible for developing system-level diagnostics and stress tools that integrate with an existing framework to detect data integrity, computational correctness, performance, and reliability issues across CPUs, AI accelerators, memory, storage, PCIe, firmware, BMC, and other platform components. Examples include silent data corruption (SDC) tests, power transient stress tools, and platform diagnostics, with opportunities to develop new diagnostics as future hardware capabilities evolve. This role requires close collaboration with hardware architects, firmware enginee
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Responsibilities and Duties We are seeking a highly skilled System Tests & Diagnostics Engineer to develop, extend, and integrate specialized silicon validation and diagnostics tools for next-generation AI SoCs. Unlike traditional validation roles focused on executing test plans, this position is responsible for developing the diagnostic software and stress tools that expose hardware failures, characterize silicon behavior, and improve platform observability throughout bring-up and validation. You will work closely with Arm engineers to understand and extend existing diagnostics technologies while developing Graphcore-specific capabilities for future AI hardware. Role Summary You will work with existing Arm-developed diagnostics technologies and extend them to support Graphcore's next-generation AI silicon. You will be responsible for developing system-level diagnostics and stress tools that integrate with an existing framework to detect data integrity, computational correctness, performance, and reliability issues across CPUs, AI accelerators, memory, storage, PCIe, firmware, BMC, and other platform components. Examples include silent data corruption (SDC) tests, power transient stress tools, and platform diagnostics, with opportunities to develop new diagnostics as future hardware capabilities evolve. This role requires close collaboration with hardware architects, firmware enginee
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Responsibilities and Duties We are seeking a highly skilled System Tests & Diagnostics Engineer to develop, extend, and integrate specialized silicon validation and diagnostics tools for next-generation AI SoCs. Unlike traditional validation roles focused on executing test plans, this position is responsible for developing the diagnostic software and stress tools that expose hardware failures, characterize silicon behavior, and improve platform observability throughout bring-up and validation. You will work closely with Arm engineers to understand and extend existing diagnostics technologies while developing Graphcore-specific capabilities for future AI hardware. Role Summary You will work with existing Arm-developed diagnostics technologies and extend them to support Graphcore's next-generation AI silicon. You will be responsible for developing system-level diagnostics and stress tools that integrate with an existing framework to detect data integrity, computational correctness, performance, and reliability issues across CPUs, AI accelerators, memory, storage, PCIe, firmware, BMC, and other platform components. Examples include silent data corruption (SDC) tests, power transient stress tools, and platform diagnostics, with opportunities to develop new diagnostics as future hardware capabilities evolve. This role requires close collaboration with hardware architects, firmware enginee
Other cities to consider
More places hiring for this role
Get new test production shift supervisor jobs in United States by email
Daily job updates · Unsubscribe anytime