About the Team The Systems Integration team is responsible for building the infrastructure, tooling, and validation systems that ensure our device software is reliable, testable, and ready to ship. We design and maintain automated test frameworks, hardware-in-the-loop labs, and release pipelines that keep quality signals trustworthy and enable rapid, safe product launches. Our work spans developer tools, automation, systems integration, and cross-team collaboration to ensure every release meets the highest standards. About the Role As a Software Engineer, Quality and Developer Tools , you will build and own the systems that validate our device software—from test frameworks and regression infrastructure to hardware-in-the-loop labs and release gates. You’ll design the tooling and automation that keep quality signals trustworthy, integrate them into CI/CD, and make it easy for engineers and QA vendor technicians to execute reliable, repeatable workflows. We’re looking for engineers with deep experience in software quality, automation, developer tooling, and hardware-software integration who thrive on building scalable, reliable systems for validation and release readiness. This role is based in San Francisco, CA. We use a hybrid work model of four days in the office per week and offer relocation assistance to new employees. In this role, you will: Test infrastructure & frameworks: Design, implement, and maintain a unified test framework for device software across unit, integration, system, and end-to-end testing, with reproducible runs and integrations with GitHub, Linear, and Slack. CI/CD integration & releases: Integrate test suites with Buildkite, enforce promotion criteria for staging and production, auto-file regressions, and publish traceable artifacts and release notes. Hardware-in-the-loop lab design & orchestration: Plan and bring up racks, power and networking systems, and orchestration for device testing; support automated flashing, provisioning
Jobs in United States
Test Analyst 3 in San Francisco
94 active opportunities · Updated October 2026
Showing
15 jobs
Explore current test analyst 3 jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples' lives. About the Role You will work at the cutting edge of OpenAI's robotics hardware, partnering with hardware and software teams to develop, build, test, and iterate on robotic systems. Working closely with engineers, you will build prototypes, execute experiments, fabricate fixtures, troubleshoot hardware, and help drive rapid iteration on new robotic technologies. Your practical problem-solving and technical judgment will help move projects from concept to reality. We are looking for versatile, hands-on generalists who enjoy solving problems across mechanical, electrical, fabrication, and experimental domains. The ideal candidate will have strong experience in high-velocity and early-stage hardware development environments, provide thoughtful feedback to engineering teams, and help improve both hardware and development processes. This role will help accelerate robotics development by enabling fast, high-quality iteration on new hardware concepts and systems. This role is based in San Francisco, CA, and is in-person 5 days a week. In this role, you will: Partner closely with engineers, researchers, and other members of the robotics team to support prototype development, experimentation, and hardware iteration. Build, modify, troubleshoot, and repair robotic systems and electromechanical assemblies spanning structural, electrical, sensing, and actuation subsystems. Design and fabricate fixtures, adapters, test equipment, and other prototype hardware that accelerate development efforts. Execute low-volume builds, engineering changes, and rework activities acro
About the Team The Human Data team at OpenAI is responsible for identifying and mitigating risks in advanced AI systems by designing evaluations, surfacing vulnerabilities, and collaborating closely with researchers to strengthen model reliability and public trust. About the Role As a Research Program Manager, you will lead initiatives that test the safety and robustness of OpenAI’s models through creative experimentation and structured evaluation. You’ll coordinate efforts across research and engineering teams to transform ambiguous risks into concrete research programs and influence future model development and deployment. We’re looking for people who are technically savvy, comfortable with ambiguity, and excited about shaping the future of safe AI. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead programs that explore unexpected model behaviors and identify failure modes. Translate vague or emergent risk signals into clear priorities and actionable research plans. Design and run creative evaluations, experiments, and red-teaming campaigns. Collaborate with research, product, and deployment teams to integrate findings into model training and deployment cycles. Develop repeatable systems for tracking model performance and understanding emerging behavior patterns. You might thrive in this role if you: Have strong experience in technical program management, with excellent organizational and communication skills. Are familiar with large language models, prompt engineering, or model evaluation techniques. Are comfortable managing fast-paced, high-uncertainty projects and shaping them from the ground up. Are creative and resourceful in devising new methods for testing model behavior and performance. Can effectively coordinate across technical and non-technical stakeholders to drive alignment and execution. About OpenAI OpenAI is an AI resear
About the Team The Scaling team is responsible for the architectural and engineering backbone of OpenAI’s infrastructure. We design and deliver advanced systems that support the deployment and operation of cutting-edge AI models. Our work spans system software, networking, platform architecture, fleet-level monitoring, and performance optimization. About the Role We’re hiring an SW Engineer to enable production workloads and end-to-end testing on new platforms. This role will include creating new test harnesses and platform stress benchmarks, porting existing inference and training workloads to new, sometimes early-access, systems/hardware, analyzing performance and bottlenecks, and characterizing the end-to-end behavior of new systems (compute, comms, storage, control plane, and failure modes). Key Responsibilities Port and validate key inference and training workloads on new platforms/SKUs as they arrive; drive correctness, performance, and stability to an internal readiness bar. Build a suite of benchmarks and stress tests that capture real E2E behavior of our workloads by exercising all aspects of a system, including CPU, GPU, memory subsystem, frontend, scale-up, and scale-out networking (including WAN traffic, NVlink and RDMA collectives), storage, thermals, and any other relevant parts. Deep-dive performance on distributed training/inference: Collective performance and tuning (across NCCL/RCCL and internal libraries) Overlap of compute/communication, kernel-level bottlenecks, memory bandwidth and scheduling effects Create repeatable test harnesses that run in CI / lab environments and produce actionable outputs (pass/fail, performance score, regression detection). Partner with systems + fleet bring-up engineers to ensure the platform is not only stable and performant, but also operationally usable and scalable (containerization, K8s integration, telemetry hooks, failure triage loops). Work cross-functionally with vendors and internal stakeholders by producing
About the Team We’re hiring software engineers to make OpenAI’s networking teams more productive. These teams build and operate the high-performance networking systems that support OpenAI’s training and inference infrastructure at frontier scale. About the Role We’re looking for someone who cares deeply about the developer experience of engineers working on complex infrastructure systems — especially around build systems, test architecture, release pipelines, and reliable development workflows. This role will be embedded with OpenAI’s networking team: making it faster, safer, and easier for engineers to build, test, validate, and ship changes across multi-server, networked, and hardware-adjacent environments. In this role you will: Improve development workflows for engineers building and operating OpenAI’s networking systems Design and improve continuous deployment, release, and validation pipelines Build and maintain test harnesses for multi-server, networked, and hardware-backed environments Improve iteration speed across C++, Python, and build-system-heavy codebases Partner with engineers to identify friction in CI, testing, debugging, and deployment workflows Drive testing and reliability strategy for infrastructure components that support large-scale training and inference workloads Work closely with centralized developer experience teams while staying deeply embedded with the networking engineers closest to the systems You might thrive in this role if: You are motivated by helping other engineers move faster and with more confidence You have experience with CI/CD, release pipelines, testing infrastructure, or build systems You are comfortable moving between C++, Python, and build systems such as CMake, Bazel, or Blaze You enjoy building test harnesses, automation, and workflow improvements for complex systems You do not need to be a networking expert, but you are excited to learn enough about the domain to make the team meaningfully more effective When you see
Strength in Trust OneTrust’s mission is to enable innovation through the responsible use of data and AI. We believe that ensuring data is trusted shouldn’t slow teams down—it should accelerate what’s possible. This led us to develop the first technology platform for responsible data use in 2016. Today, with AI representing the latest and most impactful expansion of data yet, OneTrust is once again redefining what responsible innovation looks like. OneTrust, the AI‑Ready Governance Platform™, unifies regulatory intelligence, automation, and connected governance workflows so businesses can continue to move at the speed of AI while ensuring good governance to prevent data misuse at scale. Trusted by thousands of organizations worldwide, OneTrust is shaping the future where trusted data becomes a transformative force for business and society. The Challenge As a Senior Staff Software Engineer, you will serve as a technical leader for OneTrust’s AI Governance (AIG) platform, driving the design, scalability, and reliability of systems that enable enterprises to deploy and govern AI and LLM-powered applications responsibly. You will deeply understand how customers build, deploy, and operate AI systems, and translate those needs into secure, compliant, and observable platform capabilities. Your Mission Development Lead the design and development of Java/Python microservices and shared libraries integrating with AI platforms for OneTrust’s AI Governance product. Design, build, and test cloud-native applications deployed on Microsoft Azure using Core Java, REST, and the Spring ecosystem. Lead the architecture and development of reusable AIG reporting and dashboard capabilities that integrate governance data from SQL databases and analytical platforms with runtime observability signals. Design reusable semantic-layer and metric-abstraction capabilities, including dataset contracts, metric defini
About the Team The Technology Vertical team builds within Core Products. The team’s mission is to transform every major function in a technology company with AI to drive higher economic output, while people remain owners of the outcomes: setting intent, applying judgment and craft, directing iteration, and owning the result. About the Role As a Product Manager on the Technology Vertical team, you will shape how knowledge workers use AI in their daily work. You’ll craft enterprise products at the intersection of novel research capabilities and genuine enterprise needs, driving the full feedback loop from research, to product, to OpenAI’s own enterprise functions, to the broader Technology Vertical. This role is based in San Francisco, CA. We use a hybrid work model of three days in the office per week and offer relocation assistance to new employees. In this role, you will: Own the product roadmap for a line of business within tech verticals, using OpenAI’s organization to test, refine, and scale new workflows. Develop the ecosystem strategy across technology partners in your line of business You might thrive in this role if you: Have built enterprise products for non-developer knowledge workers and understand the needs of both individual contributors and organizational buyers. Can turn complex workflows, integrations, and permission requirements into clear product decisions. Are comfortable developing a strategy while working closely with engineers, customers, and cross-functional partners to ship and learn. Bring strong judgment about which vertical-specific needs should become reusable capabilities across a broader product. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and
About the Team The Plugin Ecosystem team builds the platform and product experiences that let people extend ChatGPT and Codex. We work on plugins, skills, connectors, interactive apps, and open standards like the Model Context Protocol (MCP). We make plugins easy to discover, install, and use, ensure they’re invoked at the right time, and help people find new ways to get value from them. We want anyone to be able to turn a useful workflow into a plugin, share it, and have other people use it. A plugin can package instructions and skills with connections to the tools and data it needs. Our work spans creation and publishing, reliable execution across our products, clear permissions and approvals, and the controls admins need to bring plugins to their organizations. We work closely with research to improve plugin quality as models evolve. About the Role We’re looking for product-minded engineers to build the systems behind plugins and improve how models use them. Depending on your focus, you may scale generalist infrastructure and identity-related integrations across products, or improve plugin quality at the intersection of backend engineering and applied AI or work on the product experience itself to drive plugin usage. You’ll work across teams and own problems from diagnosis and design through implementation and release. This role is based in San Francisco. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design and ship APIs, SDKs, and services that developers use to extend ChatGPT and Codex. Build intuitive experiences that help users discover, install, and use plugins to get more done. Make plugins easier to create, test, publish, update, and share. Improve when and how models use plugins, from choosing the right plugin to completing a task. Work with Research to diagnose failures and measure improvements as models evolve. Improve plugin reliability and interaction quality acros
$900K – $1M/yr
About us EVERY™ is a leading VC-backed food tech ingredient company and market leader using precision fermentation to create animal proteins without the animal for the global food and beverage industry. EVERY™ is a team of passionate change-makers who are reimagining the factory farm model with a kinder, more sustainable alternative. Leveraging precision fermentation to produce hyper-functional and one-to-one replacement proteins from microorganisms, EVERY™ is on a mission to decouple the world’s proteins from the animals that make them. We are a passionate, determined (and fun!) team with a vital objective, and we're on the lookout for like-minded people to join our mission. For more information, visit www.every.com The Role: This unique entry-level Research Associate I position offers the rare chance to work across two core teams, Analytics and Protein Science. You’ll gain hands-on experience supporting protein development, purification, and analysis, while learning how these disciplines work together to drive innovation in precision fermentation. This is a great opportunity for someone early in their career who thrives in the lab, loves variety, and wants to learn fast in a collaborative, mission-driven environment. What you'll accomplish Generate data through basic biochemistry/molecular biology techniques (including but not limited to BCA, SDS-PAGE) Operate and maintain analytical equipment (e.g., HPLC-UV/RI and Combustion Analyzer, dynamic light scatterer, FPLC-UV, fluorescence/UV plate reader) Support protein characterization workflows through lab-scale protein powder generation involving bench-scale downstream processing unit operations (microfiltration, ultrafiltration, diafiltration) Prepare samples, reagents, and buffers to support cross-functional experiments Collaborate with scientists and engineers across teams to troubleshoot and iterate quickly as part of our Design, Build, Test, and Learn pipeline Present results
$131.1K – $163.9K/yr
Why join us Brex is the intelligent finance platform that enables companies to spend smarter and move faster in more than 200 markets. By combining global corporate cards and banking with intuitive spend management, bill pay, and travel software, Brex enables founders and finance teams to accelerate operations, gain real-time visibility, and control spend effortlessly. Brex’s AI-native automation and world-class service eliminate manual expense and accounting tasks for customers so they can focus on what matters most. Tens of thousands of the world's best companies run on Brex, including DoorDash, Coinbase, Robinhood, Zoom, Plaid, Reddit, and SeatGeek. Working at Brex allows you to push your limits, challenge the status quo, and collaborate with some of the brightest minds in the industry. We’re committed to building a diverse team and inclusive culture and believe your potential should only be limited by how big you can dream. We make this a reality by empowering you with the tools, resources, and support you need to grow your career. Marketing at Brex The Marketing team tells the Brex story — turning product innovation into customer obsession. We develop messaging, shape the brand, and bring our platform to life through thoughtful, compelling content. Our team spans Revenue Marketing, Product Marketing, and Brand Marketing, and works cross-functionally with Product, Sales, Business Development, and Design to drive growth and build trust. What you’ll do Brex is seeking a deeply analytical, growth-rooted Senior Growth Marketing Manager to own, scale, and optimize our Paid Social engine, with Meta as the core growth driver and a clear mandate to test and scale emerging channels. In this role, you will be accountable for down-funnel pipeline, efficient CAC, and customer acquisition. You won't just optimize campaigns; you will engineer signal architecture, build scalable audience models, and establish a high-velocity creative testing engine. Critically, you know how to
About the Team OpenAI is building the next generation of advertising for an AI-powered world. As AI changes how people discover information, explore possibilities, and make decisions, we believe this presents a huge opportunity for small and medium-sized businesses to connect with customers in new ways. The SMB Ads Marketing team is building this opportunity from the ground up. We define how businesses discover OpenAI’s advertising platform, become successful advertisers, and grow their investment over time. Working closely with Product, Ads Engineering, Data Science, and Sales, we combine customer insight, rigorous experimentation, and fast execution to turn early opportunities into scalable growth. About the Role We are seeking a Growth Marketing Manager to build the web and organic growth engine for SMB advertisers at OpenAI. This is a zero-to-one opportunity to shape how businesses discover our advertising platform, understand its value, and become successful advertisers. You will own the SMB Ads web strategy and experimentation roadmap, combining customer insight, analytical rigor, and product judgment to grow qualified discovery and improve conversion. Your scope spans web experiences, organic search, generative engine optimization (GEO), and the journey from first visit through signup and activation. You will work closely with Product, Design, Ads Engineering, Content, and Data Science to turn opportunities into shipped experiences, operating with the speed and ownership of a startup. In This Role, You Will Own the SMB Ads web strategy and roadmap to grow qualified discovery, signup, and advertiser activation. Grow organic discovery through search, content, and GEO. Test new ways to help businesses find and evaluate OpenAI’s advertising offering. Analyze journeys, funnels, and cohorts to identify conversion barriers and prioritize high-impact improvements. Run experiments across landing pages, messaging, navigation, personalization, offers, and conversion flo
About the Team OpenAI develops models that can reason through complex problems and hardware designed for the demands of advanced AI. AI for Chips connects these efforts: applying increasingly capable AI systems to the work of semiconductor engineering. Our goal is to help engineers develop better chips and shorten design cycles. This work brings research, model training, and hardware expertise together to build tools that engineers can use on real designs, with correctness and measurable performance at the center. About the Role We’re hiring a Research Engineer to help OpenAI models solve chip-design problems through reinforcement learning, tool use, and evaluation. You’ll own experiments from the initial idea through implementation and analysis. That means building environments and evaluations, running training, investigating failures, and using the results to decide what to try next. You’ll also build the software needed to make those experiments reliable and reproducible. We value strong coding fundamentals, careful experimental judgment, and the ability to make progress independently. Prior chip-design experience is helpful, but you can learn the domain alongside the team’s hardware specialists. In this role, you will: Build RL environments and evaluations for tasks such as RTL generation, design verification, and physical design optimization. Develop and test approaches that help models use chip-design tools and improve power, performance, and area while preserving correctness. Design experiments, establish baselines, and measure whether improvements hold up on new tasks and designs. Investigate failures across model behavior, rewards, evaluation tools, and experiment infrastructure. Improve iteration speed through better tooling, faster evaluations, and proxy rewards that reflect the outcomes we care about. Turn successful experiments into reusable research code and training workflows, working closely with researchers and engineers. You might thrive in this ro
About the Team The Emerging Products team is a lean, high-output product lab group that builds products at the forefront of model capabilities. We collaborate across all teams within the company, from research and infrastructure to consumer products. The team is responsible for identifying new product opportunities, building them quickly, dogfooding them internally, and then launching the successful products to users. We use data, user research, and analytics to inform our ideas, and make decisions on what experiments are worth iterating, stopping, or scaling. About the Role We’re looking for a senior, product-minded software engineer to own ambiguous 0-to-1 work from idea through prototype, validation, and handoff. This is a full-stack role with a strong frontend and product emphasis: you will build the interfaces and supporting backend systems needed to test new experiences quickly, while making sound architectural choices that enable successful concepts to scale. This role is based in our Mission Bay office in San Francisco. In this role, you will: Build and ship high-quality, product experiments across the full stack. Turn ambiguous user needs and emerging technical capabilities into testable product concepts, using research and metrics to guide iteration. Own technical direction for 0-to-1 projects, balancing speed, reliability, and a clear path from prototype to scalable product. Partner closely with design, product, research, and engineering teams to dogfood, evaluate, launch, and transition successful experiments. You might thrive in this role if you: Have a track record of building and shipping end-to-end products in fast-moving, startup, founder-led, growth, or other high-ownership environments. Bring strong frontend engineering skills and enough backend and systems depth to make sound full-stack architectural decisions. Pair product intuition with evidence, using user research and product data to identify opportunities and make pragmatic tradeoffs. Operat
About the Team OpenAI’s Hardware organization develops silicon and system-level solutions designed for the unique demands of advanced AI workloads. The team builds next-generation AI-native silicon and systems while working closely with software, research, and manufacturing partners to co-design hardware tightly integrated with AI models. In addition to delivering systems for OpenAI’s supercomputing infrastructure, the team develops the tools, methodologies, and strategic partnerships needed to accelerate hardware innovation. About the Role We’re seeking an experienced Hardware Strategic Sourcing Manager to own sourcing strategy and supplier partnerships for fiber and optical interconnect components across OpenAI’s next-generation AI infrastructure. Reporting to the Head of Partnerships & Strategic Sourcing, you will lead sourcing across fiber cable assemblies, internal optical harnesses, fiber shuffles, optical backplane assemblies, connectorized and standalone passive optical assemblies, fiber-array units (FAUs), fiber-to-chip and coupling interfaces, detachable connectors, optical routing, and assigned optical packaging, assembly, and test services. You will work closely with electrical engineering, optical engineering, systems engineering, mechanical and packaging engineering, quality, rack integration, data-center deployment,manufacturing, supply chain, finance, legal, and program management teams to translate demanding bandwidth, signal integrity, reliability, and scale requirements into resilient supplier partnerships and scalable commercial strategies. Your work will directly support the performance, reliability, manufacturability, and scale of the high-speed optical connectivity required for OpenAI’s next-generation AI systems. In this role, you will: Develop and execute a comprehensive sourcing strategy for fiber and optical interconnect components supporting high-bandwidth AI systems and infrastructure. Own sourcing across optical fiber cable assembli
About the Team OpenAI’s Governance team helps shape how the company responsibly develops and deploys increasingly capable AI. We bring together technical evidence, policy, and operational perspectives to help the company address emerging risks, resolve difficult questions, and make well-supported decisions. Our work includes shaping and improving governance practices, supporting effective oversight, developing clear assessments and recommendations, and ensuring that decisions lead to action. We work closely with research, safety, security, legal, and product teams to identify gaps, reconcile different views, and improve our approach as capabilities and circumstances change. About the Role We are looking for a curious, high-agency Research Program Manager who can reason from first principles, make sense of incomplete or conflicting information, and move difficult work forward. You will help shape the substance of governance reviews, connect ideas and evidence across teams, and develop recommendations that are both well-founded and practical. The work requires someone who can use established approaches where they fit, recognize when circumstances call for a different approach, and update their thinking as new evidence emerges. You should bring sound judgment, a willingness to experiment and learn, and the program-management discipline to turn good analysis into action. Depending on your experience and the team’s needs, your work may focus on safety advisory and board-level oversight, deployment governance, or standards and strategic partner commitments. In this role, you will: Bring together technical, policy, and operational inputs to develop coherent assessments, recommendations, and decision materials for governance bodies and senior leaders. Work through emerging or ambiguous questions, test assumptions, identify gaps or conflicting evidence, and help determine what additional analysis or decisions are needed. Engage critically with research, evaluations, safeguar
Other cities to consider
More places hiring for this role
Get new test analyst 3 jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime