About the team Preparedness is a critical Safety Research team at OpenAI, which is focused on mitigating AI threats to global security that could scale to an extreme level of severity. Our work involves: Measurement. Monitoring and predicting the evolving capabilities of frontier AI systems. Mitigation. Keeping misuse safeguards, alignment tools, and security measures on track to adequately address extreme threats that might arise in the future. Coordination. Setting mitigation targets by maintaining OpenAI’s preparedness framework , and partnering with other staff to achieve these targets. This is urgent, fast-paced work that has far-reaching implications for the company and for society. About the role The stakes of securing OpenAI increases as our internal coding and research becomes increasingly driven by autonomous AI agents. Compromising these agents could allow a cyber threat actor to compromise many other parts of the company. In this role, you would lead Preparedness work defending the security of our internal AI agents against insiders, Advanced Persistent Threats (APTs), or powerful AI agents. We’re looking for a strong hands-on technical executor with experience working directly with advanced cyber threat actors. In this role, you will: Develop and maintain threat models via which advanced attackers could compromise our coding assistants and automated security systems. Identify security investments that are especially critical to make in advance; for example, prioritizing by implementation lead-times, costs, and benefit. Partner with Security, Infrastructure, Research, Legal, and Preparedness to align on implementation plans and tradeoffs. Lead technical execution directly when needed, including prototyping controls, writing and reviewing software, and coordinating engineers across teams. Work with penetration testers to close gaps in defenses. You might thrive in this role if you: Are an exceptional hands-on technical executor. Have worked with advanced
Jobs in United States
Tester in San Francisco
12 active opportunities · Updated September 2026
Showing
12 jobs
Explore current tester jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is seeking talented and experienced Software Engineers to join our Platform team within the Infrastructure organization. As a senior member of Baseten's Platform Team, you will own the systems that let every engineer at Baseten prove their code works before it reaches production. Our product runs mission-critical AI inference for customers who measure downtime in dollars per second, which means our internal bar for correctness, performance, and failure tolerance has to be exceptional. Your focus is the full testing stack: fast and reliable unit test tooling, integration harnesses that spin up realistic environments on demand, load and performance testing for GPU-backed inference workloads, and resilience testing that deliberately breaks things so our customers never have to find out what happens when a node dies mid-request. This is a builder role with org-wide leverage. You won't be writing tests for other teams — you'll be building the frameworks, harnesses, and feedback loops that make writing good tests the path of least resistance, and you'll set the standards for what "well-tested" means at Baseten. RESPONSIBILITIES Own Baseten's testing strategy end to end — define the standards, the tiers, and the tooling that engineering teams build against. Build and maintain unit, integration, load and performance testing frameworks Design end to end test infrastructure that provisions realistic dependencies
About the Role The Mechanical Commissioning Project Engineer owns mechanical commissioning planning, readiness, quality coordination, and test execution oversight for the project. This role ensures mechanical systems are installed, inspected, started, balanced, controlled, and tested in a way that supports reliable integrated facility performance. Reports to the Commissioning Project Lead and partners closely with mechanical contractors, equipment vendors, design/engineering teams, the electrical commissioning lead, controls stakeholders, and vendor field/test engineers. Key Responsibilities Develop and maintain the mechanical commissioning scope, readiness criteria, inspection strategy, and discipline test execution plan. Review mechanical design packages, specifications, submittals, method statements, controls narratives, sequence assumptions, and testing requirements for commissionability and risk. Coordinate mechanical QA/QC inspections with contractors and vendor field/test engineers, including installation checks, pre-functional readiness, deficiency capture, and closeout tracking. Own mechanical commissioning procedure development and review, including equipment startup, functional testing, controls verification, balancing prerequisites, failure mode validation, and integrated systems testing inputs. Coordinate with equipment vendors on factory/site acceptance requirements, startup support, test prerequisites, documentation packages, and vendor participation during critical tests. Support readiness and execution for cooling, ventilation, hydronic, pumping, heat rejection, controls, and other project-specific mechanical systems. Lead discipline-level review of mechanical test results, deficiencies, corrective actions, retest requirements, trend logs, and acceptance evidence. Maintain mechanical commissioning dashboards and status inputs for the Commissioning Project Lead, including risk items, resource needs, test readiness, and issue aging. Partner with the E
About the Team OpenAI’s Industrial Compute organization is building the infrastructure required to support the next generation of frontier AI systems. Through a combination of strategic partnerships and self-built data center campuses, we are scaling the physical infrastructure needed to deliver compute at unprecedented scale. The Commissioning organization is responsible for ensuring this infrastructure is safely tested, validated, integrated, and transitioned into reliable operations. As the portfolio grows, the team is building common standards, processes, tools, and reporting systems that allow commissioning programs to operate consistently across projects while giving teams and leadership clear visibility into readiness, risk, and execution. About the Role We are seeking a Commissioning Program Manager to build and scale the operating systems behind OpenAI’s infrastructure commissioning programs. You will own the development and continuous improvement of commissioning standards, processes, tools, dashboards, and KPIs across the infrastructure portfolio. You will work closely with commissioning and construction teams to translate field execution needs into practical playbooks, workflows, templates, metrics, and reporting mechanisms that teams can use from construction readiness through testing and turnover. This role sits at the intersection of infrastructure delivery, program management, process design, and data. The ideal candidate understands how complex construction projects operate and can turn fragmented workflows and project data into repeatable systems that improve execution without creating unnecessary administrative burden. Key Responsibilities Develop and maintain commissioning program standards, playbooks, process maps, templates, checklists, stage gates, and acceptance criteria across infrastructure projects. Establish consistent workflows for commissioning planning, construction readiness, QA/QC, issue management, document control, testing evidence
About the Team OpenAI’s Industrial Compute organization is building the infrastructure required to support the next generation of frontier AI systems. Through a combination of strategic partnerships and self-built data center campuses, we are scaling the power, cooling, electrical, mechanical, and controls infrastructure needed to deliver compute at unprecedented scale. The Commissioning organization is responsible for ensuring this infrastructure is safely tested, validated, integrated, and transitioned into reliable operations. For our self-build campuses, the team operates through a hybrid delivery model: OpenAI provides commissioning leadership, discipline ownership, governance, and project integration, while commissioning partners provide field and test engineering capacity to support inspections, startup, testing, and turnover. About the Role We are seeking a Commissioning Project Lead to own the commissioning strategy and execution for a large-scale, self-build data center project. You will lead the overall commissioning program from early construction planning through startup, functional testing, integrated systems testing, and final turnover. You will establish the commissioning execution plan, integrate commissioning activities into the master project schedule, coordinate multidisciplinary readiness, and lead the vendor commissioning partners providing field and test engineering capacity. This role serves as the primary commissioning interface to project leadership, construction management, contractors, equipment vendors, operations, and commissioning partners. You will be responsible for creating clarity across organizations, identifying readiness and schedule risks early, and ensuring the facility progresses through testing and turnover against clearly defined acceptance criteria. The role will initially support planning and coordination in a hybrid capacity and transition to full-time onsite presence as construction, inspections, startup, testing, and t
About the Team The Consumer Devices team at OpenAI builds end-to-end hardware and software systems that bring AI into the physical world. We work at the intersection of custom silicon, embedded systems, operating systems, cloud services, mechanical engineering, electrical engineering, and product design to deliver reliable, production-ready devices at scale. Within Consumer Devices, Hardware Engineering eXperience, or HEX, is a new bootstrapped team building the environments, applications, compute, product-data systems, and workflows that let hardware engineers do their work without needing to troubleshoot the machinery underneath. HEX owns virtual engineering environments, HPC/GPU compute, storage, networking, licensing, MCAD/ECAD/CAE applications, PLM, product data, automation, validation, and support as one connected system. About the Role As a Staff PLM & Engineering Applications Engineer, you will be one of the first technical builders of HEX and the primary counterpart to the HEX lead. You will own the engineering-application and product-data side of the hardware engineering experience, with an initial focus on NX, Teamcenter, licensing, parts import, integrations, packaging, validation, and user workflows. This is not a traditional Teamcenter administration role and not a Corporate IT application-support role. You will take complex, fragile workflows and turn them into reliable engineering systems. This role is highly hands-on and systems-oriented. You will not inherit a mature environment and support queue. You will help build a fresh one, replacing manual setup guides, tribal knowledge, repeated support issues, and team handoffs with tested automation and reliable workflows. In This Role, You Will Own the technical architecture, deployment, configuration, integration, validation, and long-term operation of NX and Teamcenter. Build reliable workflows for parts import, product-data migration, metadata quality, BOMs, revisions, lifecycle states, and releas
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Seattle, Washington D.C., Raleigh, London, and Amsterdam. The Integrations Operations Engineering (IOE) team strengthens Plaid's network with financial institutions by directly resolving integration issues to increase the reliability and quality of data, and by building new integrations to grow our financial network. We sit at the nexus of engineering, a deep understanding of Plaid's products and customers, and direct financial-institution relationships: we write and ship the code that keeps connectivity healthy, we use data to focus on the issues with the greatest customer impact, and we work directly with data partners (the financial institutions and platforms themselves) to resolve the problems that can't be fixed from our side alone. The team plays a mission-critical role in ensuring industry-leading connectivity so our customers can meet ever-expanding financial-services use cases and reach as many users as possible. What You'll Do Investigate and resolve the highest-impact integration issues by writing maintainable, tested code and deploying it to production, then monitor for regression or degradation after your changes ship. Prioritize by customer impact. The team runs a business-value-based prioritization model that automatically
$106K – $142K/yr
About Taskrabbit: Taskrabbit is a marketplace platform that conveniently connects people with Taskers to handle everyday home to-do’s, such as furniture assembly, handyman work, moving help, and much more. At Taskrabbit, we want to transform lives one task at a time. As a company we celebrate innovation, inclusion and hard work. Our culture is collaborative, pragmatic, and fast-paced. We’re looking for talented, entrepreneurially minded and data-driven people who also have a passion for helping people do what they love. Together with IKEA, we’re creating more opportunities for people to earn a consistent, meaningful income on their own terms by building lasting relationships with clients in communities around the world. Taskrabbit is a hybrid company with employees distributed across the US and EU and a Built In — Best Places to Work (2022, 2023, 2024, 2025) continually ranked across multiple national and regional categories. Join us at Taskrabbit, where your work will be meaningful, your ideas valued, and your potential unleashed! This role operates on a hybrid schedule requiring two days of in-office collaboration per week. The position must be based in the San Francisco Bay Area. About the Role We're hiring a Software Engineer II within our Fulfillment organization — the backend systems that get the right job to the right Tasker and see it through to completion. You'll join Fulfillment Lifecycle, the team that decides how jobs are matched to Taskers for our partner and marketplace business, increasingly using unstructured data and experimentation to make matching smarter and fulfillment more reliable. The team is part of a company-wide platform modernization effort, breaking a legacy monolith into well-bounded, API-first services. We're hiring for a strong backend engineer who thrives on complex, data-intensive problems, is comfortable with ambiguity, and takes pride in well-tested, observable, production-ready code. What You'll Work On B
$155K – $400K/yr
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About this role Are you ready to redefine the future of JavaScript development? Are you a seasoned JavaScript expert who gets a thrill from tackling complex challenges across the entire JavaScript ecosystem? Do you believe that AI can be a powerful partner in crafting elegant, high-impact code? If you're ready to leave the mundane behind and join a team that's shaping the tools used by millions of developers globally, we've got an opportunity for you at Sentry. This isn't your typical Senior Software Engineer position. As a key member of our growing JavaScript SDK team, you'll be at the forefront of innovation, working on everything from our cutting-edge SDKs for Node.js, Bun, Deno, Cloudflare Workers, and other modern server runtimes. You won't just be maintaining code; you'll be pushing the boundaries of what's possible in developer tooling across the rapidly evolving server-side JavaScript landscape. In this role you will Join our JavaScript SDK team and get ready to build the future. You'll be at the forefront, working on: A Universe of JavaScript Challenges: Dive deep into our extensive suite of JavaScript SDKs, with a sharp focus on server-side and edge runtimes — from the battle-tested Node.js ecosystem to cutting-edge alternatives like Bun and Deno, and distributed edge environments like Cloudflare Workers. Your work will directly empower millions of developers to build better, more reliable software, no matter which runtime powers their stack End-to-End Ownership: We believe in giving our engineers the autonomy to see their vision through. You'll have the freedom to plan, implement, and ship your code, from writing
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Plaid’s mission is to unlock financial freedom for everyone by making money movement and access to financial data simple and secure. As a Software Engineer, you will design and build the systems that power how millions of people connect to their finances. You will work across the stack, from reliable backend services and APIs to intuitive applications that bring those systems to life. You will collaborate with engineers, product managers, and designers to ship products that make financial services more accessible and transparent. At Plaid, engineers take ownership early, grow quickly, and see their work reach millions of users. Responsibilities: System Design & Development: Build and maintain scalable, reliable backend or fullstack systems and APIs that power Plaid’s products. Collaboration: Work closely with product managers, designers, and other engineers to define and deliver features that solve real customer problems. Code Quality: Write clean, efficient, and well-tested code. Participate in reviews to maintain high engineering standards. Testing & Debugging: Build automated tests, monitor system performance, and troubleshoot issues in production environments. Continuous Improvement: Con
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. This internship will take place from May 17 - August 6 or June 14 - Sept 3 (based on your summer schedule) and you will need to be able to work out of our NY or SF office during this time. What You'll Achieve: Write clean, secure, tested, and documented code. Design & enhance the Notion product with new capabilities with an AI first mindset Develop, fix and debug software for across our full stack: web services, databases, applications, tools while leveraging modern frameworks and tooling. Qualifications: Pursuing a bachelor's or master’s degree in computer science, engineering, or another related field. Must graduate before Summer 2028. This internship will take place from May 17 - August 6 or June 14 - Sept 3 (based on your summer schedule) and you will need to be able to work out of our NY or SF office during this time. Previous internship experience. Working towards a proficiency of one or more programming languages such as Typescript, Node.js, or Python. You find large challenges exciting and enjoy discovering problems as much as solving them. You are able to problem-solve and adapt to changing priorities in a fast-paced, dy
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. This internship will take place from January 25 - April 16 and you will need to be able to work out of our NY or SF office during this time. What You'll Achieve: Write clean, secure, tested, and documented code. Design & enhance the Notion product with new capabilities with an AI first mindset Develop, fix and debug software for across our full stack: web services, databases, applications, tools while leveraging modern frameworks and tooling. Qualifications: Pursuing a bachelor's or master’s degree in computer science, engineering, or another related field. Must graduate before December 2027. This internship will take place from January 25 - April 16 and you will need to be able to work out of our NY or SF office during this time. Previous internship experience. Working towards a proficiency of one or more programming languages such as Typescript, Node.js, or Python. You find large challenges exciting and enjoy discovering problems as much as solving them. You are able to problem-solve and adapt to changing priorities in a fast-paced, dynamic environment. Skills You’ll Need To Bring: Thoughtful problem-solving: For you, problem-s
Other cities to consider
More places hiring for this role
Get new tester jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime