AI Systems Engineer - Codex Core Agents About The Team The Codex Core Agents team builds the agent harness that turns model capability into real-world action. We own the systems around the model: prompting and interpreting model outputs, executing actions safely in real environments, and feeding production experience back into better models and better agent behavior. This team sits close to research and works across the stack: harness, model interaction, inference, sandboxed execution, orchestration, evals, production reliability, and the performance envelope around tokens, latency, cost, capacity, and quality. The harness is open source and increasingly part of how models are trained and evaluated, making this one of the highest-leverage layers in Codex. About The Role We’re looking for engineers to build the AI systems that make Codex agents dependable in production. The ideal candidate is an agent-systems builder: hands-on across low-level systems and ML workflows, able to debug Codex behavior end to end across the harness, model behavior, inference/runtime stack, GPU fleet, and product surface. You’ll work with research, infrastructure, and product to design agent harness capabilities, run experiments and ablations across the model + system prompt + harness stack, build frameworks for assessing production agent performance, and turn messy failures into durable improvements. What You’ll Do Design and build the core agent harness and execution loop that lets Codex agents interpret model outputs, use tools, execute code, and complete long-horizon tasks safely. Build sandboxing, isolation, orchestration, state, and workflow infrastructure for agents operating in real development environments. Develop evaluation, experimentation, and debugging systems that distinguish harness issues, model behavior, inference/runtime issues, and product failures. Run ablations across prompts, model-facing interfaces, context construction, tool-use strategies, and harness behavior to
Jobs in United States
Open Availability Am Pm in San Francisco
106 active opportunities · Updated October 2026
Showing
15 jobs
Explore current open availability am pm jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team The Intelligence and Investigations team seeks to rapidly detect and disrupt abuse in AI technologies to ensure their safe use. We are dedicated to identifying emerging abuse trends, analyzing risks, and working with our internal partners to implement effective mitigation strategies to protect against misuse. Our efforts contribute to OpenAI's overarching goal of developing AI that benefits all of humanity. About the Role As an Intelligence Systems Engineer, you’ll be focused on advancing our Intelligence & Investigations efforts at OpenAI, ensuring the safe and responsible use of AI across our products and services. We are seeking a self-starter to prototype, develop, and maintain new tools and processes that integrate OpenAI’s models and infrastructure to enable internal teams to make sense of large, open-domain datasets, fight abuse, and inform high-stakes decisions. You will be a crucial technical bridge between our data scientists and subject matter experts and technical teams like Platform Integrity, Safety Systems, and Research by leading the development of innovative tools and processes that bolster goals in scaled collections, investigations, and analysis. The ideal candidate has strong analytical and data skills, with a background in both prototyping and building scalable systems that can swiftly detect emerging threats, process vast amounts of information, and deliver insights to stakeholders. We value professionals with outstanding communication skills, a commitment to continuous learning, and who are dedicated to promoting the responsible use of AI. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Prototype, build, and maintain at-scale intelligence systems that detect, triage, and monitor targeted signals from both open-source and internal data Analyze requirements and deliver end-to-end solutions that address
About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. Role summary We are seeking a Networking Operating System Firmware Engineer to help bootstrap and scale the switching layer of our AI supercomputers. In this role, you will build and maintain custom NOS images from scratch, using open source components from SONiC, SAI, FRR, and related networking stacks while working across the Linux kernel, switch ASIC SAI/SDKs, platform drivers, control-plane services, and orchestration layers. This is a software engineering role that requires a deep understanding of networking, NOS internals, switch hardware, and production systems. You will design, implement, test, and debug production NOS software across platform drivers, routing and control-plane state, ASIC programming, observability, and fleet integration. The engineer in this role should be able to work through ambiguous, open-ended technical problems and drive feature development across software, hardware, and vendor boundaries. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will Design, develop, and maintain custom NOS images for large-scale AI fabrics, using open source components from SONiC, FRR, and related networking stacks. Integrate, build and configure Linux kernel components, device drivers, switch ASIC SDKs, and SAI layers. Bring up new switch platforms, including thermal and fan control, power monitoring, transceiver management, watchdogs, OSFP CMIS, L
About the Team Our Executive Operations team includes Executive Business Partners and Administrative Business Partners, who serve as trusted advisors and collaborators to OpenAI's executives and leaders, focused on strong communication and operational excellence across teams. With a focus on elevating the impact and efficiency of leadership, we anticipate needs, streamline processes, and provide comprehensive support to ensure our executives can focus on high-impact initiatives. We are pivotal in driving success and achieving key milestones by cultivating strong relationships and leveraging our deep understanding of business objectives. With a commitment to excellence and a proactive approach, we are dedicated to empowering our executives and contributing to the overall growth and success of the company. Our leadership team reflects OpenAI’s culture and core values and is a mission-driven, kind, and thoughtful group. We take pride in creating a work environment that fosters collaboration, open communication, and authenticity, making OpenAI an excellent place to work for highly accomplished professionals. About the Role: This posting is part of a shared hiring process for Executive Business Partner and Administrative Business Partner opportunities at OpenAI. Rather than hiring for a specific team, we consider candidates across multiple opportunities and identify the best fit based on your experience, interests, and business needs as you progress through the interview process. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Manage complex calendars, balancing competing priorities while ensuring leaders’ time is aligned with business needs. Coordinate internal and external meetings, resolve scheduling conflicts, and facilitate effective communication across stakeholders. Plan and manage domestic and international travel, ensuring seamless logis
About the Team The Synthetic RL team develops reinforcement learning methods that leverage synthetic data, environments, and feedback to train and evaluate frontier AI models. The team explores approaches such as self-play, simulators, and other synthetic evaluations to push model capability, generalization, and alignment beyond what is possible with the current prevailing methodology. About the Role As a Research Scientist on the Synthetic RL team, you will develop novel reinforcement learning techniques that use synthetic environments and feedback to improve large-scale models. You’ll work closely with other researchers to design experiments, analyze learning dynamics, and translate research insights into training approaches used in production systems. We’re looking for researchers who enjoy working on open-ended problems, value fast iteration, and want their work to directly shape how frontier models are trained. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Research and develop reinforcement learning algorithms Design and run experiments to study training dynamics and model behavior at scale Collaborate with engineers and researchers to integrate successful approaches into model training pipelines You might thrive in this role if you: Have a strong background in reinforcement learning, machine learning research, or related fields Have strong engineering and statistical analysis skills Enjoy exploring new problem spaces where data, objectives, and evaluation are imperfect or evolving Are motivated by seeing research ideas influence real-world AI systems About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an ex
About the Team The Agent Enablement team works across engineering, product, design, and research to bring our technology to the world. We seek to learn from deployment and broadly distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. We aim to make our innovative tools globally accessible, transcending geographic, economic, and platform barriers. Our commitment is to facilitate the use of AI to enhance lives, supported by rigorous insights into how people use our products. About the Role We are looking for experienced full-stack engineers to join our new Agent Enablement team. Our goal is to design and grow an open ecosystem of agent-enabled sites and services. This is a wide-ranging role: you’ll build new user and agent identity protocols, user experiences to control and observe agents across web, desktop, and mobile, and much more. We will rely on you to drive our technical decisions while also steering our product and partnership direction, optimizing for both short-term impact and long-term success of the ecosystem. We value engineers who are impact-driven, autonomous, and adept at removing barriers to forward progress. In this role, you will: Design the primitives and protocols for an open agent ecosystem, enabling our users’ agents to make the best use of sites and services across the internet. Build a next-generation user experience to observe and control agents, across web, desktop, and mobile. Evolve our approach to token consumption across subscriptions and API customers. Execute on fast-paced projects in collaboration with research, design, data science and other product engineering teams. Work closely with our strategic customers and partners to grow the ecosystem. You might thrive in this role if you: Have strong full-stack engineering skills and experience shipping customer-facing products from concept to production. You’re comfortable working across frontend, backend, APIs, data models, and product desig
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . The Team Pinterest's Data Engineering organization builds and operates the data platforms that power every Pinterest product — the batch and streaming pipelines that produce training data for our ML models, the storage and table formats that back our data lake, the workflow orchestration that runs it all, and the analytics platforms that fuel experimentation and decision-making. We're in the middle of a multi-year modernization effort: moving to streaming-first ingestion (CDC, Kafka, Flink), open table formats (Iceberg), a consolidated workflow platform, and retiring legacy footprints along the way. We work closely with ML, product, and analytics teams to make Pinterest's data platforms faster, more reliable, and more cost-efficient. What You'll Do: Lead a multi-quarter portfolio of data platform modernization programs — spanning ingestion (CDC/
Drata is building the trust layer between great companies - automating compliance, managing risk, and helping organizations prove trust continuously as they scale. We're Dratanauts: a global crew of 600+ professionals united by a culture that rewards integrity, ownership, and raising the bar, no matter where in the world we're working from. Why Join the Drata Team? At Drata, you're not maintaining legacy compliance software - you're building the agentic AI platform defining what trust looks like for the next generation of companies. Here's what makes the work itself worth showing up for: Problems without a playbook: You'll work at the edge of AI and security, building agentic governance, continuous compliance, and real-time trust verification to solve problems that don't have an established answer yet. You're writing it as you go. Real ownership, not just process: Our values center on owning outcomes and raising the bar, not checking boxes. You're expected to have opinions and back them. A seat at the table: Your perspective is unique and valued. Open debate and diverse viewpoints are built into how decisions actually get made here, at every level. Growth at rocketship speed: Drata is scaling fast, which means scope grows fast too. High performers get more ownership, visibility, and experience. A crew, not just coworkers: Dratanauts consistently describe a "come as you are" culture with sharp, curious people—the kind of team that makes hard problems genuinely fun to solve. See what they say here and follow us on LinkedIn for company news, employee stories, and career updates. Job Summary: As Revenue Operations Manager, you’ll help build the operational foundation for a fast-scaling GTM team. What you'll do: Own and operate some of our core GTM systems — including Salesforce, sales engagement, marketing automation, and BI tools. Translate GTM requirements into scalable workflows, automations, and processes across Sales, Marketing, Partnerships, and Account Management
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We're looking for a senior social media manager to own how Baseten shows up on social. Our audience is ML engineers, infrastructure teams, and technical founders, and most of them meet us first on X or LinkedIn, around a model launch, a benchmark, an open source release, or a customer result. This role decides what that first impression is. This is a senior individual contributor role. You set the strategy and you write the posts. Day to day you'll work with product marketing, comms, design, our engineers, and our founders. This role relies on technical credibility. You don't need an engineering background, but you do need to understand what we're claiming and why it matters. We post about latency, throughput, and GPU cost, and we hold ourselves to getting those details right. RESPONSIBILITIES This role is the face of the Baseten brand on our social channels and builds our direct line of communication with the community across X, LinkedIn, YouTube, and the communities where our audience already spends time. Track the conversation across AI and open source, and move quickly when we have something useful to add. Set the social strategy: what we post where, how each account grows, and how we measure it, with a clear point of view on which channels deserve investment and which don't. Translate technical work into posts worth sharing: model launches, benchmark results, open source projects, engineering deep dives, an
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About this role This is a chance to help shape the next chapter of WRITER's product at a moment when generative AI is redefining how work gets done. As senior staff product manager, you'll report directly to Wil Pong, VP of Product, and operate as a force-multiplier across the product organization — setting direction where it's unclear, aligning stakeholders, and making sure our most critical initiatives ship with excellence. You'll own ambiguous, high-impact problem spaces that span multiple product areas and teams, and you'll be equal parts strategist and operator. One moment you're shaping a multi-quarter roadmap; the next you're rolling up your sleeves to unblock execution. We're looking for someone who thrives at the intersection of customer empathy, technical depth, and business judgment — and who is energized by turning complex, evolving AI capabilities into products customers love. We are open to hiring for this role at multiple levels, with compensation ranges scaled
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role We're looking for an exceptional software engineer to join our rapidly evolving team at WRITER. In this pivotal role, you'll be at the forefront of expanding human capacity by building the next generation of AI-powered solutions that transform how leading enterprises operate. You'll dive deep into developing a state-of-the-art platform that leverages cutting-edge generative AI technologies, from large language models to sophisticated agentic workflows, delivering seamless, scalable, and secure applications that redefine enterprise productivity. This is an unparalleled opportunity to make a tangible impact, shaping the future of AI and contributing to a product that’s changing how the world works. This role is hybrid, based out of our San Francisco, New York City, or Seattle hubs. You'll report to our senior director, engineering. We are open to hiring for this role at multiple levels, with compensation ranges scaled to match your experience, expertise, and the
$165K – $270K/yr
Drata is building the trust layer between great companies - automating compliance, managing risk, and helping organizations prove trust continuously as they scale. We're Dratanauts: a global crew of 600+ professionals united by a culture that rewards integrity, ownership, and raising the bar, no matter where in the world we're working from. Why Join the Drata Team? At Drata, you're not maintaining legacy compliance software - you're building the agentic AI platform defining what trust looks like for the next generation of companies. Here's what makes the work itself worth showing up for: Problems without a playbook: You'll work at the edge of AI and security, building agentic governance, continuous compliance, and real-time trust verification to solve problems that don't have an established answer yet. You're writing it as you go. Real ownership, not just process: Our values center on owning outcomes and raising the bar, not checking boxes. You're expected to have opinions and back them. A seat at the table: Your perspective is unique and valued. Open debate and diverse viewpoints are built into how decisions actually get made here, at every level. Growth at rocketship speed: Drata is scaling fast, which means scope grows fast too. High performers get more ownership, visibility, and experience. A crew, not just coworkers: Dratanauts consistently describe a "come as you are" culture with sharp, curious people—the kind of team that makes hard problems genuinely fun to solve. See what they say here and follow us on LinkedIn for company news, employee stories, and career updates. Job Summary: Drata is looking for a Staff Data Engineer! This person will be a key member of the growing data team, supporting one of the fastest growing B2B SaaS startups to achieve unicorn status. At Drata, we’re on a mission to help build trust across the internet! Data accuracy is an essential cornerstone of our mission. We are looking for a senior data engineer that can help us strategize
About the Team The Product Policy team develops and implements policies that shape how OpenAI’s technology is built and used. We work with teams across the company to turn complex questions about AI’s benefits and risks into practical guidance for responsible research, product development, and deployment. About the Role As Product Policy Manager, Regulation, you will map external regulatory and legislative trends and synthesize them into concrete research, product, and policy requirements. You will connect deep expertise in regulation and public policy with a sophisticated understanding of AI and strong product judgment. Your remit will span frontier AI, youth, wellbeing, use in high-stakes domains, deceptive use, and other emerging issues. Working closely with Legal and Global Affairs, you will help Product, Engineering, Research, and Safety teams understand what external developments mean for their work, what decisions or evidence are needed, and how to act. We’re looking for someone who combines deep regulatory and public-policy expertise with an in-depth understanding of AI and a demonstrated ability to advise product and engineering teams. You should be able to move from complex external developments to clear decisions and requirements that teams can implement. This role can be based in San Francisco or remote. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Map and prioritize regulatory and legislative developments across jurisdictions and AI issue areas. Distinguish enacted requirements, proposals, guidance, and emerging expectations; assess their relevance, uncertainty, timing, and implications for our models and products. Translate external developments into actionable research questions, evaluation needs, product requirements, and policy recommendations. Identify gaps in evidence, define the intended outcome and acceptance criteria, and make tradeoffs and open questions c
From $160K/yr
About Flexport: At Flexport, we believe global trade can move the human race forward. That’s why it’s our mission to make global commerce so easy there will be more of it. We’re shaping the future of a $10T industry with solutions powered by innovative technology and exceptional people. Today, companies of all sizes—from emerging brands to Fortune 500s—use Flexport technology to move more than $19B of merchandise across 112 countries a year. The recent global supply chain crisis has put Flexport center stage as we continue to play a pivotal role in how goods move around the world. We are proud to have the support of the best investors in the game who believe in our mission, solutions and people. Ready to tackle global challenges that impact business, society, and the environment? Come join us. The opportunity Join the Office of the CEO and work directly on Flexport's highest-priority problems alongside the CEO, the President, and the leaders who run our regions and modes of freight. You will pick up a mode or a region, ocean, air, trucking, omnichannel, or customs, and learn it from the inside by solving real problems in it. Pricing in a mode that is losing money. Why a region misses its plan. Whether we buy, build, or partner. Then you go run part of it. Some people in this role move into a business unit as the number two to the leader who runs it. Some stay at the center and take on more scope. The operating job at the end is the point of the role, not a side benefit. You will: Own our biggest open questions end to end and come back with recommendations the CEO can act on. Build the analysis yourself. Pull the data, build the model, find the error before someone else does. Go to where the work happens, from warehouses and ports to a customer's supply chain team. Opinions about freight formed in a conference room are usually wrong. Run the operating cadence for a freight mode or region. Weekly metrics, quarterly plan, and the follow-ups nobody wants to
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. ABOUT THE TEAM Supply is responsible for knowing everything happening in the compute market: who's building, who's buying, and on what terms. This role owns a specific and fast-moving slice of that map — emerging clouds and international markets — and owns the full relationship lifecycle in that space, from first outreach through to closed terms. RESPONSIBILITIES Build and maintain a real-time picture of the emerging cloud and international compute landscape — who's active, what they're building, and what terms are available Own the full partnership lifecycle in this space — from identifying and sourcing new providers, to negotiating terms, to ongoing relationship management Develop and manage relationships across a broad set of emerging and international providers, from account reps up through leadership Identify, structure, and help close opportunities where Baseten can move quickly to secure favorable capacity terms Define compelling value propositions tailored to different types of providers, rather than a one-size-fits-all pitch Partner closely with others in the team already covering this space to build out a durable, well-organized intelligence and relationship function Collaborate with the broader Supply and Deals functions to bring opportunities to the table and support negotiation when it's time to close WHAT WE’RE LOOKING FOR Equal parts relationship-builder and operator — you can open a door and also drive it
Other cities to consider
More places hiring for this role
Get new open availability am pm jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime